Logo
Explore Help
Sign In
Karylab-cklius/vllm
Watch 1
Star 0
Fork 1
Code Issues Pull Requests 1 Actions Packages Projects Releases Wiki Activity
Files
2fa1f8ec00cf85b15422cb4c0e8eb3632ee13ea8
vllm/tests/models/language/generation
T
History
Artem PerevedentsevandGitHub b92ef9ec5a [Perf] Enable FlashInfer top-k/top-p sampler by default (#40376)
Signed-off-by: Artem Perevedentsev <aperevedents@nvidia.com>
2026-04-29 19:10:34 +04:00
..
__init__.py
[CI/Build] Reorganize models tests (#17459)
2025-04-30 23:03:08 -07:00
conftest.py
[ROCm][CI] Disable skinny GEMMs in language model standard tests to fix non-determinism (#35152)
2026-03-02 15:04:18 +08:00
test_common.py
[CPU] Refactor CPU affinity and memory management (#39781)
2026-04-17 21:01:08 +08:00
test_gemma.py
[Model] Revert PR #26715: Restore custom PaliGemma and Gemma3-MM impl… (#27309)
2025-10-22 10:05:34 -07:00
test_granite.py
[CI/Build] Improve stability of CPU tests (#39966)
2026-04-16 21:50:36 +08:00
test_grok.py
[Model] Add Grok-2 (#31847)
2026-01-08 04:59:48 -08:00
test_hybrid.py
[Perf] Enable FlashInfer top-k/top-p sampler by default (#40376)
2026-04-29 19:10:34 +04:00
test_mistral.py
[Bugfix][CI] fix typos (#34934)
2026-03-05 17:05:46 +00:00
test_phimoe.py
[CI] Skip Phi-MoE test due to old API util (#31632)
2026-01-05 08:52:07 +08:00
Powered by Gitea Version: 1.27.1 Page: 277ms Template: 2ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API