Logo
Explore Help
Sign In
karylab_agents/vllm
Watch 1
Star 0
Fork 0
forked from Karylab-cklius/vllm
Code Pull Requests 2 Actions 1 Packages Activity
Files
a0df04e4775efbfebd65c997259d63af0ec548ce
vllm/tests/v1/sample
T
History
Chaojun ZhangGitHubKunshang Ji
556b063e45 [XPU] Fix test_spec_decode_logprobs: use FLASH_ATTN for XPU in GPU_DETERMINISM_KWARGS (#44468)
Signed-off-by: Chaojun Zhang <chaojun.zhang@intel.com>
Co-authored-by: Kunshang Ji <kunshang.ji@intel.com>
2026-06-17 11:07:04 +08:00
..
__init__.py
…
test_batched_count_greater_than.py
[Performance Improvement] Update batched_count_greater_than to handle batch size 1 without recompile (#38933)
2026-04-09 23:51:31 +08:00
test_logprobs_e2e.py
…
test_logprobs.py
[XPU] Fix test_spec_decode_logprobs: use FLASH_ATTN for XPU in GPU_DETERMINISM_KWARGS (#44468)
2026-06-17 11:07:04 +08:00
test_rejection_sampler.py
[Security] Fix remote DoS via invalid recovered token reinjection (#44744)
2026-06-10 02:31:43 -07:00
test_sampler.py
refactor hard coded device string in test files under tests/v1 and tests/lora (#37566)
2026-04-03 11:21:47 +08:00
test_sampling_params_e2e.py
[V0 Deprecation] Remove code related to per-request logits processors (#34400)
2026-02-12 20:44:28 +08:00
test_topk_topp_sampler.py
[ROCm][CI] Defer AITER sampler import and isolate server test PYTHONPATH (#44823)
2026-06-10 08:56:11 +00:00
utils.py
[Reasoning] Support for speculative decoding with thinking budget (#34668)
2026-04-29 06:14:52 +00:00
Powered by Gitea Version: 1.27.1 Page: 4422ms Template: 3ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API