Logo
Explore Help
Sign In
karylab_agents/vllm
Watch 1
Star 0
Fork 0
forked from Karylab-cklius/vllm
Code Pull Requests 2 Actions 1 Packages Activity
Files
f1599ca55d79cb686cb94dc3ff2f65d82db94940
vllm/tests/v1/e2e
T
History
Charlie FuandGitHub 6af70e11a0 [ROCm][CI] Fix test_max_len.py for Rocm (#29916)
Signed-off-by: charlifu <charlifu@amd.com>
Signed-off-by: Charlie Fu <Charlie.Fu@amd.com>
2025-12-08 16:58:30 -05:00
..
__init__.py
[V1] Implement Cascade Attention (#11635)
2025-01-01 21:56:46 +09:00
test_async_scheduling.py
[Core] Support logprobs with spec decode + async scheduling (#29223)
2025-11-25 12:55:24 -08:00
test_cascade_attention.py
[V0 Deprecation] Remove VLLM_USE_V1 from tests (#26341)
2025-10-07 15:42:31 +00:00
test_context_length.py
[Bugfix] Fix validate model input for decoder models (#27099)
2025-11-13 10:18:47 -08:00
test_correctness_sliding_window.py
[CI][ROCm] Fix test_correctness_sliding_window (#29243)
2025-12-02 04:53:27 +00:00
test_kv_sharing_fast_prefill.py
[CI][ROCm][tests/v1/e2e] Fix multiprocessing launch for the test (#29123)
2025-12-02 20:46:10 +00:00
test_lora_with_spec_decode.py
[Misc] remove useless v1 env (#29164)
2025-11-21 01:41:20 -08:00
test_min_tokens.py
Update Optional[x] -> x | None and Union[x, y] to x | y (#26633)
2025-10-12 09:51:31 -07:00
test_pooling_chunked_prefill.py
Add tests for chunked prefill and prefix cache with causal pooling models (#26526)
2025-10-14 07:45:04 +08:00
test_spec_decode.py
[ROCm][CI] Fix test_max_len.py for Rocm (#29916)
2025-12-08 16:58:30 -05:00
Powered by Gitea Version: 1.27.1 Page: 234ms Template: 1ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API