Logo
Explore Help
Sign In
karylab_agents/vllm
Watch 1
Star 0
Fork 0
forked from Karylab-cklius/vllm
Code Pull Requests 2 Actions 1 Packages Activity
Files
231c2c63e4decd0cbf863690dfffe88e1d97a003
vllm/tests/v1/core
T
History
Chen ZhangandGitHub 9607d5eb44 [Hybrid Allocator] Support full attention with different hidden size (#25101)
Signed-off-by: Chen Zhang <zhangch99@outlook.com>
2025-09-19 23:43:59 -07:00
..
__init__.py
…
test_async_scheduler.py
[Spec Decode] Make propose_draft_token_ids non-blocking for lower TTFT (#23041)
2025-08-18 17:20:38 -07:00
test_encoder_cache_manager.py
[Multimodal] Remove legacy multimodal fields in favor of MultiModalFeatureSpec (#24548)
2025-09-12 21:42:23 +08:00
test_kv_cache_utils.py
[Hybrid Allocator] Support full attention with different hidden size (#25101)
2025-09-19 23:43:59 -07:00
test_prefix_caching.py
[Tests] fix initialization of kv hash in tests (#24273)
2025-09-15 21:48:27 +00:00
test_scheduler_e2e.py
…
test_scheduler.py
[Chore] Cleanup guided namespace, move to structured outputs config (#22772)
2025-09-18 09:20:27 +00:00
test_single_type_kv_cache_manager.py
[Core] Use sha256 bytes instead of BlockHash to reduce GC overhead (#23673)
2025-09-08 21:34:37 -07:00
utils.py
[Core] Use sha256 bytes instead of BlockHash to reduce GC overhead (#23673)
2025-09-08 21:34:37 -07:00
Powered by Gitea Version: 1.27.1 Page: 2881ms Template: 3ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API