This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
59c62cd2b4214dacb00a8efdfbee7af69058c1cf
vllm
/
tests
/
v1
/
e2e
/
spec_decode
T
History
Giancarlo Delfin
and
GitHub
272abd5f48
[Tests][Spec Decode] Add gemma4 MTP acceptance rates test (
#47920
)
...
Signed-off-by: Giancarlo Delfin <
gdelfin@inferact.ai
>
2026-07-27 16:15:54 -07:00
..
__init__.py
[CI] Split V1 e2e + engine (1 GPU) into separate jobs (
#36945
)
2026-03-13 14:16:02 -07:00
test_async_spec_decode.py
[CI] Add MTP coverage: Qwen3.5 correctness + no-sync spec decode (
#40472
)
2026-04-30 12:24:09 -07:00
test_laguna_dflash.py
Add Laguna XS.2.1 DFlash drafter support (
#46853
)
2026-07-02 18:09:27 -07:00
test_lora_with_spec_decode.py
[CI] Fix test_lora_with_spec_decode on V2 model runner (
#43314
)
2026-05-22 14:24:18 +08:00
test_mtp_parallel_load.py
[Test] Add DeepSeek MTP parallel-load tests (
#41653
)
2026-07-21 10:38:20 -04:00
test_spec_decode.py
[Tests][Spec Decode] Add gemma4 MTP acceptance rates test (
#47920
)
2026-07-27 16:15:54 -07:00