This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
59c62cd2b4214dacb00a8efdfbee7af69058c1cf
vllm
/
tests
/
v1
/
e2e
T
History
Giancarlo Delfin
and
GitHub
272abd5f48
[Tests][Spec Decode] Add gemma4 MTP acceptance rates test (
#47920
)
...
Signed-off-by: Giancarlo Delfin <
gdelfin@inferact.ai
>
2026-07-27 16:15:54 -07:00
..
general
[ROCm][CI] Ensure sliding window tests release GPU memory (
#49055
)
2026-07-18 20:44:05 +00:00
spec_decode
[Tests][Spec Decode] Add gemma4 MTP acceptance rates test (
#47920
)
2026-07-27 16:15:54 -07:00
__init__.py
[V1] Implement Cascade Attention (
#11635
)
2025-01-01 21:56:46 +09:00
test_cpu_linear_attn_chunked_prefix.py
[CPU] Enable chunked prefill and prefix caching for qwen3.5 (
#46202
)
2026-06-25 03:49:21 +00:00
test_hybrid_chunked_prefill.py
[ROCm][CI] Enable hybrid chunked prefill test (
#38317
)
2026-03-30 10:30:26 +08:00
test_replayssm_decode.py
[Kernel] ReplaySSM: cache SSM inputs for faster Mamba2 standard decode (
#48018
)
2026-07-24 09:39:49 -07:00