This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
afd7b1dce94fed484351fafd5bf5ea6601ac621e
vllm
/
tests
/
v1
/
e2e
T
History
roikoren755
and
GitHub
737bfa3a43
[Bugfix][Hybrid][NemotronH] Fix mamba_cache_mode=all + speculative decoding crash (
#41233
)
...
Signed-off-by: Roi Koren <
roik@nvidia.com
>
2026-05-18 14:54:00 +03:00
..
general
[Bugfix][Hybrid][NemotronH] Fix mamba_cache_mode=all + speculative decoding crash (
#41233
)
2026-05-18 14:54:00 +03:00
spec_decode
[CI] Migrate remaining B200 jobs to b200-k8s with test fixes (
#42387
)
2026-05-12 02:00:37 -07:00
__init__.py
[V1] Implement Cascade Attention (
#11635
)
2025-01-01 21:56:46 +09:00
test_hybrid_chunked_prefill.py
[ROCm][CI] Enable hybrid chunked prefill test (
#38317
)
2026-03-30 10:30:26 +08:00