This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
dependabot/github_actions/actions/setup-python-6.3.0
vllm
/
tests
/
v1
/
e2e
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
Delete Directory
History
zhrrr
and
GitHub
61ab70ec3b
[Model Runner V2] support mamba hybrid models align prefix cache (
#42406
)
...
Signed-off-by: zhuhaoran <
zhuhaoran.zhr@alibaba-inc.com
>
2026-06-29 14:09:16 -07:00
..
general
[Model Runner V2] support mamba hybrid models align prefix cache (
#42406
)
2026-06-29 14:09:16 -07:00
spec_decode
[Hardware][AMD][CI] Fix Spec Decode Eagle test group (
#46018
)
2026-06-21 17:40:02 -05:00
__init__.py
[V1] Implement Cascade Attention (
#11635
)
2025-01-01 21:56:46 +09:00
test_cpu_linear_attn_chunked_prefix.py
[CPU] Enable chunked prefill and prefix caching for qwen3.5 (
#46202
)
2026-06-25 03:49:21 +00:00
test_hybrid_chunked_prefill.py
[ROCm][CI] Enable hybrid chunked prefill test (
#38317
)
2026-03-30 10:30:26 +08:00