This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
ae098abe3fffedd7acecc04265a10dc5624efdca
vllm
/
tests
/
v1
/
e2e
T
History
BadrBasowid
and
GitHub
740f379fae
[ROCm][AITER] Directly Implement AITER Custom All-reduce in CudaCommunicator (
#46065
)
...
Signed-off-by: BadrBasowid <
badr.basowid@gmail.com
>
2026-07-06 12:16:32 +00:00
..
general
[ROCm][AITER] Directly Implement AITER Custom All-reduce in CudaCommunicator (
#46065
)
2026-07-06 12:16:32 +00:00
spec_decode
Add Laguna XS.2.1 DFlash drafter support (
#46853
)
2026-07-02 18:09:27 -07:00
__init__.py
[V1] Implement Cascade Attention (
#11635
)
2025-01-01 21:56:46 +09:00
test_cpu_linear_attn_chunked_prefix.py
[CPU] Enable chunked prefill and prefix caching for qwen3.5 (
#46202
)
2026-06-25 03:49:21 +00:00
test_hybrid_chunked_prefill.py
[ROCm][CI] Enable hybrid chunked prefill test (
#38317
)
2026-03-30 10:30:26 +08:00