This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
68ee8300a047db78fb52bac477daaaac7be11216
vllm
/
csrc
/
libtorch_stable
/
attention
T
History
Yifan Qiao
and
GitHub
aa4990a9a2
[Attention] Re-enable cross-layer KV cache layout for MLA via stride-aware kernels (
#45111
)
...
Signed-off-by: Yifan Qiao <
yifanqiao@inferact.ai
>
2026-06-22 06:57:02 -07:00
..
mla
[Attention] Re-enable cross-layer KV cache layout for MLA via stride-aware kernels (
#45111
)
2026-06-22 06:57:02 -07:00
attention_kernels.cuh
[9/n] Migrate attention and cache kernels to torch stable ABI (continued) (
#43717
)
2026-05-29 04:44:45 +00:00
attention_utils.cuh
[9/n] Migrate attention and cache kernels to torch stable ABI (continued) (
#43717
)
2026-05-29 04:44:45 +00:00
merge_attn_states.cu
[9/n] Migrate attention and cache kernels to torch stable ABI (continued) (
#43717
)
2026-05-29 04:44:45 +00:00
paged_attention_v1.cu
[9/n] Migrate attention and cache kernels to torch stable ABI (continued) (
#43717
)
2026-05-29 04:44:45 +00:00
paged_attention_v2.cu
[9/n] Migrate attention and cache kernels to torch stable ABI (continued) (
#43717
)
2026-05-29 04:44:45 +00:00