This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
76e4dcf225e4de115bdc20b00a78d49bec767c09
vllm
/
vllm
/
platforms
T
History
Matthew Bonanni
and
GitHub
684f254585
Prefer FlashAttention MLA as default over FlashMLA (
#27363
)
...
Signed-off-by: Matthew Bonanni <
mbonanni@redhat.com
>
2025-11-11 17:13:51 +00:00
..
__init__.py
[Misc] Clean up more utils (
#27567
)
2025-10-27 15:30:38 +00:00
cpu.py
[Attention] Refactor CUDA attention backend selection logic (
#24794
)
2025-11-11 07:40:44 -05:00
cuda.py
Prefer FlashAttention MLA as default over FlashMLA (
#27363
)
2025-11-11 17:13:51 +00:00
interface.py
[Attention] Refactor CUDA attention backend selection logic (
#24794
)
2025-11-11 07:40:44 -05:00
rocm.py
[Attention] Refactor CUDA attention backend selection logic (
#24794
)
2025-11-11 07:40:44 -05:00
tpu.py
[Attention] Refactor CUDA attention backend selection logic (
#24794
)
2025-11-11 07:40:44 -05:00
xpu.py
[Attention] Refactor CUDA attention backend selection logic (
#24794
)
2025-11-11 07:40:44 -05:00