This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
2
Packages
Activity
Files
ce88f01c9ac4fcde9dd43a983074d4e893cde65d
vllm
/
vllm
/
platforms
T
History
Li, Jiang
and
GitHub
b4601ad43f
[CPU] Add fused GDN support for AMX CPU platform (
#42707
)
...
Signed-off-by: jiang1.li <
jiang1.li@intel.com
>
2026-05-18 03:04:36 -07:00
..
__init__.py
[Platform] Fix RISC-V platform detection (lscpu parsing + non-NUMA meminfo) (
#40427
)
2026-04-24 04:33:05 +00:00
cpu.py
[CPU] Add fused GDN support for AMX CPU platform (
#42707
)
2026-05-18 03:04:36 -07:00
cuda.py
[MLA Attention Backend] Add TOKENSPEED_MLA backend for DSR1/Kimi K25 prefill + decode on Blackwell (
#41778
)
2026-05-13 23:48:02 -07:00
interface.py
platforms: add uses_cpu_device() hook to Platform for DeviceConfig (
#42313
)
2026-05-12 12:39:17 -07:00
rocm.py
[Quant] Consolidate GPTQ: rename gptq_marlin.py to auto_gptq.py (
#38288
)
2026-05-15 08:25:52 +08:00
tpu.py
[Refactor][TPU] Remove torch_xla path and use tpu-inference (
#30808
)
2026-01-07 16:07:16 +08:00
xpu.py
[XPU] disable fusion pattern support on XPU platform (
#39789
)
2026-04-23 10:07:45 +08:00
zen_cpu.py
[ZenCPU] AMD Zen CPU Backend with supported dtypes via zentorch weekly (
#39967
)
2026-04-18 06:22:37 +00:00