This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
887eaedb3de3baa6da3bf70a511c76b038592de2
vllm
/
csrc
/
libtorch_stable
/
quantization
T
History
Wentao Ye
and
GitHub
a55cb117a7
Merge branch 'main' into wentao-optimize-per-token-group-quant
2026-05-29 11:12:12 -04:00
..
awq
[5/n] Migrate CUTLASS MLA, hadamard, awq, allspark and DSV3 fused a gemm to torch stable ABI (continued) (
#42339
)
2026-05-13 07:24:39 +00:00
cutlass_w4a8
[4/n] Migrate FP4/W4A8 CUTLASS kernels to torch stable ABI (
#37503
)
2026-03-31 10:21:13 -07:00
fp4
[9/n] Migrate attention and cache kernels to torch stable ABI (continued) (
#43717
)
2026-05-29 04:44:45 +00:00
fused_kernels
[7/n] Migrate pos_encoding and norm kernels to libtorch stable ABI (continued) (
#43209
)
2026-05-23 13:20:00 +08:00
gguf
[6/n] Migrate activation kernels, gptq, gguf, non cutlass w8a8 to libtorch stable ABI (continued) (
#42663
)
2026-05-20 00:18:12 -07:00
gptq
[6/n] Migrate activation kernels, gptq, gguf, non cutlass w8a8 to libtorch stable ABI (continued) (
#42663
)
2026-05-20 00:18:12 -07:00
gptq_allspark
[5/n] Migrate CUTLASS MLA, hadamard, awq, allspark and DSV3 fused a gemm to torch stable ABI (continued) (
#42339
)
2026-05-13 07:24:39 +00:00
hadamard
/hadacore
[5/n] Migrate CUTLASS MLA, hadamard, awq, allspark and DSV3 fused a gemm to torch stable ABI (continued) (
#42339
)
2026-05-13 07:24:39 +00:00
w8a8
refactor
2026-05-28 22:48:37 +00:00
vectorization_utils.cuh
[2/n] Migrate per_token_group_quant to torch stable ABI (
#36058
)
2026-03-25 10:15:13 -07:00
vectorization.cuh
[2/n] Migrate per_token_group_quant to torch stable ABI (
#36058
)
2026-03-25 10:15:13 -07:00