This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
0f199f197b4e7a835ccc5b4d15363f8faa7824c8
vllm
/
csrc
/
quantization
T
History
3 people
shixianc
GitHub
Shixian Cui
5780121c95
[Perf] Add swap_ab to SM90 FP8 non-block CUTLASS moe grouped gemm (
#20911
)
...
Signed-off-by: Shixian Cui <
shixian@amazon.com
> Co-authored-by: Shixian Cui <
shixian@amazon.com
>
2025-07-18 04:34:43 +00:00
..
aqlm
…
awq
…
compressed_tensors
[Perf] Optimize Vectorization Utils for Int 8 Quantization Kernels (
#20331
)
2025-07-04 15:06:24 +08:00
cutlass_w8a8
[Perf] Add swap_ab to SM90 FP8 non-block CUTLASS moe grouped gemm (
#20911
)
2025-07-18 04:34:43 +00:00
fp4
[Kernel] Basic tuned configs for NVFP4 CUTLASS dense GEMM (
#20646
)
2025-07-11 10:05:33 -06:00
fp8
…
fused_kernels
…
gguf
…
gptq
…
gptq_allspark
…
gptq_marlin
…
machete
[feat]: CUTLASS block scaled group gemm for SM100 (
#19757
)
2025-07-04 12:58:04 -06:00
marlin
…
activation_kernels.cu
…
utils.cuh
…
vectorization_utils.cuh
[Perf] Optimize Vectorization Utils for Int 8 Quantization Kernels (
#20331
)
2025-07-04 15:06:24 +08:00
vectorization.cuh
…