This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
16282a9c4ee754bedd67f72b01ac76ae8568dbdd
vllm
/
csrc
/
cutlass_extensions
T
History
Chris Leonard
and
GitHub
56aff0dd15
[10/n] Migrate cuda_view and silu_and_mul_per_block_quant kernels to torch stale ABI. (
#44334
)
2026-06-04 20:14:43 -07:00
..
epilogue
Migrate header files to torch stable abi (
#44013
)
2026-06-02 08:09:52 -07:00
cute_utils.cuh
[4/n] Migrate FP4/W4A8 CUTLASS kernels to torch stable ABI (
#37503
)
2026-03-31 10:21:13 -07:00
torch_utils.hpp
[6/n] Migrate activation kernels, gptq, gguf, non cutlass w8a8 to libtorch stable ABI (continued) (
#42663
)
2026-05-20 00:18:12 -07:00
vllm_collective_builder.cuh
[Perf] Use upstream CUTLASS for SM90 Block FP8 kernel (
#23280
)
2025-09-11 15:43:14 -07:00
vllm_custom_types.cuh
[Kernel] (1/N) Machete - Hopper Optimized Mixed Precision Linear Kernel (
#7174
)
2024-08-20 07:09:33 -06:00
vllm_cutlass_library_extension.py
Update
Optional[x]
->
x | None
and
Union[x, y]
to
x | y
(
#26633
)
2025-10-12 09:51:31 -07:00
vllm_numeric_conversion.cuh
[Kernel] Initial Machete W4A8 support + Refactors (
#9855
)
2024-11-18 12:59:29 -07:00
vllm_type_utils.cuh
[Kernel] Initial Machete W4A8 support + Refactors (
#9855
)
2024-11-18 12:59:29 -07:00