This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
36bbecd6436d0dd4c7a27fbb09a787e00534d647
vllm
/
vllm
/
model_executor
/
warmup
T
History
Roberto L. Castro
and
GitHub
eddfd4cf21
[Perf][2/N] Expand Triton kernel warmup coverage, Qwen (
#46750
)
...
Signed-off-by: LopezCastroRoberto <
rocastro@redhat.com
>
2026-06-29 10:10:07 +00:00
..
__init__.py
[Docs] Fix warnings in docs build (
#22588
)
2025-08-10 05:49:51 -07:00
deep_gemm_warmup.py
Enable DeepSeek V4 and GLM-5.1 on SM120 (
#43477
)
2026-06-22 11:54:14 -07:00
deepseek_v4_mhc_warmup.py
Enable DeepSeek V4 and GLM-5.1 on SM120 (
#43477
)
2026-06-22 11:54:14 -07:00
flashinfer_autotune_cache.py
Enable DeepSeek V4 and GLM-5.1 on SM120 (
#43477
)
2026-06-22 11:54:14 -07:00
flashinfer_sparse_mla_warmup.py
Enable DeepSeek V4 and GLM-5.1 on SM120 (
#43477
)
2026-06-22 11:54:14 -07:00
kernel_warmup.py
[Perf][2/N] Expand Triton kernel warmup coverage, Qwen (
#46750
)
2026-06-29 10:10:07 +00:00
minimax_m3_msa_warmup.py
[Model] Add MiniMax M3 support (
#45381
)
2026-06-16 01:01:25 +08:00
qwen_triton_warmup.py
[Perf][2/N] Expand Triton kernel warmup coverage, Qwen (
#46750
)
2026-06-29 10:10:07 +00:00