This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
ae098abe3fffedd7acecc04265a10dc5624efdca
vllm
/
vllm
/
model_executor
/
warmup
T
History
Harry Mellor
and
GitHub
ae098abe3f
[CI] Fix some errors on
main
(
#47726
)
...
Signed-off-by: Harry Mellor <
19981378+hmellor@users.noreply.github.com
>
2026-07-06 19:40:23 +00:00
..
__init__.py
[Docs] Fix warnings in docs build (
#22588
)
2025-08-10 05:49:51 -07:00
cutedsl_warmup.py
[Feat][1/N] CuTeDSL warmup infrastructure, FA4 MLA (
#46182
)
2026-06-30 12:17:34 -07:00
deep_gemm_warmup.py
Enable DeepSeek V4 and GLM-5.1 on SM120 (
#43477
)
2026-06-22 11:54:14 -07:00
deepseek_v4_mhc_warmup.py
Enable DeepSeek V4 and GLM-5.1 on SM120 (
#43477
)
2026-06-22 11:54:14 -07:00
fa4_cutedsl_config.py
[Feat][1/N] CuTeDSL warmup infrastructure, FA4 MLA (
#46182
)
2026-06-30 12:17:34 -07:00
flashinfer_autotune_cache.py
Enable DeepSeek V4 and GLM-5.1 on SM120 (
#43477
)
2026-06-22 11:54:14 -07:00
flashinfer_sparse_mla_warmup.py
[BugFix] Gate MRV2 mixed sparse-MLA warmup on
max_num_seqs
> 1 (
#47050
)
2026-06-30 16:31:27 +01:00
kernel_warmup.py
[CI] Fix some errors on
main
(
#47726
)
2026-07-06 19:40:23 +00:00
minimax_m3_msa_warmup.py
[Model] Add MiniMax M3 support (
#45381
)
2026-06-16 01:01:25 +08:00
qwen_triton_warmup.py
[CI Bugfix] Lazily import Qwen warmup dependencies (
#47539
)
2026-07-03 23:10:49 +08:00
sparse_mla_triton_warmup.py
[Model Runner V2][Perf] Warm up GLM-5.2 DSA indexer prefill metadata kernel (
#47285
)
2026-07-02 16:31:31 +00:00
v1_block_table_warmup.py
[Perf][1/N] Expand Triton kernel warmup coverage, DSv4 (
#46634
)
2026-06-29 16:40:34 +00:00