Logo
Explore Help
Sign In
Karylab-cklius/vllm
Watch 1
Star 0
Fork 1
Code Issues Pull Requests 1 Actions Packages Projects Releases Wiki Activity
Files
143e4dccdfd8293c70c76f8d32a60ce23ecc23ea
vllm/vllm/model_executor
T
History
arloandGitHub 8c29042bb9 [Feature] Add InstantTensor weight loader (#36139)
2026-03-14 18:05:23 +01:00
..
kernels
[Misc] Use envs module to get VLLM_DISABLED_KERNELS (#35776)
2026-03-11 13:37:46 +00:00
layers
Enable loading of fused expert weights in the Transformers modelling backend (#36997)
2026-03-14 07:01:06 +00:00
model_loader
[Feature] Add InstantTensor weight loader (#36139)
2026-03-14 18:05:23 +01:00
models
[Misc] Clean up Kimi-audio whisper encoder loading (#36903)
2026-03-14 23:37:52 +08:00
offloader
[UX] Remove NoOpOffloader log (#35678)
2026-03-04 12:13:40 -08:00
warmup
[Feature]: Remove Chunking From FusedMoE (#34086)
2026-03-12 14:24:38 -04:00
__init__.py
[Platform] Deprecate seed_everything (#31659)
2026-01-04 18:34:04 -08:00
custom_op.py
[MM][OOT] Support CPU seq_lens for OOT MMEncoderAttention kernels (#36605)
2026-03-12 03:28:23 -07:00
parameter.py
[QeRL] Layerwise Reloading (#32133)
2026-01-30 08:50:05 -07:00
utils.py
[BugFix] Fix EPLB fail for MoeFP4 model with Marlin backend (#33262)
2026-01-29 16:52:11 +08:00
Powered by Gitea Version: 1.27.1 Page: 253ms Template: 5ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API