Logo
Explore Help
Sign In
Karylab-cklius/vllm
Watch 1
Star 0
Fork 1
Code Issues Pull Requests 1 Actions Packages Projects Releases Wiki Activity
Files
08e50675615fc0b3e4e458645d8dd16cefb371bd
vllm/vllm/model_executor/model_loader/reload
T
History
alexxu-robloxGitHubYQ-WangalexhxuCursor
d96aee0951 [Bugfix] Re-sync parameter tp_rank after process_weights_after_loading (fix replicated / disable_tp weight reload) (#48025)
Signed-off-by: Alex Xu <alexxu@roblox.com>
Co-authored-by: YQ-Wang <yiqingwang@roblox.com>
Co-authored-by: alexhxu <alex.xu1015@gmail.com>
Co-authored-by: Cursor <cursoragent@cursor.com>
2026-07-18 08:40:06 +00:00
..
__init__.py
[Quantization] - Layerwise reloading of Attention/KV quantized models (#38995)
2026-04-15 18:03:32 -07:00
layerwise.py
[Bugfix] Re-sync parameter tp_rank after process_weights_after_loading (fix replicated / disable_tp weight reload) (#48025)
2026-07-18 08:40:06 +00:00
meta.py
[Bugfix] Preserve unloaded non-persistent buffers during layerwise reload (#44371)
2026-07-14 17:46:29 -07:00
sanitize.py
Speed up docs build (#44635)
2026-06-05 14:51:44 +00:00
torchao_decorator.py
[QeRL] Layerwise Reloading (#32133)
2026-01-30 08:50:05 -07:00
types.py
[Bugfix] Preserve unloaded non-persistent buffers during layerwise reload (#44371)
2026-07-14 17:46:29 -07:00
utils.py
Speed up docs build (#44635)
2026-06-05 14:51:44 +00:00
Powered by Gitea Version: 1.27.1 Page: 1087ms Template: 2ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API