This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
4b87b3e845fcf2ce54f74fa3178a714d755c3a6a
vllm
/
tests
/
model_executor
/
model_loader
T
History
JartX
and
GitHub
7b476c8f14
[ROCm][CI] Skip fp8 reload tests on gfx90a (MI250) (
#44369
)
...
Signed-off-by: JartX <
sagformas@epdcenter.es
>
2026-06-02 22:27:14 -05:00
..
fastsafetensors_loader
…
instanttensor_loader
[Feature] Add InstantTensor weight loader (
#36139
)
2026-03-14 18:05:23 +01:00
runai_streamer_loader
Fix RunAI streamer tensor buffer reuse during weight loading (
#43464
)
2026-05-27 19:16:52 -07:00
tensorizer_loader
[Core] Move
max_concurrent_batches
to
VllmConfig
(
#44274
)
2026-06-02 08:57:25 -07:00
__init__.py
…
test_ep_weight_filter.py
[Performance][Model Loader] Skip non-local expert weights during EP model loading (
#37136
)
2026-03-16 01:33:36 -07:00
test_modelexpress_loader.py
[Core] Add native ModelExpress load format (
#43105
)
2026-05-21 16:05:01 -04:00
test_registry.py
…
test_reload.py
[ROCm][CI] Skip fp8 reload tests on gfx90a (MI250) (
#44369
)
2026-06-02 22:27:14 -05:00
test_sharded_state_loader.py
[ROCm][CI] Force max_num_seqs=1 on ROCm In test_sharded_state_loader to reduce flakiness (
#33277
)
2026-01-31 12:28:29 +08:00