This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
ffce72c0415701eb10af86668cf504a5eeba2a72
vllm
/
vllm
/
model_executor
/
models
/
transformers
T
History
Harry Mellor
and
GitHub
af9f583344
Revert "[Bugfix][CI] Gemma3 Transformers multimodal encoder profiling and build prompt-embedding fixtures" (
#45029
)
...
Signed-off-by: Harry Mellor <
19981378+hmellor@users.noreply.github.com
>
2026-06-10 01:37:03 -07:00
..
__init__.py
Don't compile vision encoder for Transformers backend (
#30518
)
2026-04-02 12:42:29 +00:00
base.py
[ROCm][CI] Fix ROCm LoRA Transformers fallback with full CUDA graphs (
#41577
)
2026-05-23 04:31:32 +00:00
causal.py
Fix pipeline parallel with multimodal models with the Transformers modelling backend (
#37057
)
2026-03-16 10:20:37 +00:00
legacy.py
[Bugfix] Fix RoBERTa position_ids accumulation on CUDA graph padding (
#37884
)
2026-03-23 15:15:12 +00:00
moe.py
[MoE Refactor] FusedMoE/MoERunner inversion refactor (
#41184
)
2026-06-08 10:42:58 -04:00
multimodal.py
Revert "[Bugfix][CI] Gemma3 Transformers multimodal encoder profiling and build prompt-embedding fixtures" (
#45029
)
2026-06-10 01:37:03 -07:00
pooling.py
[Doc] Fix duplicate words in comments (
#36713
)
2026-03-10 21:28:31 -07:00
utils.py
[Bugfix] Convert Gemma4-MM ViT linear layers to vllm native impl (
#43798
)
2026-06-01 21:41:16 -07:00