This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
2
Packages
Activity
Files
deepseek_optimizations_alex_rob
vllm
/
tests
/
v1
/
spec_decode
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
Delete Directory
History
Eldar Kurtić
and
GitHub
e439c784fa
Add support for Eagle with separate lm-head and embed_tokens layers (
#28549
)
...
Signed-off-by: Eldar Kurtic <
8884008+eldarkurtic@users.noreply.github.com
>
2025-11-15 06:12:02 -08:00
..
test_eagle.py
Add support for Eagle with separate lm-head and embed_tokens layers (
#28549
)
2025-11-15 06:12:02 -08:00
test_max_len.py
[Bugfix] Spec decode + structured output + spec model max len edge case (
#28298
)
2025-11-08 19:44:25 +00:00
test_mtp.py
Add support for Eagle with separate lm-head and embed_tokens layers (
#28549
)
2025-11-15 06:12:02 -08:00
test_ngram.py
[Redo]
#26368
(
#28771
)
2025-11-14 22:47:41 -08:00
test_speculators_eagle3.py
[Speculators] Move tests + fix integration (
#27308
)
2025-10-29 00:54:21 -07:00
test_tree_attention.py
[Attention] Refactor CUDA attention backend selection logic (
#24794
)
2025-11-11 07:40:44 -05:00