Logo
Explore Help
Sign In
karylab_agents/vllm
Watch 1
Star 0
Fork 0
forked from Karylab-cklius/vllm
Code Pull Requests 2 Actions 1 Packages Activity
Files
2b93162fb0bc91c8e1e397cb2a385e8febce11e2
vllm/vllm/engine
T
History
Varun Sundar RabindranathGitHubVarun Sundar Rabindranath
79455cf421 [Misc] Enable V1 LoRA by default (#15320)
Signed-off-by: Varun Sundar Rabindranath <varun@neuralmagic.com>
Co-authored-by: Varun Sundar Rabindranath <varun@neuralmagic.com>
2025-04-01 16:53:56 +08:00
..
multiprocessing
[FEAT]Support reset prefix cache by specified device (#15003)
2025-03-19 10:54:41 -07:00
output_processor
[Bugfix] EAGLE output norm bug (#14464)
2025-03-15 06:50:33 +00:00
__init__.py
Change the name to vLLM (#150)
2023-06-17 03:07:40 -07:00
arg_utils.py
[Misc] Enable V1 LoRA by default (#15320)
2025-04-01 16:53:56 +08:00
async_llm_engine.py
[Misc] Update guided decoding logs to debug (#15310)
2025-03-24 04:25:20 -07:00
async_timeout.py
[Misc] Add SPDX-License-Identifier headers to python source files (#12628)
2025-02-02 11:58:18 -08:00
llm_engine.py
[V1] [Feature] Collective RPC (#15444)
2025-03-29 03:39:14 -07:00
metrics_types.py
[V1][Metrics] Support vllm:cache_config_info (#13299)
2025-02-22 00:20:00 -08:00
metrics.py
[V0][Metrics] Deprecate some questionable request time metrics (#14135)
2025-03-04 15:11:33 +00:00
protocol.py
[V1] Avoid redundant input processing in n>1 case (#14985)
2025-03-20 22:24:10 -07:00
Powered by Gitea Version: 1.27.1 Page: 213ms Template: 2ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API