Logo
Explore Help
Sign In
Karylab-cklius/vllm
Watch 1
Star 0
Fork 1
Code Issues Pull Requests 1 Actions Packages Projects Releases Wiki Activity
Files
df29793dc73a83f3c86c19de967adffda1a28a93
vllm/vllm/engine
T
History
Ronen SchafferGitHubRobert ShawRobert Shaw
bf480c5302 Add more Prometheus metrics (#2764)
Co-authored-by: Robert Shaw <114415538+robertgshaw2-neuralmagic@users.noreply.github.com>
Co-authored-by: Robert Shaw <rshaw@neuralmagic.com>
2024-04-28 15:59:33 -07:00
..
output_processor
[Core] Refactoring sampler and support prompt logprob for chunked prefill (#4309)
2024-04-26 13:02:02 +00:00
__init__.py
Change the name to vLLM (#150)
2023-06-17 03:07:40 -07:00
arg_utils.py
[Kernel] Full Tensor Parallelism for LoRA Layers (#3524)
2024-04-27 00:03:48 -07:00
async_llm_engine.py
[Bugfix][Core] Fix get decoding config from ray (#4335)
2024-04-27 11:30:08 +00:00
llm_engine.py
Add more Prometheus metrics (#2764)
2024-04-28 15:59:33 -07:00
metrics.py
Add more Prometheus metrics (#2764)
2024-04-28 15:59:33 -07:00
Powered by Gitea Version: 1.27.1 Page: 121ms Template: 2ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API