Logo
Explore Help
Sign In
Karylab-cklius/vllm
Watch 1
Star 0
Fork 1
Code Issues Pull Requests 1 Actions Packages Projects Releases Wiki Activity
Files
2a543d6efecc4e0fe391cbccb68d99ab42e37c33
vllm/vllm/engine
T
History
Woosuk KwonandGitHub a463c333dd Use CuPy for CUDA graphs (#2811)
2024-02-13 11:32:06 -08:00
..
__init__.py
Change the name to vLLM (#150)
2023-06-17 03:07:40 -07:00
arg_utils.py
Remove hardcoded device="cuda" to support more devices (#2503)
2024-02-01 15:46:39 -08:00
async_llm_engine.py
fix some bugs (#2689)
2024-01-31 10:09:23 -08:00
llm_engine.py
Use CuPy for CUDA graphs (#2811)
2024-02-13 11:32:06 -08:00
metrics.py
Refactor Prometheus and Add Request Level Metrics (#2316)
2024-01-31 14:58:07 -08:00
ray_utils.py
[Ray] Integration compiled DAG off by default (#2471)
2024-02-08 09:57:25 -08:00
Powered by Gitea Version: 1.27.1 Page: 65ms Template: 1ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API