This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
6650e6a930dbdf1cd4def9b58e952376400ccfcf
vllm
/
docs
/
source
T
History
3 people
kakao-kevin-us
GitHub
Kevin-Yang
6650e6a930
[Model] Add classification Task with Qwen2ForSequenceClassification (
#9704
)
...
Signed-off-by: Kevin-Yang <
ykcha9@gmail.com
> Co-authored-by: Kevin-Yang <
ykcha9@gmail.com
>
2024-10-26 17:53:35 +00:00
..
_static
[Docs] Add RunLLM chat widget (
#6857
)
2024-07-27 09:24:46 -07:00
_templates
/sections
[Doc] Guide for adding multi-modal plugins (
#6205
)
2024-07-10 14:55:34 +08:00
assets
[Doc] add visualization for multi-stage dockerfile (
#4456
)
2024-04-30 17:41:59 +00:00
automatic_prefix_caching
[Doc] Add an automatic prefix caching section in vllm documentation (
#5324
)
2024-06-11 10:24:59 -07:00
community
Add NVIDIA Meetup slides, announce AMD meetup, and add contact info (
#8319
)
2024-09-09 23:21:00 -07:00
dev
[Core] Rename input data types (
#8688
)
2024-10-16 10:49:37 +00:00
getting_started
[Doc] Improve quickstart documentation (
#9256
)
2024-10-25 14:32:10 -07:00
models
[Model] Add classification Task with Qwen2ForSequenceClassification (
#9704
)
2024-10-26 17:53:35 +00:00
performance_benchmark
[Doc] fix 404 link (
#7966
)
2024-08-28 13:54:23 -07:00
quantization
[Hardware][CPU] Support AWQ for CPU backend (
#7515
)
2024-10-09 10:28:08 -06:00
serving
[Bugfix]: Make chat content text allow type content (
#9358
)
2024-10-24 05:05:49 +00:00
conf.py
[model] Support for Llava-Next-Video model (
#7559
)
2024-09-10 22:21:36 -07:00
generate_examples.py
Add example scripts to documentation (
#4225
)
2024-04-22 16:36:54 +00:00
index.rst
[Hardware][Intel CPU][DOC] Update docs for CPU backend (
#6212
)
2024-10-22 10:38:04 -07:00