This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
bcf3c8230d23941cd1d3098ea89e54ed96b8be9e
vllm
/
docs
/
source
/
features
T
History
Harry Mellor
and
GitHub
d6484ef3c3
Add full API docs and improve the UX of navigating them (
#17485
)
...
Signed-off-by: Harry Mellor <
19981378+hmellor@users.noreply.github.com
>
2025-05-03 19:42:43 -07:00
..
quantization
Add NVIDIA TensorRT Model Optimizer in vLLM documentation (
#17561
)
2025-05-02 11:36:46 -07:00
automatic_prefix_caching.md
[Doc] Convert docs to use colon fences (
#12471
)
2025-01-29 11:38:29 +08:00
compatibility_matrix.md
Add full API docs and improve the UX of navigating them (
#17485
)
2025-05-03 19:42:43 -07:00
disagg_prefill.md
[Doc] Add two links to disagg_prefill.md (
#17168
)
2025-04-25 10:23:57 +00:00
lora.md
[Misc] Enable vLLM to Dynamically Load LoRA from a Remote Server (
#10546
)
2025-04-15 22:31:38 +00:00
reasoning_outputs.md
[Feature][Frontend]: Deprecate --enable-reasoning (
#17452
)
2025-05-01 06:46:16 -07:00
spec_decode.md
[V1][Spec Decode] Remove deprecated spec decode config params (
#15466
)
2025-03-31 09:19:35 -07:00
structured_outputs.md
[Docs] Update structured output doc for V1 (
#17135
)
2025-04-26 15:12:18 +00:00
tool_calling.md
Add chat template for Llama 4 models (
#16428
)
2025-04-24 20:19:36 +00:00