This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
4b0da7b60e35bbeb2cbd656c54966ee2ceb4c4bd
vllm
/
docs
/
source
/
features
/
quantization
T
History
3 people
Reid
GitHub
reidliu41
7377dd0307
[doc] update the issue link (
#17782
)
...
Signed-off-by: reidliu41 <
reid201711@gmail.com
> Co-authored-by: reidliu41 <
reid201711@gmail.com
>
2025-05-07 20:29:05 +08:00
..
auto_awq.md
[doc] update wrong hf model links (
#17184
)
2025-04-25 16:40:54 +00:00
bitblas.md
[doc] update wrong hf model links (
#17184
)
2025-04-25 16:40:54 +00:00
bnb.md
[doc] update wrong hf model links (
#17184
)
2025-04-25 16:40:54 +00:00
fp8.md
[doc] update the issue link (
#17782
)
2025-05-07 20:29:05 +08:00
gguf.md
doc: fix some typos in doc (
#16154
)
2025-04-07 05:32:06 +00:00
gptqmodel.md
[doc] update wrong model id (
#17287
)
2025-04-28 04:20:51 -07:00
index.md
Add NVIDIA TensorRT Model Optimizer in vLLM documentation (
#17561
)
2025-05-02 11:36:46 -07:00
int4.md
[doc] update the issue link (
#17782
)
2025-05-07 20:29:05 +08:00
int8.md
[doc] update the issue link (
#17782
)
2025-05-07 20:29:05 +08:00
modelopt.md
Add NVIDIA TensorRT Model Optimizer in vLLM documentation (
#17561
)
2025-05-02 11:36:46 -07:00
quantized_kvcache.md
[doc] add install tips (
#17373
)
2025-04-30 17:02:41 +00:00
quark.md
[doc] add install tips (
#17373
)
2025-04-30 17:02:41 +00:00
supported_hardware.md
Add NVIDIA TensorRT Model Optimizer in vLLM documentation (
#17561
)
2025-05-02 11:36:46 -07:00
torchao.md
[doc] update wrong hf model links (
#17184
)
2025-04-25 16:40:54 +00:00