This website requires JavaScript.
Explore
Help
Sign In
Karylab-cklius
/
vllm
Watch
1
Star
0
Fork
1
Code
Issues
Pull Requests
1
Actions
Packages
Projects
Releases
Wiki
Activity
Files
c188749bcdaa2c72cc3c8a4a28e722af2abc4bb8
vllm
/
tests
/
models
/
quantization
T
History
Micah Williamson
and
GitHub
e7213003cb
[ROCm][CI] Fix TP size issue for
test_gpt_oss
(
#35887
)
...
Signed-off-by: Micah Williamson <
micah.williamson@amd.com
>
2026-03-03 20:57:34 +00:00
..
__init__.py
…
test_awq.py
[Renderer] Define
render_cmpl
and
render_chat
(
#34039
)
2026-02-07 05:24:40 -08:00
test_bitsandbytes.py
…
test_fp8.py
[1/N][Attention] Restructure attention: move files (
#31916
)
2026-01-09 13:10:24 -08:00
test_gguf.py
[Bugfix][Quantization] Support BF16 tensors on GGUF (
#29948
)
2025-12-03 10:33:46 +00:00
test_gpt_oss.py
[ROCm][CI] Fix TP size issue for
test_gpt_oss
(
#35887
)
2026-03-03 20:57:34 +00:00
test_gptq_marlin.py
…
test_modelopt.py
…
test_mxfp4.py
…
test_nvfp4.py
[Kernel][Performance] Enable smaller Scaling Factor tiling for NVFP4 small-batch decoding (
#30885
)
2026-01-13 15:22:53 -08:00