This website requires JavaScript.
Explore
Help
Sign In
karylab_agents
/
vllm
Watch
1
Star
0
Fork
0
forked from
Karylab-cklius/vllm
Code
Pull Requests
2
Actions
1
Packages
Activity
Files
fbc3a1907aeb6beff59461e535045f17ac14306e
vllm
/
tests
/
quantization
T
History
Yuwen Zhou
and
GitHub
0cd9b7af25
[CPU] Support CPU W4A16 INT4 MoE (
#43409
)
...
Signed-off-by: yuwenzho <
yuwen.zhou@intel.com
>
2026-06-12 07:12:37 +00:00
..
__init__.py
…
fp_quant.py
…
reference_mxfp4.py
…
test_auto_gptq.py
…
test_auto_round.py
…
test_blackwell_moe.py
…
test_compressed_tensors.py
…
test_configs.py
…
test_cpu_offload.py
…
test_cpu_wna16.py
[CPU] Support CPU W4A16 INT4 MoE (
#43409
)
2026-06-12 07:12:37 +00:00
test_cutlass_w4a16.py
…
test_experts_int8.py
…
test_fp8_per_channel.py
…
test_fp8.py
…
test_gfx950_moe.py
…
test_gptq_dynamic.py
…
test_gptq_v2.py
…
test_lm_head.py
…
test_mixed_precision.py
…
test_modelopt.py
…
test_moe_wna16.py
…
test_online.py
…
test_per_token_kv_cache.py
…
test_quantization_config_args.py
…
test_quark.py
…
test_register_quantization_config.py
…
test_torchao.py
…
test_trtllm_nvfp4_hidden_dim_padding.py
…
test_turboquant.py
…
utils.py
…