Files
vllm/tests/kernels
Roger WangandOpenAI Codex 21eaa58f71 [Bugfix] Fix NVFP4 CuteDSL batched all2all selection
Promote --moe-backend=flashinfer_cutedsl to the batched NVFP4 backend when the selected all2all path requires batched expert activations, and make flashinfer_nvlink_one_sided prepare/finalize speak the batched expert contract.

Also add regression coverage for the backend selection and one-sided regroup/reduce helpers.

Co-authored-by: OpenAI Codex <codex@openai.com>
Signed-off-by: Roger Wang <hey@rogerw.io>
2026-04-18 17:47:35 -07:00
..
2026-04-16 13:06:01 -07:00