SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[Bugfix] [AITER] [ROCm] Fix Quark MoE Quant Config and AITER Fused MoE quant type logic (#27029)

Signed-off-by: vllmellm <vllm.ellm@embeddedllm.com>
V
vllmellm committed
e33ee23ee3cde5aa69912d9d03aa31421851662c
Parent: b10c64c
Committed by GitHub <noreply@github.com> on 10/17/2025, 6:51:10 PM