SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[Model]Improve Qwen3VLMoeForConditionalGeneration packed_modules_mapping (#27096)

Signed-off-by: Jee Jee Li <pandaleefree@gmail.com>
J
Jee Jee Li committed
daec4d2624cb816b92d5463c7f47878a342c7e76
Parent: 6c9fdbf
Committed by GitHub <noreply@github.com> on 10/17/2025, 11:47:00 AM