SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[MM][Core] Decouple ViT backend from LM backend (#27061)

Signed-off-by: Roger Wang <hey@rogerw.io>
R
Roger Wang committed
c3a2c6ac5f9a13c5c407c69199dd0fa68145acad
Parent: 72f431e
Committed by GitHub <noreply@github.com> on 10/21/2025, 7:30:10 AM