SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[MM][Bugfix] Replace `PatchEmbed`'s conv3d to linear layer (#27418)

Signed-off-by: Isotr0py <mozf@mail2.sysu.edu.cn>
Co-authored-by: Roger Wang <hey@rogerw.io>
I
Isotr0py committed
42efe609ba75eb1b0bc06ae635778b2bc0aa4e7a
Parent: 88d3141
Committed by GitHub <noreply@github.com> on 10/24/2025, 7:32:47 AM