SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[Frontend] Enforce tokenize=False when applying chat template (#27205)

Signed-off-by: Isotr0py <mozf@mail2.sysu.edu.cn>
Co-authored-by: Isotr0py <mozf@mail2.sysu.edu.cn>
R
Russell Bryant committed
3ada34f9cb4d1af763fdfa3b481862a93eb6bd2b
Parent: 0eb8f2b
Committed by GitHub <noreply@github.com> on 10/21/2025, 2:57:34 AM