SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[BugFix] Fix failing gemma-3-1b-it test: `test_lm_eval_accuracy_v1_engine[google/gemma-3-1b-it]` (#27111)

Signed-off-by: Lucas Wilkinson <lwilkins@redhat.com>
L
Lucas Wilkinson committed
9f020f4f31094d129671037dd65f6b10aae5e814
Parent: 3b45075
Committed by GitHub <noreply@github.com> on 10/18/2025, 6:44:39 PM