SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[Attention] Add missing kv cache scale setup (#27490)

Signed-off-by: Matthew Bonanni <mbonanni@redhat.com>
M
Matthew Bonanni committed
a99564ac5b2bc86d97cb18a7e18086b4ba94466a
Parent: 4c5f632
Committed by GitHub <noreply@github.com> on 10/25/2025, 7:12:49 AM