SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[Prefix Cache] Use LoRA name for consistent KV-cache block hashing (#27211)

Signed-off-by: Sage Ahrac <sagiahrak@gmail.com>
S
Sage committed
1651003c35d263fbc7d87c2e75b9cca5d590eb27
Parent: 1cb8c6c
Committed by GitHub <noreply@github.com> on 10/22/2025, 6:13:03 PM