SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[KVConnector] Migrate the LMCache integration code to be vLLM native (#25542)

Signed-off-by: ApostaC <yihua98@uchicago.edu>
Y
Yihua Cheng committed
83f478bb19489b41e9d208b47b4bb5a95ac171ac
Parent: 269c4db
Committed by GitHub <noreply@github.com> on 10/25/2025, 12:23:53 AM