SIGN IN SIGN UP

A high-throughput and memory-efficient inference and serving engine for LLMs

0 0 122 Python

[NIXL][BUGFIX] delay done_recving queue cleanup to bottom of get_finished (#27297)

Signed-off-by: Chendi Xue <chendi.xue@intel.com>
C
Chendi.Xue committed
699d62e6cf4d9a570e2441141c65f687b3b1eef2
Parent: cd390b6
Committed by GitHub <noreply@github.com> on 10/24/2025, 5:01:41 PM