llama : add abort_callback to interrupt computation (#5409)
* using abort_callback from ggml to stop llama computation * format fix * a brief explaining comment --------- Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
M
Michael Podvitskiy committed
4a6e2d6142ab815c964924896891e9ab3e050632
Parent: 494c870
Committed by GitHub <noreply@github.com>
on 3/2/2024, 7:52:25 PM