SIGN IN SIGN UP

llama : add abort_callback to interrupt computation (#5409)

* using abort_callback from ggml to stop llama computation

* format fix

* a brief explaining comment

---------

Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
M
Michael Podvitskiy committed
4a6e2d6142ab815c964924896891e9ab3e050632
Parent: 494c870
Committed by GitHub <noreply@github.com> on 3/2/2024, 7:52:25 PM