backend : add eval callback (#4935)
* backend : add eval callback ggml-ci * backend : group nodes in a single compute when user don't need them * backend : clean-up the implementation ggml-ci * simple : do not perform tensor data copy if not needed * simple : fix * simple : no need for ggml_is_contiguous + fix bool parse * llama : fix callback placement in llama_context_params * backend : avoid double-ask callback calls * simple : restore examples, imatrix will serve as a demo
G
Georgi Gerganov committed
44a1a4a41a4c0b03afaa7d9e06bcbc7cf95aa1e6
Parent: c918fe8
Committed by GitHub <noreply@github.com>
on 1/17/2024, 4:39:41 PM