imatrix : offload to GPU support (#4957)
* backend : add eval callback ggml-ci * backend : group nodes in a single compute when user don't need them * backend : clean-up the implementation ggml-ci * simple : do not perform tensor data copy if not needed * simple : fix * imatrix : offload to GPU support * imatrix : fix ggml_mul_mat_id hanlding ggml-ci * ci : add imatrix test ggml-ci * ci : rearrange output ggml-ci
G
Georgi Gerganov committed
ba69bbc84ced580fe4fdb0713ca2d95634325b7a
Parent: 44a1a4a
Committed by GitHub <noreply@github.com>
on 1/17/2024, 4:46:30 PM