SIGN IN SIGN UP

imatrix : offload to GPU support (#4957)

* backend : add eval callback

ggml-ci

* backend : group nodes in a single compute when user don't need them

* backend : clean-up the implementation

ggml-ci

* simple : do not perform tensor data copy if not needed

* simple : fix

* imatrix : offload to GPU support

* imatrix : fix ggml_mul_mat_id hanlding

ggml-ci

* ci : add imatrix test

ggml-ci

* ci : rearrange output

ggml-ci
G
Georgi Gerganov committed
ba69bbc84ced580fe4fdb0713ca2d95634325b7a
Parent: 44a1a4a
Committed by GitHub <noreply@github.com> on 1/17/2024, 4:46:30 PM