SIGN IN SIGN UP

Q4_1 quantization (#193)

* Add AVX2 version of ggml_vec_dot_q4_1

* Small optimisations to q4_1 dot product (@Const-me)

* Rearrange Q4_1 quantization to work for multipart models. (Fix #152)

* Fix ggml_vec_mad_q4_1 too

* Fix non-vectorised q4_1 vec mul
M
Matvey Soloviev committed
904d2a8d6acd667c9633138d45a361d40fbf76d0
Parent: 7213110
Committed by GitHub <noreply@github.com> on 3/17/2023, 4:48:39 AM