Q4_1 quantization (#193)
* Add AVX2 version of ggml_vec_dot_q4_1 * Small optimisations to q4_1 dot product (@Const-me) * Rearrange Q4_1 quantization to work for multipart models. (Fix #152) * Fix ggml_vec_mad_q4_1 too * Fix non-vectorised q4_1 vec mul
M
Matvey Soloviev committed
904d2a8d6acd667c9633138d45a361d40fbf76d0
Parent: 7213110
Committed by GitHub <noreply@github.com>
on 3/17/2023, 4:48:39 AM