SIGN IN SIGN UP

Q6_K AVX improvements (#10118)

* q6_k instruction reordering attempt

* better subtract method

* should be theoretically faster

small improvement with shuffle lut, likely because all loads are already done at that stage

* optimize bit fiddling

* handle -32 offset separately. bsums exists for a reason!

* use shift

* Update ggml-quants.c

* have to update ci macos version to 13 as 12 doesnt work now. 13 is still x86
E
Eve committed
340736477651095a98a3b10e19b038ec62593a1d
Parent: d5a409e
Committed by GitHub <noreply@github.com> on 11/4/2024, 10:06:31 PM