llama : quantize up to 31% faster on Linux and Windows with mmap (#3206)
* llama : enable mmap in quantize on Linux -> 31% faster * also enable mmap on Windows --------- Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
C
Cebtenzzre committed
2777a84be429401a2b7d33c2b6a4ada1f0776f1b
Parent: 0a4a4a0
Committed by GitHub <noreply@github.com>
on 9/29/2023, 1:48:45 PM