quantize: options for output and token embedding tensors qtype (#6239)
* quantize: be able to specify the output tensor type * quantize: be able to specify the token embedding tensor type --------- Co-authored-by: Iwan Kawrakow <iwan.kawrakow@gmail.com>
K
Kawrakow committed
1d0331c12a2f2a6296b471232bd4e66fbf06e6a1
Parent: dba1af6
Committed by GitHub <noreply@github.com>
on 3/22/2024, 6:47:14 PM