SIGN IN SIGN UP

quantize: options for output and token embedding tensors qtype (#6239)

* quantize: be able to specify the output tensor type

* quantize: be able to specify the token embedding tensor type

---------

Co-authored-by: Iwan Kawrakow <iwan.kawrakow@gmail.com>
K
Kawrakow committed
1d0331c12a2f2a6296b471232bd4e66fbf06e6a1
Parent: dba1af6
Committed by GitHub <noreply@github.com> on 3/22/2024, 6:47:14 PM