SIGN IN SIGN UP

gguf-py, convert-hf : model conversion support for T5 and FLAN-T5 model variants (#5763)

* gguf-py : add T5 model architecture

* gguf-py : add separate tensors for encoder and decoder

* gguf-py : add new model header parameters: decoder_start_token_id, attention.relative_buckets_count, tokenizer.ggml.remove_extra_whitespaces, tokenizer.ggml.precompiled_charsmap

* convert-hf : add model conversion support for T5ForConditionalGeneration and T5WithLMHeadModel

---------

Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com>
F
fairydreaming committed
de0d6a68ac99f307fe889c48e21124bc3b7ca29a
Parent: 95f57bb
Committed by GitHub <noreply@github.com> on 6/24/2024, 5:06:05 AM