SIGN IN SIGN UP

llama : add `gemma` model (#5631)

There are couple things in this architecture:

1. Shared input and output embedding parameters.
2. Key length and value length are not derived from `n_embd`.

More information about the models can be found at
https://ai.google.dev/gemma. GGUFs can be downloaded from
https://huggingface.co/google.
P
postmasters committed
580111d42b3b6ad0a390bfb267d6e3077506eb31
Parent: 88c46cb
Committed by GitHub <noreply@github.com> on 2/21/2024, 1:08:22 PM