SIGN IN SIGN UP

Add llama4 (#3145)

* initial changes

* Add support for other vlm

* cleanup comment

* Improve attn_implementation

* Add comments for support of models

* add model

* add model

* fixes and improvements

* update docker

* Add cache position

* Add tests

* remove redundant changes

* remove tr version

* Upgrade doc + fix linting.

* Fixing the CI.

---------

Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com>
M
Mohit Sharma committed
d9bb9bebc9eba54287c4f759cbef70543f63ef2f
Parent: 3d059f9
Committed by GitHub <noreply@github.com> on 4/6/2025, 8:20:22 AM