Add llama4 (#3145)
* initial changes * Add support for other vlm * cleanup comment * Improve attn_implementation * Add comments for support of models * add model * add model * fixes and improvements * update docker * Add cache position * Add tests * remove redundant changes * remove tr version * Upgrade doc + fix linting. * Fixing the CI. --------- Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com>
M
Mohit Sharma committed
d9bb9bebc9eba54287c4f759cbef70543f63ef2f
Parent: 3d059f9
Committed by GitHub <noreply@github.com>
on 4/6/2025, 8:20:22 AM