SIGN IN SIGN UP

llama : only copy used KV cache in get / set state (#1272)

* llama : only copy used KV cache in get / set state

* switch to ggml for copying k, v

* avoid designated initializers
E
Evan Jones committed
e216aa04633892b972d013719e38b59fd4917341
Parent: 2485d7a
Committed by GitHub <noreply@github.com> on 5/3/2023, 2:26:13 AM