SIGN IN SIGN UP

llama : offload to RPC in addition to other backends (#7640)

* llama : offload to RPC in addition to other backends

* - fix copy_tensor being called on the src buffer instead of the dst buffer

- always initialize views in the view_src buffer

- add RPC backend to Makefile build

- add endpoint to all RPC object names

* add rpc-server to Makefile

* Update llama.cpp

Co-authored-by: slaren <slarengh@gmail.com>

---------

Co-authored-by: slaren <slarengh@gmail.com>
R
Radoslav Gerganov committed
bde7cd3cd949c1a85d3a199498ac98e78039d46f
Parent: a5735e4
Committed by GitHub <noreply@github.com> on 6/3/2024, 5:03:26 PM