llama : offload to RPC in addition to other backends (#7640)
* llama : offload to RPC in addition to other backends * - fix copy_tensor being called on the src buffer instead of the dst buffer - always initialize views in the view_src buffer - add RPC backend to Makefile build - add endpoint to all RPC object names * add rpc-server to Makefile * Update llama.cpp Co-authored-by: slaren <slarengh@gmail.com> --------- Co-authored-by: slaren <slarengh@gmail.com>
R
Radoslav Gerganov committed
bde7cd3cd949c1a85d3a199498ac98e78039d46f
Parent: a5735e4
Committed by GitHub <noreply@github.com>
on 6/3/2024, 5:03:26 PM