ci : add test-llama-archs tensor split for Metal (#27598)
Run test-llama-archs with 1 to 4 GGML_METAL_DEVICES, mirroring the existing CUDA runs, and dispatch the job unconditionally since the per-backend guards now decide what to run. Assisted-by: llama.cpp:DeepSeek-v4-Flash-0731
G
Georgi Gerganov committed
95b8e33e16bb9a60de780a70930ebf729db6a90a
Parent: a278dce
Committed by GitHub <noreply@github.com>
on 8/23/2026, 12:57:07 PM