llama.cpp releases
· Infrastructure
b9788
sycl : support --split-mode tensor (#24152) Sycl tp stage1 (#1) SYCL: tensor parallelism (--split-mode tensor) for dual-GPU Adds the comm_init/comm_free/comm_allreduce_tensor trio that the meta-backend queries via get_proc_address to enable backend-specific all-reduce, mirroring the pattern used by ggml-cuda.cu. For N=