r/LocalLLaMA
· Communities
Recent llama.cpp updates for SYCL/Intel
Some fixes & boost(pp) for SYCL/Intel. Merged PRs: [SYCL] Flash Attention with XMX engine via oneDNN graph API (SDPA) on KV f16 for Xe2 ; Qwen3.6-27b-Q8_0 prefill speed up x1.21 at p=512 and x4.26 at p=80k #25222 sycl: Increase minimum buffer size for USM system allocations #25525 [SYCL] Support OP XIELU #25550 [SYCL]