Skip to content
llama.cpp releases · Infrastructure

b9968

opencl: add int8 dp4 dense and MoE prefill optimization for Adreno GPUs (#25537) opencl: add int8 dp4 dense and moe GEMM opencl: refactor Co-authored-by: Li He lih@qti.qualcomm.com macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubun