Skip to content
llama.cpp releases · Infrastructure

b9874

cuda : concat implementation for quantized types (#25303) cuda : concat implementation for quantized types chore : apply am17an clever suggestion to shorten the code Co-authored-by: Stanisław Szymczyk sszymczy@gmail.com macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS