llama.cpp releases
· Infrastructure
b10068
model: rotate injected K/V cache for DFlash (#25823) dflash: rotate injected K/V cache when using K/V quantization Update src/models/dflash.cpp Co-authored-by: Georgi Gerganov ggerganov@gmail.com clearer format remove trailing whitespace Co-authored-by: Georgi Gerganov ggerganov@gmail.com Website: https://llama.app mac