Skip to content
llama.cpp releases · Infrastructure

b10068

model: rotate injected K/V cache for DFlash (#25823) dflash: rotate injected K/V cache when using K/V quantization Update src/models/dflash.cpp Co-authored-by: Georgi Gerganov ggerganov@gmail.com clearer format remove trailing whitespace Co-authored-by: Georgi Gerganov ggerganov@gmail.com Website: https://llama.app mac