Skip to content
llama.cpp releases · Infrastructure

b10206

llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized (#25871) llama : enforce the same K and V cache types for DeepSeek V4; enable FA if V cache is quantized llama : enforce the same K and V cache types for MLA models Co-authored-by: Stanisław Szymczyk sszymczy@gmail.com Websi