llama.cpp releases
· Infrastructure
b10201
ggml-webgpu: improve flash_attn_vec for quantized KV at long contexts (#25956) improve fa of quantized kv cache Fix some bugs and some comments. fix v type check and some comments Fix build error caused by rebasing editorconfig checking pass Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple