Skip to content
r/LocalLLaMA · Communities

llamacpp performing slower then Ollama

Hi. So I just setup llamacpp for the first time. I'm using the model : "Huihui-Qwen3.6-35B-A3B-abliterated-ggml-model-Q4_K.gguf". When I test this in llamacpp server GUI I get about 55tps, while in ollama default GUI i get about 61tps. (Tho Prompt processing is slower in ollama, overall ollama is still faster) Im using