Skip to content
r/LocalLLaMA · Communities

A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone

Liquid AI released LFM2.5-2.6B today, and this might be more relevant to local AI than another massive model most people cannot run. The model is only 2.69B parameters, has 128K context, supports tool calling and was post-trained specifically for multi-step agent workflows. The official Q4_K_M GGUF is around 1.67 GB an