Skip to content
r/LocalLLaMA · Communities

Caveman reasoning

Since I've seen this come up a couple of times with finetunes like Grug, I wonder how you feel about it now that official models have released that implement it (muse glimmer, deepseek v4 pro ga). I can understand it benefits agentic use as it wastes less tokens on the task. Personally I'm not a big fan as it seems to