r/LocalLLaMA
· Communities
Caveman reasoning
Since I've seen this come up a couple of times with finetunes like Grug, I wonder how you feel about it now that official models have released that implement it (muse glimmer, deepseek v4 pro ga). I can understand it benefits agentic use as it wastes less tokens on the task. Personally I'm not a big fan as it seems to