r/LocalLLaMA
· Communities
Qwen 3.8-27b unusable long thinking?
I have a small coding test, where I ask a model to implement a simple CLI from a spec file. Qwen3.6-27b can do it in ~50k tokens. Qwen3.8-27b uses an absurd amount of thinking. I did a few runs, but never finished a single one, because after 50k tokens it usually didn't even finish the planning phase. I only have a sin