r/LocalLLaMA
· Communities
Best choice of model 40B+ Parameters
currently using Qwen3.6 35B as my main assistant model + coding agent but I think sometimes it misses basical general knowledge things, and it is more like executioner that assistant. That's why I though should I go with bigger models, But I don't want to lose speed I am on Strix Halo Having 30-40 t/s roughly on 131k c