Skip to content
r/LocalLLaMA · Communities

I’ve had ling-3.0-flash and glm-5.2 both in my executor slot for a few weeks. They don’t split the way the benchmarks predict

Same harness, same task set, same agent scaffold, the only thing I swapped was the executor. Not a proper benchmark, no clean tok/s numbers, this is a workflow read not a leaderboard. glm-5.2 is the better model and it shows on anything that needs an actual decision. When the plan is loose or the step is ambiguous it f