r/LocalLLaMA
· Communities
qwen agentworld can self-correct in reasoning traces
decided to mess around with it to see how the world model training affects it, found a system prompt that massively improves reasoning: predict your own response, then analyze your prediction for any errors. Use the analysis to craft the final response. respond only with the final response, and do not repeat thoughts.