arXiv cs.CL
· Papers
Self-Correction Bench: Uncovering and Addressing the Self-Correction Blind Spot in Large Language Models
arXiv:2507.02778v3 Announce Type: replace Abstract: Although large language models (LLMs) have transformed AI, they still make errors and follow unproductive reasoning paths. Self-correction is vital for safety-critical applications, but studying it requires disentangling activation failure from knowledge deficiency: w