Skip to content
arXiv cs.CL · Papers

CC-OCR V2: Fine-Grained Attribution of LMM Failures in Real-World Visual Document Understanding

arXiv:2605.03903v2 Announce Type: replace Abstract: Recent Large Multimodal Models (LMMs) have achieved remarkable progress on OCR-centric document understanding and processing tasks. Existing benchmarks primarily evaluate LMMs across diverse tasks to reflect practical document-processing workflows or analyze how docum