Skip to content
arXiv cs.CV · Papers

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

arXiv:2607.24957v1 Announce Type: new Abstract: We introduce PerceptionBench, a benchmark specifically designed to evaluate the atomic visual perception capabilities of Multimodal Large Language Models (MLLMs). Existing benchmarks often fail to isolate perception: holistic evaluations conflate perceptual errors with fa