arXiv cs.CV
· Papers
PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models
arXiv:2607.24957v1 Announce Type: new Abstract: We introduce PerceptionBench, a benchmark specifically designed to evaluate the atomic visual perception capabilities of Multimodal Large Language Models (MLLMs). Existing benchmarks often fail to isolate perception: holistic evaluations conflate perceptual errors with fa