arXiv cs.CV
· Papers
VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models
arXiv:2512.22539v4 Announce Type: replace-cross Abstract: While Vision-Language-Action models (VLAs) are rapidly advancing toward generalist robot policies, quantitatively characterizing their capability boundaries and failure modes remains challenging. To address this, we introduce VLA-Arena, a comprehensive benchmark