Skip to content
arXiv cs.CV · Papers

VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models

arXiv:2512.22539v4 Announce Type: replace-cross Abstract: While Vision-Language-Action models (VLAs) are rapidly advancing toward generalist robot policies, quantitatively characterizing their capability boundaries and failure modes remains challenging. To address this, we introduce VLA-Arena, a comprehensive benchmark