arXiv cs.AI
· Papers
ABot-N1: Toward a General Visual Language Navigation Foundation Model
arXiv:2607.10383v3 Announce Type: replace-cross Abstract: Visual Language Navigation foundation models aim to unify deep reasoning for grounded spatial decisions with broad versatility for diverse embodied tasks. Current approaches typically achieve this integration via monolithic policies that map observations directl