Skip to content
arXiv cs.AI · Papers

VoLN: Vision-Only Long-Horizon Navigation—Paradigm, Benchmark, and Method

arXiv:2607.21400v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) enables embodied agents to follow natural-language instructions. However, route-level instructions commonly encode spatial priors, such as orientation, distance, and layout, that are not explicitly available from onboard sensing at d