arXiv cs.AI
· Papers
VoLN: Vision-Only Long-Horizon Navigation—Paradigm, Benchmark, and Method
arXiv:2607.21400v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) enables embodied agents to follow natural-language instructions. However, route-level instructions commonly encode spatial priors, such as orientation, distance, and layout, that are not explicitly available from onboard sensing at d