This paper introduces HarnessVLN, a training-free framework for embodied navigation that uses a unified tool interface to validate proposed actions against spatial evidence and task progress, allowing agents to generalize and learn from multimodal large language models.
Firehose
Filtered to Papers, tagged “spatial reasoning” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives