This paper introduces HarnessVLN, a training-free framework for embodied navigation that uses a unified tool interface to validate proposed actions against spatial evidence and task progress, allowing agents to generalize and learn from multimodal large language models.
Firehose
Filtered to Papers, tagged “training-free methods” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives