Back to search

Article

Large-Scale Model-Enhanced Vision-Language Navigation: Recent Advances, Practical Applications, and Future Challenges

2026-02-10

Abstract excerpt

The ability to autonomously navigate and explore complex 3D environments in a purposeful manner, while integrating visual perception with natural language interaction in a human-like way, represents a longstanding research objective in Artificial Intelligence (AI) and embodied cognition. Vision-Language Navigation (VLN) has evolved from geometry-driven to semantics-driven and, more recently, knowledge-driven appro...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
bd9dff26-3210-50a4-9589-fcd9eba5325a
DOI
10.20944/preprints202602.0768.v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Large-Scale Model-Enhanced Vision-Language Navigation: Recent Advances, Practical Applications, and Future ChallengesDOI 10.20944/preprints202602.0768.v1
Select a neighboring publication to make it the new centre.