SparseNav: Instruction-conditioned Sparse Semantic Perception for Training-Free Vision-Language Navigation
Map-based vision-language navigation (VLN) relies on persistent spatial representations to connect language understanding with geometric planning. However, acquiring semantics beyond the needs of the current instruction can introduce unnecessary perception cost and irrelevant annotations. Continuously accumulating unre...