Skip to content

RLLMNav: boosting object goal navigation through seamless integration of historical experience with large language model

Jul 2026 · International Journal of Machine Learning and Cybernetics · Vol 17 · 0 citations · 55 references
Computer Science

TL;DR

This work proposes RLLMNav, a method that integrates historical experience of RL with commonsense reasoning of LLM through a confidence-gated routing mechanism, and utilizes commonsense knowledge extracted from an LLM to suggest frontiers.

View source

Similar papers

Preprint Sep 2026

"Dear LLaVA, Please Drive": A Depth-Aware Vision-Language Agent for Closed-Loop Robotic Control

This work proposes a parameter-efficient approach to fine-tune a pretrained VLM for autonomous navigation using an Imperative Learning paradigm, and introduces a unified end-to-end navigation pipeline for natural-language-driven robotic control.

Sebastian Berger, Katharina Winter, Fabian B. Flohr · 0 citations
Preprint Sep 2026

HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness

Embodied navigation requires agents to ground instructions or object goals in spatial observations and translate plans into successful execution. As multimodal large language models (MLLMs) become increasingly capable, they offer stronger support for navigation without task-specific training; however, improved semantic...

Yang Chen, Li-Rong Che, Zhen-Yu Huang et al. · 5 citations · ⚡1
#reinforcement learning Conference Sep 2026

A hierarchical reinforcement learning approach for robot navigation integrating LiDAR priors and the options framework

This method formalizes the navigation task as a semiMarkov decision process and constructs a two-layer decision architecture with collaboration between a high-level manager and a low-level worker with collaboration between a high-level manager and a low-level worker.

Qi-Ming Chen · 0 citations
#machine learning Preprint Sep 2026

Prioritized Rollouts for Efficient World Model-based Vision-Language-Action Policy Optimization

Vision-Language-Action (VLA) models have emerged as a powerful paradigm for embodied intelligence, but fine-tuning them with reinforcement learning (RL) remains constrained by the cost of real-world robot interaction. Model-based reinforcement learning (MBRL) reduces this cost by using a learned world model to generate...

Yi-Fei Sheng, Hao-Xiang Ren, Zhilong Zhang et al. · 0 citations
Open access Aug 2026

DevGRU: Depth-Guided Visual Navigation Using a Collision-Aware Recurrent Model

The proposed DevGRU navigation system employs an action predictor that generates collision-aware future trajectories, enabling effective avoidance of immediate obstacles and has a relatively small number of trainable parameters, resulting in the fastest inference time among the baselines.

Kyung Min Han, Eunsom Kim, Young J. Kim · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.