AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning

📅 2026-06-09
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the challenge of long-term robot navigation in dynamic environments, where persistent scene understanding and generalization capabilities are often lacking. To this end, the authors propose AllDayNav, a novel framework that introduces implicit, memory-driven reinforcement learning to lifelong navigation for the first time. Departing from conventional explicit maps or scene graphs, AllDayNav employs a self-evolving multimodal memory mechanism that integrates visual keyframes, semantic descriptions, and temporal context. Leveraging large language models, the system autonomously generates open-vocabulary instructions, image-based goals, and structured rewards. Experiments demonstrate that AllDayNav achieves near-perfect success rates in both simulated and real-world environments, significantly outperforming strong baselines—including map-based, vision-language, and reinforcement learning approaches—in terms of path efficiency and robustness, while enabling cross-room, cross-task, and cross-time navigation.
📝 Abstract
Lifelong embodied navigation in dynamic environments requires robots to form persistent scene understanding from fragmentary observations, which remains difficult for existing methods that rely on explicit maps or scene graphs and struggle to generalize beyond structured settings. We propose AllDayNav, a lifelong self-learning navigation framework that implicitly encodes scene dynamics into the billion-scale parameters of a large model via reinforcement learning, powered by a self-evolving multimodal memory that maintains and updates visual keyframes, semantic descriptions, and temporal context while autonomously generating open-vocabulary instructions, image goals, and structured rewards. Experiments in both synthetic and real-world environments across cross-room, cross-episode, and cross-task scenarios show that AllDayNav achieves success rates approaching $100\%$ and consistently surpasses strong map-based, VLM, and RL baselines in path efficiency and robustness, demonstrating implicit, memory-driven reinforcement learning as a scalable alternative to explicit mapping for reliable lifelong navigation.
Problem

Research questions and friction points this paper is trying to address.

lifelong navigation
embodied AI
dynamic environments
persistent scene understanding
fragmentary observations
Innovation

Methods, ideas, or system contributions that make the work stand out.

lifelong navigation
implicit scene representation
self-evolving multimodal memory
reinforcement learning
open-vocabulary instruction
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
Hang Yin
Hang Yin
Department of Mathematical Sciences, Tsinghua University
machine learningknowledge graphsknowledge representation and reasoning
Y
Yinan Liang
Tsinghua University, Beijing 100084, China; Galbot Robotics, Beijing, China
Jiazhao Zhang
Jiazhao Zhang
Peking University
Embodied AINavigation3D Vision
J
Jiahang Liu
Galbot Robotics, Beijing, China
M
Minghan Li
Galbot Robotics, Beijing, China
Z
Zhizheng Zhang
Galbot Robotics, Beijing, China; Beijing Academy of Artificial Intelligence, Beijing, China
He Wang
He Wang
Assistant Professor of Computer Science, Peking University
Embodied AIComputer VisionRobotics