Benchmarking Reinforcement Learning Techniques for Autonomous Navigation

VIEW PUBLICATION

Zifan Xu*

Bo Liu*

Xuesu Xiao*

Anirudh Nair*

Peter Stone

* External authors

ICRA 2023

2023

Abstract

Deep reinforcement learning (RL) has broughtmany successes for autonomous robot navigation. However,there still exists important limitations that prevent real-worlduse of RL-based navigation systems. For example, most learningapproaches lack safety guarantees; and learned navigationsystems may not generalize well to unseen environments.Despite a variety of recent learning techniques to tackle thesechallenges in general, a lack of an open-source benchmarkand reproducible learning methods specifically for autonomousnavigation makes it difficult for roboticists to choose whatlearning methods to use for their mobile robots and for learningresearchers to identify current shortcomings of general learningmethods for autonomous navigation. In this paper, we identifyfour major desiderata of applying deep RL approaches forautonomous navigation: (D1) reasoning under uncertainty, (D2)safety, (D3) learning from limited trial-and-error data, and (D4)generalization to diverse and novel environments. Then, weexplore four major classes of learning techniques with thepurpose of achieving one or more of the four desiderata:memory-based neural network architectures (D1), safe RL (D2),model-based RL (D2, D3), and domain randomization (D4). Bydeploying these learning techniques in a new open-source large-scale navigation benchmark and real-world environments, weperform a comprehensive study aimed at establishing to whatextent can these techniques achieve these desiderata for RL-based navigation systems

Related Publications

Human-Interactive Robot Learning: Definition, Challenges, and Recommendations

THRI, 2025
Kim Baraka, Ifrah Idrees, Taylor Kessler Faulkner, Erdem Biyik, Serena Booth*, Mohamed Chetouani, Daniel Grollman, Akanksha Saran, Emmanuel Senft, Silvia Tulli, Anna-Lisa Vollmer, Antonio Andriella, Helen Beierling, Tiffany Horter, Jens Kober, Isaac Sheidlower, Matthew Taylor, Sanne van Waveren, Xuesu Xiao*

Robot learning from humans has been proposed and researched for several decades as a means to enable robots to learn new skills or adapt existing ones to new situations. Recent advances in artificial intelligence, including learning approaches like reinforcement learning and…

ProtoCRL: Prototype-based Network for Continual Reinforcement Learning

RLC, 2025
Michela Proietti*, Peter R. Wurman, Peter Stone, Roberto Capobianco

The purpose of continual reinforcement learning is to train an agent on a sequence of tasks such that it learns the ones that appear later in the sequence while retaining theability to perform the tasks that appeared earlier. Experience replay is a popular method used to mak…

Automated Reward Design for Gran Turismo

NeurIPS, 2025
Michel Ma, Takuma Seno, Kaushik Subramanian, Peter R. Wurman, Peter Stone, Craig Sherstan

When designing reinforcement learning (RL) agents, a designer communicates the desired agent behavior through the definition of reward functions - numerical feedback given to the agent as reward or punishment for its actions. However, mapping desired behaviors to reward func…

SEE ALL

HOME
Publications
Benchmarking Reinforcement Learning Techniques for Autonomous Navigation

JOIN US

Shape the Future of AI with Sony AI

We want to hear from those of you who have a strong desire
to shape the future of AI.

LEARN MORE