cs.ROAug 26, 2024

A Survey on Reinforcement Learning Applications in SLAM

Authors: Mohammad Dehghani Tezerjani, Mohammad Khoshnazar, Mohammadhamed Tangestanizadeh, Arman Kiani, Qing Yang

Organizations: Computer Science and Engineering, University of North Texas, Denton, USA · Institute of Artificial Intelligence, University of Bremen, Bremen, Germany · Computer Science and Engineering, University of California Santa Cruz, Santa Cruz, USA · Electrical and Computer Engineering, University of Maine, Maine, USA

Abstract

Simultaneous localization and mapping (SLAM) allows a mobile robot or autonomous vehicle to build a map of an unknown environment while estimating its own pose within that map. Reinforcement learning (RL), in which an agent learns a decision policy from interaction and reward, has been applied to decide how such systems move, explore, and recognize places they have visited before. This survey reviews the applications of RL in SLAM. We first distinguish passive SLAM, in which the robot's motion is not chosen by the SLAM system, from active SLAM, in which it is, and summarize the sensors that provide the input to SLAM. We then introduce the RL methods used in this literature, from value-based and policy-based methods to actor-critic and deep RL. Next, we classify RL applications in SLAM into three categories: path planning, including environment exploration and obstacle avoidance; loop closure detection; and active SLAM. Thirteen representative studies are compared in terms of their simulation environment, deep learning method, SLAM method, and RL algorithm. Most of these studies are evaluated mainly in simulation, and value-based methods from the deep Q-network family are the most common. Finally, we discuss the challenges of applying RL to SLAM, namely computational demands, safety, generalization, high-dimensional state and action spaces, sample efficiency, and sensor and actuator delays, and we outline directions for future research.

Figures & tables

Explore similar work

CardsList
  1. SLAM as a Stochastic Control Problem with Partial Information: Optimal Solutions and Rigorous Approximations

    Apr 23, 2026Ilir Gusija, Fady Alajaji, Serdar YükselSimultaneous Localization And MappingPartially Observable Markov Decision Process

  2. SLAMSqueezeBench: Comparing SLAM Systems under Resource Constraints

    Sep 17, 2026Mohamed Hefny, Karthik Dantu, Steven Y. KoSimultaneous Localization And Mapping

  3. Why does Deep Learning Improve Visual SLAM?

    Jul 7, 2026Giovanni Cioffi, Davide ScaramuzzaOrb-Slam2Robotic Perception