cs.ROApr 16, 2025

A Graph-Based Reinforcement Learning Approach with Frontier Potential Based Reward for Safe Cluttered Environment Exploration

Authors: Gabriele CalzolariVidya SumathyChristoforos KanellakisGeorge Nikolakopoulos

Organizations: Luleå University of Technology

Abstract

Autonomous exploration of cluttered environments requires efficient exploration strategies that guarantee safety against potential collisions with unknown random obstacles. This paper presents a novel approach combining a graph neural network-based exploration greedy policy with a safety shield to ensure safe navigation goal selection. The network is trained using reinforcement learning and the proximal policy optimization algorithm to maximize exploration efficiency while reducing the safety shield interventions. However, if the policy selects an infeasible action, the safety shield intervenes to choose the best feasible alternative, ensuring system consistency. Moreover, this paper proposes a reward function that includes a potential field based on the agent's proximity to unexplored regions and the expected information gain from reaching them. Overall, the approach investigated in this paper merges the benefits of the adaptability of reinforcement learning-driven exploration policies and the guarantee ensured by explicit safety mechanisms. Extensive evaluations in simulated environments demonstrate that the approach enables efficient and safe exploration in cluttered environments.

Explore similar work

CardsList
  1. Safety-aware Skill Adaptation for Reinforcement Learning in Dynamic Environments

    Sep 11, 2026A K M Nadimul Haque, Sheila Sutjipto, Marc G. Carmichael +1Obstacle AvoidanceEnvironment

  2. Sampling-Based Safe Reinforcement Learning

    May 19, 2026Luca Vignola, Bruce D. Lee, Manish Prajapat +4Safety ConstraintsEfficient Exploration

  3. Easy-to-Use Shielding for Reinforcement Learning

    Jun 2, 2026Stefan Pranger, Bettina KönighoferShielding