cs.ROSep 30, 2026

RL-Guided PAC-NMPC for Probabilistically-Safe Perception-Based Navigation in Unknown Environments

Authors: Adam Polevoy, Dillon Capalongo, Katherine Tang, Mark Gonzales, Marin Kobilarov, Joseph Moore

Organizations: Johns Hopkins University Applied Physics Laboratory, Laurel, MD 20723, USA. · Department of Mechanical Engineering, Johns Hopkins University, Baltimore, MD 21218, USA.

Abstract

In this paper, we present an approach for combining stochastic nonlinear model predictive control (SNMPC) and reinforcement learning (RL) to enable probabilistically-safe perception-based navigation in unknown environments. Our method first uses RL to train probabilistic actor-critic and sensor prediction models. We then leverage these probabilistic models in a sampling-based SNMPC framework known as Probably Approximately Correct (PAC)-NMPC, which uses hard constraints to enforce finite-time statistical guarantees on the probability of collision and value function improvement. By ensuring that our finite-horizon SNMPC policies decrease the value function in expectation, we can approach the long-horizon performance of the RL approach while satisfying probabilistic safety constraints. Through simulation experiments, we show that our approach can improve the safety of perception-based RL navigation policies and scale to high dimensional systems with large sensor input spaces and complex nonlinear dynamics. We also demonstrate our approach through hardware experiments, showing improved performance for vision-based navigation with an agile fixed-wing aerial vehicle in unknown environments.

Figures & tables

Explore similar work

CardsList
  1. Online, Reachability-Aware, Sampling-Based Motion Planning

    Sep 8, 2026Brendan Gould, Zhiyuan Zhang, Panagiotis Tsiotras +1Model Predictive ControlMotion Planning

  2. Probabilistic Recursively Feasible Motion Planning Under Uncertain Environments

    May 18, 2026Hyeontae Sung, Hyeongchan Ham, Junyoung Park +2Model Predictive ControlMotion Planning

  3. Interactive Trajectory Planning with Learning-based Distributionally Robust Model Predictive Control and Markov Systems

    May 8, 2026Erik Börve, Nikolce Murgovski, Morteza Haghir Chehreghani +1Model Predictive ControlRobust Trajectory