cs.ROOct 5, 2026

End-to-End Safe Social Navigation via Multi-Task Reinforcement Learning and Probabilistic Perception

Authors: Tommaso Van Der Meer, Andrea Garulli, Antonio Giannitrapani, Renato Quartullo, Alberto Vaglio, Alexandre Alahi

Organizations: Dipartimento di Ingegneria dell’Informazione e Scienze Matematiche, Universit`a di Siena, Siena, Italy · Uninettuno University, Rome, Italy · Visual Intelligence for Transportation (VITA) Lab, ´Ecole Polytechnique F´ed´erale de Lausanne (EPFL), Lausanne, Switzerland

Abstract

Autonomous social navigation requires balancing efficiency, physical safety, and social compliance. Reinforcement Learning (RL) methods provide a viable and effective solution but often rely on unrealistic assumptions, such as the knowledge of humans' position and velocity. In this paper, we introduce JESSI (JAX-based E2E Safe Social Interpretable navigation), a lightweight end-to-end RL framework that maps raw LiDAR scans directly to kinematically feasible control commands. JESSI enhances safety via Dirichlet-parameterized continuous action spaces and deterministic bounding, while an integrated attention-based perception module extracts probabilistic human states for interpretable, socially aware decision-making. Through extensive simulations and real-world deployment on a differential-drive robot, we demonstrate that jointly optimizing the RL policy with a supervised perception signal in a multi-task paradigm enhances social behavior. Ultimately, JESSI is able to balance high navigation success rates and superior social behaviors compared to state-of-the-art baselines.

Figures & tables

Explore similar work

CardsList
  1. KinematicRL: A Sim-to-Real Reinforcement Learning Framework For Social Navigation With Kinodynamic Feasibility

    Jun 10, 2026Zhiming Xu, Haodong Yang, Chengju Liu +2Robot NavigationSim-To-Real Reinforcement Learning

  2. Learning Social Robot Navigation By Sensing Human Legs

    Jul 30, 2026Alberto Vaglio, Andrea Garulli, Antonio Giannitrapani +2Robot NavigationTwo-Dimensional Light Detection And Ranging

  3. Think When It Matters: Conditional VLM Reasoning for Social Navigation with RL Policies

    Jul 13, 2026Ali Ahmadi, Hamed Rahimi, Adrien Jacquet Cretides +3Robot NavigationSafe Navigation