Partially Observable RL

RL: Reinforcement Learning

Latest papers 81

All topics
CardsList
  1. Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability

    May 14, 2026Yushen Liu, Yin-Jen Chen, Ziyi Chen +4Safety-Critical ControlPartially Observable RL

  2. On the Importance of Multistability for Horizon Generalization in Reinforcement Learning

    May 12, 2026Asad Bakija, Florent De Geeter, Julien Brandoit +2Dynamical SystemsReinforcement Learning

  3. Agent-BRACE: Decoupling Beliefs from Actions in Long-Horizon Tasks via Verbalized State Uncertainty

    May 12, 2026Joykirat Singh, Zaid Khan, Archiki Prasad +5Partially Observable RLBayesian Inference

  4. SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning

    May 8, 2026Nikunj Gupta, James Zachary Hare, Jesse Milzman +2Multi-Agent CoordinationPartially Observable RL

  5. Causal Reinforcement Learning for Complex Card Games: A Magic The Gathering Benchmark

    May 7, 2026Cristiano da Costa Cunha, Ajmal Mian, Tim French +1RL BenchmarksStructural Causal Models

  6. Hidden States as Value Gradients: The Pontryagin Structure of Recurrent Policies

    May 6, 2026David Leeftink, Max Hinne, Marcel van GervenRL ControlReinforcement Learning

  7. Recurrent Deep Reinforcement Learning for Chemotherapy Control under Partial Observability

    May 4, 2026Firas Mohamed Elamine Kiram, Imane Youkana, Rachida Saouli +2Reinforcement LearningHealthcare

  8. Spatial-Temporal Learning-Based Distributed Routing for Dynamic LEO Satellite Networks

    May 4, 2026Po-Heng Chou, Chiapin Wang, Shou-Yu Chen +1Deep Q-LearningPartially Observable RL

  9. Closed-Loop CO2 Storage Control With History-Based Reinforcement Learning and Latent Model-Based Adaptation

    May 4, 2026Sofianos Panagiotis Fotias, Vassilis GaganisRL ControlLatent Dynamics Modeling

  10. Forager: a lightweight testbed for continual learning with partial observability in RL

    May 1, 2026Steven Tang, Xinze Xiong, Anna Hakhverdyan +7RL BenchmarksLoss of Plasticity

  11. NASimJax: A GPU-Accelerated Policy Learning Framework for Penetration Testing

    Mar 20, 2026Raphael Simon, José Carrasquel, Elli Makdis Antoun +2RL BenchmarksPartially Observable RL

  12. Model-Free Output Feedback Stabilization via Policy Gradient Methods

    Jan 27, 2026Ankang Zhang, Ming Chi, Xiaoling Wang +1Dynamical SystemsPartially Observable RL

  13. Learning Robust Penetration Testing Policies under Partial Observability: A systematic evaluation

    Sep 24, 2025Raphael Simon, Pieter Libin, Wim MeesReinforcement LearningPartially Observable RL

  14. Quantum Bayesian Networks Can Speed up Reinforcement Learning in Partially Observable Environments

    Jul 24, 2025Gilberto Cunha, Alexandra Ramôa, André Sequeira +2Partially Observable RLBayesian RL

  15. SUB-PLAY: Adversarial Policies against Partially Observed Multi-Agent Reinforcement Learning Systems

    Feb 6, 2024Oubo Ma, Yuwen Pu, Linkang Du +5Adversarial AttacksPartially Observable RL

  16. Robust Recurrent Reinforcement Learning under Evolving Hidden Disturbances with Application to Rover Wheel Slip

    Jul 29, 2023Saki Omi, Hyo-Sang Shin, Namhoon Cho +2RL for RoboticsReinforcement Learning