Partially Observable RL

RL: Reinforcement Learning

Latest papers 81

All topics
CardsList
  1. Robustness of Reinforcement Learning-Based Congestion Management in Low-Voltage Grids

    Jul 17, 2026Josef Hoppe, Sarra Bouchkati, Farah Nasr +7Partially Observable RLRobust RL

  2. Flow-aware Optimal Navigation in Unsteady Flows through Reinforcement Learning

    Jul 15, 2026Andrea Maria Braghin, Nicolò Botteghi, Matteo Tomasetto +2Partially Observable RLRobot Navigation

  3. EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents

    Jul 7, 2026Mianqiu Huang, Taofeng Xue, Chong Peng +12RL for GUI AgentsPartially Observable RL

  4. Relational Multi-Agent Reinforcement Learning for Dynamic Pricing in High-Speed Railway Markets

    Jul 6, 2026Enrique Adrian Villarrubia-Martin, David Muñoz-Valero, Luis Rodriguez-Benitez +2Multi-Agent System OptimizationIntelligent Transportation Systems

  5. ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability

    Jul 2, 2026Juarez Monteiro, Nathan Gavenski, Guilherme Lima +3Partially Observable RLLLM Prompting

  6. HyPOLE: Hyperproperty-Guided Multi-Agent Reinforcement Learning under Partial Observation

    Jun 29, 2026Arshia Rafieioskouei, Tzu-Han Hsu, Matthew Lucas +1Reinforcement LearningPartially Observable RL

  7. Adversarial observations in probabilistic State-Space Models for robust Reinforcement Learning

    Jun 18, 2026M. Santos-Pascual, D. Ríos InsuaAdversarial AttacksPartially Observable RL

  8. Direct Advantage Estimation for Scalable and Sample-efficient Deep Reinforcement Learning

    Jun 18, 2026Hsiao-Ru Pan, Bernhard SchölkopfPartially Observable RLSample-Efficient RL

  9. Learning-Based Decision Making for Combustion Phasing Control in Multi-Fuel CI Engines with Latent Fuel Reactivity Estimation

    Jun 16, 2026Rajasree Sarkar, Aditya Satish Patil, Arunava Banerjee +4Reinforcement LearningPartially Observable RL

  10. Learning Red Agent Policy from Observations for Neurosymbolic Autonomous Cyber Agents

    Jun 16, 2026Ankita Samaddar, Sandeep Neema, Daniel Balasubramanian +1Opponent ModelingAutonomous Cyber Defense

  11. Minimax-Optimal Policy Regret in Partially Observable Markov Games

    Jun 1, 2026Raman AroraRegret Minimization in RLImperfect-Information Games

  12. Why Linear Recurrent Memory Works in Partially Observable Reinforcement Learning

    May 29, 2026Yike Zhao, Onno Eberhard, Malek Khammassi +2Partially Observable RLMarkov Models

  13. The Challenges of Using Reinforcement Learning for Controlling Industrial Energy Systems

    May 29, 2026Tobias Lademann, Théo Vincent, Jan Peters +1RL ControlPartially Observable RL

  14. Commit to the Bit: Reactive Reinforcement Learning Done Right

    May 27, 2026Onno Eberhard, Claire Vernade, Michael MuehlebachReinforcement LearningPartially Observable RL

  15. Grow-Prune-Freeze Networks: Adaptive & Continual Learning Technique for Olfactory Navigation

    May 24, 2026Kordel K. France, Ovidiu DaescuNon-Stationary RLPartially Observable RL

  16. Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning

    May 23, 2026Noah Farr, Aryaman Reddi, Carlo D'Eramo +1Reinforcement LearningPartially Observable RL

  17. Short-Term-to-Long-Term Memory Transfer for Knowledge Graphs under Partial Observability

    May 21, 2026Taewoon Kim, Vincent François-Lavet, Michael CochezPartially Observable RLQ-Learning

  18. Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents

    May 19, 2026Wenjie Tang, Minne Li, Sijie Huang +2Reinforcement LearningPartially Observable RL

  19. JAXenstein: Accelerated Benchmarking for First-Person Environments

    May 19, 2026Ruo Yu Tao, George KonidarisRL BenchmarksPartially Observable RL

  20. Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

    May 19, 2026Nasehatul Mustakim, Lucas LehnertPartially Observable RLOOD Generalization

  21. Learning to Hand Off: Provably Convergent Workflow Learning under Interface Constraints

    May 18, 2026Jiayu Li, Enpei Zhang, Dawei Zhou +2Sequential MARLSequential Decision Making

  22. When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

    May 18, 2026Peiying Zhu, Sidi ChangPartially Observable RLDynamic Pricing

  23. Privacy Preserving Reinforcement Learning with One-Sided Feedback

    May 18, 2026Lin William Cong, Guangyan Gan, Hanzhang Qin +1Reinforcement LearningPartially Observable RL

  24. Dynamic Plasma Shape Control with Arbitrary Sensor Subsets

    May 15, 2026D. Sorokin, M. Stokolesov, A. Granovskiy +7RL ControlPartially Observable RL

  25. DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments

    May 15, 2026Anindya Sarkar, Srikumar Sastry, Aleksis Pirinen +2Visual NavigationAerial Robotics

  26. CaMeRL: Collision-Aware and Memory-Enhanced Reinforcement Learning for UAV Navigation in Multi-Scale Obstacle Environments

    May 14, 2026Hong Hong, Feiyu Liao, Yongheng Liang +3RL for RoboticsReinforcement Learning

  27. Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning

    May 14, 2026Luca Marzari, Enrico MarchesiniNeural Network VerificationPartially Observable RL