Bayesian RL

RL: Reinforcement Learning

Momentum

2 papers in the last four weeks, against 1 the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 17

All topics
CardsList
  1. MaPP: A Unified Marginalized Posterior-Predictive Framework for Data-Efficient RLVR

    Sep 28, 2026Yangyang Ren, Haodong Zhu, Sheng Xu +4Bayesian RLRL for Language Model Reasoning

  2. Efficient Bayes-Adaptive Reinforcement Learning with Temporal Logic Specifications

    Sep 17, 2026Jonathan Hau, Alessandro AbateBayesian RLSafe RL

  3. Subspace Inference Enables Efficient Active Reward Learning from Preferences

    Sep 3, 2026Yutai Zhou, Erdem BıyıkBayesian RLPreference Learning

  4. Full Bayesian Reinforcement Learning via LF-IBIS

    Jul 2, 2026Stefano Masini, Cecilia Viscardi, Michela BacciniSimulation-Based InferenceReinforcement Learning

  5. Agentic Monte Carlo: Simulating Reinforcement Learning for Black-Box Agents

    Jun 3, 2026Dae Yon Hwang, Raunaq Suri, Valentin Villecroze +4Reinforcement LearningBayesian RL

  6. Bayesian learning for the stochastic shortest path problem

    Jun 3, 2026Chon Wai Ho, Sumeetpal S. Singh, Jiaqi GuoMarkov Decision ProcessesReinforcement Learning

  7. Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief

    May 30, 2026Hongqiang Lin, Pengfei Wang, Nenggan ZhengReinforcement LearningBayesian RL

  8. Information-Directed Offline-to-Online Reinforcement Learning

    May 28, 2026Keru ChenRegret Minimization in RLBayesian RL

  9. Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs

    May 23, 2026Meichen Song, Yuhao Wang, Enlu ZhouReinforcement LearningBayesian RL

  10. Bayesian policy gradient and actor-critic algorithms

    Apr 30, 2026Mohammad Ghavamzadeh, Yaakov Engel, Michal ValkoBayesian RLTemporal-Difference Learning

  11. Posterior Sampling Reinforcement Learning with Gaussian Processes for Continuous Control: Sublinear Regret Bounds for Unbounded State Spaces

    Mar 9, 2026Hamish Flynn, Joe Watson, Ingmar Posner +1RL ControlRegret Minimization in RL

  12. A Model-Free Universal AI

    Feb 26, 2026Yegon Kim, Juho LeeReinforcement LearningBayesian RL

  13. Meta-RL with Bayesian Linear Task Models

    Dec 24, 2025Jingyang You, Hanna KurniawatiRepresentation LearningReinforcement Learning

  14. Quantum Bayesian Networks Can Speed up Reinforcement Learning in Partially Observable Environments

    Jul 24, 2025Gilberto Cunha, Alexandra Ramôa, André Sequeira +2Partially Observable RLBayesian RL

  15. Fully Offline Reinforcement Learning

    May 28, 2025Mattie Fellows, Clarisse Wibault, Uljad Berdica +3Bayesian RLOffline RL

  16. Thompson Sampling for Infinite-Horizon Discounted Decision Processes

    May 14, 2024Daniel Adelman, Cagla Keceli, Alba V. Olivares-NadalThompson SamplingRegret Minimization in RL