Sample-Efficient RL

RL: Reinforcement Learning

Momentum

12 papers in the last four weeks, up 140% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 168

All topics
CardsList
  1. Agile Reinforcement Learning through Separable Neural Architecture and Applications

    Jan 30, 2026Rajib Mostakim, Reza T. Batley, Sourav SahaRL ControlNeural Network Approximation Theory

  2. HeaPA: Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning

    Jan 30, 2026Weiqi Wang, Xin Liu, Binxuan Huang +13RL for Language ModelsRL for Language Model Reasoning

  3. CLEANER: Self-Purified Trajectories Boost Agentic Reinforcement Learning

    Jan 21, 2026Tianshi Xu, Yuteng Chen, Meng LiAgentic RLReinforcement Learning

  4. Reinforcement Learning in the Real World: A Survey of Statistical Challenges and Future Directions

    Jan 21, 2026Asim H. Gazi, Yongyi Guo, Daiqi Gao +3Non-Stationary RLReinforcement Learning

  5. Agentic Episodic Control

    Jun 2, 2025Xidong Yang, Wenhao Li, Junjie Sheng +4Agentic RLReinforcement Learning

  6. Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

    May 18, 2025Zachary Roch, George Atia, Yue WangReinforcement LearningAverage-Reward RL

  7. Adaptive Resolving Methods for Markov Decision Processes with Function Approximations

    May 17, 2025Jiashuo Jiang, Yinyu Ye, Yiming ZongMarkov Decision ProcessesReinforcement Learning

  8. Sample Complexity of Linear Quadratic Regulator Without Initial Stability

    Feb 20, 2025Amirreza Neshaei Moghaddam, Alex Olshevsky, Bahman GharesifardRL ControlSample Complexity

  9. Efficient Diversity-based Experience Replay for Deep Reinforcement Learning

    Oct 27, 2024Kaiyan Zhao, Yiming Wang, Yuyang Chen +3Experience ReplayDeterminantal Point Process

  10. Adaptive teachers for amortized samplers

    Oct 2, 2024Minsu Kim, Sanghyeok Choi, Taeyoung Yun +7Amortized InferenceRL Exploration

  11. Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning

    Sep 22, 2024Jia Wan, Sean R. Sinclair, Devavrat Shah +1Regret Minimization in RLMarkov Decision Processes

  12. ProSpec RL: Plan Ahead, then Execute

    Jul 31, 2024Liangliang Liu, Yi Guan, BoRan Wang +5Reinforcement LearningRisk-Sensitive RL

  13. Reflective Policy Optimization

    Jun 6, 2024Yaozhong Gan, Renye Yan, Zhe Wu +1Policy OptimizationPolicy Gradient Methods

  14. Provably Efficient Off-Policy Adversarial Imitation Learning with Convergence Guarantees

    May 26, 2024Yilei Chen, Vittorio Giammarino, James Queeney +1Adversarial Imitation LearningImitation Learning

  15. Distributionally Robust Reinforcement Learning with Interactive Data Collection: Fundamental Hardness and Near-Optimal Algorithms

    Apr 4, 2024Miao Lu, Han Zhong, Tong Zhang +1Distributionally Robust RLRobust RL

  16. ETHER: Aligning Emergent Communication for Hindsight Experience Replay

    Jul 28, 2023Kevin Yandoka Denamganaï, Daniel Hernandez, Ozan Vardal +2Reinforcement LearningExperience Replay

  17. Reinforcement Learning with Temporal-Logic-Based Causal Diagrams

    Jun 23, 2023Yash Paliwal, Rajarshi Roy, Jean-Raphaël Gaglione +5Reinforcement LearningCausal RL