Model-Based RL

RL: Reinforcement Learning

Momentum

24 papers in the last four weeks, up 243% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 142

All topics
CardsList
  1. RIWANav: Recursive World-Action Models with Self-Improvement for Urban Navigation

    Oct 6, 2026Jing Xie, Shouwei Ruan, Yubin Wang +4Robot Policy LearningWorld Model Learning

  2. Considering Context: When World Models Need Context Encoders

    Oct 5, 2026Oleg Smirnov, Sofiane Ennadir, John Pertoft +2Reinforcement LearningWorld Model Learning

  3. Mind the Execution Gap: Action-Semantic Mismatch in World-Model Control

    Oct 5, 2026Shengtao Wen, Xiang Chen, Yu Tian +3World Model-Based PlanningAction-Conditioned World Models

  4. Optimal Control with Learned Critics under Unmodeled State Dependencies

    Oct 4, 2026Philipp Schoch, Markus RyllResidual Dynamics LearningRobot Dynamics Learning

  5. R2R^2-WAM: Repair-and-Reject Post-Training for World Action Models

    Oct 4, 2026Ruiyan Xu, Haisheng Su, Sixu Lin +4World Action ModelsAction-Conditioned World Models

  6. In CEM, a World Model Is Also a Proposal Mechanism

    Oct 1, 2026Oliver Obst, Frieder StolzenburgBlack-Box OptimizationModel-Based Planning

  7. Lucid Dreaming for World Models: Learning to Doubt Imagination and Decide by Trust

    Sep 29, 2026Ziqi Wen, Ting Xu, Lianyu Wang +5Uncertainty QuantificationWorld Model-Based Planning

  8. Model-Informed Safe Reinforcement Learning for Bipedal Locomotion via Step-to-Step Prediction

    Sep 28, 2026Victor Paredes, Ayonga HereidHumanoid Robot LocomotionControl Barrier Functions

  9. MA-JEPA: Joint-Embedding World Models for Multi-Agent Reinforcement Learning

    Sep 27, 2026Brandon Gary Kaplowitz, Osaze James Obahor, Christian Schroeder de WittGame-Playing AgentsWorld Model Learning

  10. Think Fast, Plan Selectively: Adaptive Deliberation for Efficient Data-Driven MPC

    Sep 26, 2026Yi Xian Goh, Sze Jue Yang, Hao LuanData-Driven ControlModel Predictive Control

  11. Not All Errors Matter: Decision-Relevant Prediction Error Predicts Planning Quality

    Sep 26, 2026Linhao Wang, Yiyan Fan, Dongjin HuangWorld Model-Based PlanningModel-Based Planning

  12. Dual-Frontier: When Can an Agent Trust Its World Model?

    Sep 22, 2026Huatai Zhu, Qiang Chen, Ziqian Kou +5World Model-Based PlanningModel-Based Planning

  13. Imagine-RL: Residual-Confidence-Guided Cross-Attention for World-Model-Augmented VLA Reinforcement Learning

    Sep 21, 2026Kejia Hu, Wentong Zhai, Bo Zhao +1Robotic ManipulationVision-Language-Action Models

  14. Efficient Bayes-Adaptive Reinforcement Learning with Temporal Logic Specifications

    Sep 17, 2026Jonathan Hau, Alessandro AbateBayesian RLSafe RL

  15. A Convergence Framework for Deep VV-Learning: Error Propagation and Sharp Action-Gap Bounds

    Sep 16, 2026Yury KolomeytsevReinforcement LearningPolicy Learning

  16. Characterizing Replay Retention Under Dynamics Shift in Model-Based Reinforcement Learning

    Sep 16, 2026Everest Yang, Skye Thompson, George D. KonidarisNon-Stationary RLContinual Robot Learning

  17. ProxiDex: Learning Dynamics-Guided Proximity Policy for Dexterous Manipulation

    Sep 15, 2026Yushan Bai, Boyu Zheng, Zhiyang Mao +4Robot Policy LearningContact-Rich Robotic Manipulation

  18. Distributed Optimization of Modular Production Systems using Model-based Reinforcement Learning with Inverse Models

    Sep 12, 2026Andreas Schwung, Steve Yuwono, Sofiene Lassoued +1ManufacturingRL Control

  19. Amortized Low-Rank Adaptation for Model-Based Reinforcement Learning

    Sep 10, 2026Fernando Palafox, David Fridovich-KeilTest-Time AdaptationWorld Model Learning

  20. Near-Optimal Reinforcement Learning with Multi-Step Transition Lookahead

    Sep 10, 2026Corentin Pla, Hugo Richard, Marc Abeille +1Reinforcement LearningApproximation Algorithms

  21. HaWMPO: Hallucination-Aware World Model-based Policy Optimization for Generalist Robot Policy

    Sep 9, 2026Zengjue Chen, Peidong Liu, Jiawei Li +1Robot Policy LearningWorld Models

  22. CAST: Alternating State-Value Targets and Expanded Policy Gradients for Model-Based Reinforcement Learning

    Sep 8, 2026Pietro Noah Crestaz, Mohamed Yassine Kabouri, Nicolas Mansard +1RL for RoboticsReinforcement Learning

  23. World-Model-Augmented Visual Locomotion for Humanoids on Foothold-Constrained Terrain

    Sep 2, 2026Yuxi Liu, Lijun Han, Ziming Wang +3Humanoid Robot LocomotionTerrain-Aware Robot Locomotion

  24. NashDreamer: Model-Based Reinforcement Learning for Zero-Sum Imperfect-Information Games

    Sep 1, 2026Tomáš Holeček, Viliam LisýImperfect-Information GamesZero-Sum Games

  25. Reinforced Planning with Latent World Models

    Aug 19, 2026Armin Sommer, Jannik SchillingWorld Model-Based PlanningRobotic RL