Reward Learning

Recent momentum

emerging

2 papers in the last 28 days · 0.1% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-14

2 new papers

A weekly snapshot of new work published in Reward Learning.

Period ending 2026-09-07

2 new papers

A weekly snapshot of new work published in Reward Learning.

35 papers

Latest in Reward Learning

Open your feed →
CardsList
  1. Efficient Exploration Is Enough

    Sep 7, 2026Mikel Malagón, Jon Vadillo, Josu Ceberio +2Efficient ExplorationReward Learning

  2. Can In-Context Learning Support Intrinsic Curiosity?

    Jun 17, 2026Eric Elmoznino, Sangnie Bhardwaj, Johannes von Oswald +5In-Context LearningReward Learning

  3. A No-Regret Framework for Adaptive Incentive Design

    Jun 1, 2026Georgios Vasileiou, Lantian Zhang, Silun ZhangIncentivesOptimal Policies

  4. Exploratory Experience Shapes the Geometry of Predictive Representations

    May 27, 2026Kseniia Shilova, Abdelrahman Sharafeldin, Advay Balakrishnan +1Predictive CodingPerception-Action Loop

  5. Finite-Time Regret Analysis of Retry-Aware Bandits

    May 20, 2026Bingkui Tong, Junpei Komiyama, Soichiro Nishimori +1Sublinear RegretRegret

  6. A single algorithm for both restless and rested rotting bandits

    Apr 23, 2026Julien Seznec, Pierre Ménard, Alessandro Lazaric +1BanditsReward Learning

  7. Best of both worlds: Stochastic & adversarial best-arm identification

    Apr 16, 2026Yasin Abbasi-Yadkori, Peter L. Bartlett, Victor Gabillon +2Generalized Linear BanditBandits