Group Relative Policy Optimization

Recent momentum

emerging

0 papers in the last 28 days · 0.0% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this field, kept on the site without email delivery.

Period ending 2026-09-21

14 new papers

A weekly snapshot of new work published in Group Relative Policy Optimization.

Period ending 2026-09-14

6 new papers

A weekly snapshot of new work published in Group Relative Policy Optimization.

Period ending 2026-09-07

21 new papers

A weekly snapshot of new work published in Group Relative Policy Optimization.

Inside this field

Focused directions

532 papers

Latest in Group Relative Policy Optimization

  1. Poly-EPO: Training Exploratory Reasoning Models

    Apr 19, 2026Ifdita Hasan Orney, Jubayer Ibn Hamid, Shreya S Ramanujam +5Multi-Objective Policy OptimizationEfficient Exploration

  2. TADPO: Reinforcement Learning Goes Off-road

    Mar 6, 2026Zhouchonghao Wu, Raymond Song, Vedant Mundheda +3Proximal Policy OptimizationTerrain

  3. Constrained Group Relative Policy Optimization

    Feb 5, 2026Roger Girgis, Rodrigue de Schaetzen, Luke Rowe +3Group Relative Policy Optimization

  4. LLP: LLM-Based Product Pricing in E-commerce

    Oct 10, 2025Hairu Wang, Sheng You, Qiheng Zhang +5PricesProducts

  5. ESSA: Evolutionary Strategies for Scalable Alignment

    Jul 6, 2025Daria Korotyshova, Boris Shaposhnikov, Alexey Malakhov +7Large Language Model AlignmentLarge Models