On-Policy Distillation

Recent momentum

-66%

21 papers in the last 28 days · 0.5% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-14

13 new papers

A weekly snapshot of new work published in On-Policy Distillation.

226 papers

Latest in On-Policy Distillation

Open your feed →
CardsList
  1. Scaling Self-Play for End-to-End Driving

    Jun 17, 2026Luke Rowe, Roger Girgis, Rodrigue de Schaetzen +6Autonomous Driving SimulationAutonomous Driving

  2. Covert Influence Between Language Models

    Jun 2, 2026Avidan Shah, Jay Chooi, Jinghua Ou +1InfluenceProtein Language Model

  3. Trust Region On-Policy Distillation

    May 31, 2026Xingrun Xing, Haoqing Wang, Boyan Gao +2On-Policy DistillationToken-Level Supervision

  4. Trust-Region Behavior Blending for On-Policy Distillation

    May 29, 2026Daniil Plyusov, Alexey Gorbatovski, Alexey Malakhov +4On-Policy DistillationTrust Region