On-Policy Distillation

Recent momentum

-52%

33 papers in the last 28 days · 0.5% of indexed attention

Twelve weeks of publication activity for this topic as it is defined today.

Weekly history

Recent digests

What was published in this topic, kept on the site without email delivery.

Period ending 2026-09-21

9 new papers

A weekly snapshot of new work published in On-Policy Distillation.

Period ending 2026-09-14

13 new papers

A weekly snapshot of new work published in On-Policy Distillation.

291 papers

Latest in On-Policy Distillation

  1. Trajectory-Refined Distillation

    Jun 7, 2026Li Jiang, Haoran Xu, Yichuan Ding +1On-Policy DistillationTool Failures

  2. OPRD: On-Policy Representation Distillation

    Jun 4, 2026Shenzhi Yang, Guangcheng Zhu, Bowen Song +8On-Policy DistillationHidden States

  3. Covert Influence Between Language Models

    Jun 2, 2026Avidan Shah, Jay Chooi, Jinghua Ou +1InfluenceTraining Language Models

  4. Trust Region On-Policy Distillation

    May 31, 2026Xingrun Xing, Haoqing Wang, Boyan Gao +2On-Policy DistillationToken-Level Supervision

  5. Trust-Region Behavior Blending for On-Policy Distillation

    May 29, 2026Daniil Plyusov, Alexey Gorbatovski, Alexey Malakhov +4On-Policy DistillationTrust Region