Teacher-Student Learning

Momentum

7 papers in the last four weeks, against 2 the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 162

All topics
CardsList
  1. Localizing Credit at the Divergence: Path-Conditioned Self-Distillation for LLM Reasoning

    Jun 14, 2026Yu Li, Shu Hong, Tian LanRL for Language Model ReasoningCredit Assignment in RL

  2. Teacher-Student Structure for Domain Adaptation in Ensemble Audio-Visual Video Deepfake Detection

    Jun 13, 2026Elham Abolhasani, Maryam Ramezani, Hamid R. RabieeAudio Deepfake DetectionDomain Adaptation

  3. Robust Fall Recovery for Armless Bipedal-Wheeled Robots Via Force-Guided Learning

    Jun 12, 2026Haidong Hou, Zhangguo Yu, Tao Han +6Robot Failure RecoveryRobot Locomotion

  4. Dense Supervision, Sparse Updates: On the Sparsity and Geometry of On-Policy Distillation

    Jun 11, 2026Guo Yu, Wenlin Liu, Yulan Hu +3Sparse Fine-TuningOn-Policy Distillation

  5. A solvable model for unsupervised federated learning

    Jun 11, 2026Giovanni Catania, Aurélien Decelle, Gianluca Manzan +2Statistical Physics of LearningUnsupervised Learning

  6. Beyond Dark Knowledge: Mixup-Based Distillation for Reliable Predictions

    Jun 10, 2026José Medina, Paul Honeine, Abdelaziz Bensrhair +1Data AugmentationModel Calibration

  7. Hey Chat, Can You Teach Me? Structuring Socratic Dialogue for Human Learning in the Wild

    Jun 10, 2026Sidney Tio, Arunesh Sinha, Pradeep VarakanthamIntelligent Tutoring SystemsAI in Education

  8. RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation

    Jun 10, 2026Leyi Pan, Shuchang Tao, Yunpeng Zhai +5RL for Language Model ReasoningLanguage Model Distillation

  9. Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation

    Jun 9, 2026Uwe König, Hamza Kazmi, Ruizhe Li +1Subliminal LearningLanguage Model Distillation

  10. PADD: Path-Aligned Decompression Distillation for Non-Router Teacher to Guide MoE Student Learning

    Jun 9, 2026Xinyue Peng, Yi Qian, Jiaojiao Lin +2Mixture-of-Experts Language ModelsExpert Routing

  11. Escaping the KL Agreement Trap in On-Policy Distillation

    Jun 8, 2026Haoran Xin, Anhao Zhao, Ying Sun +3On-Policy DistillationTeacher-Student Learning

  12. SG-OPD: Sign-Gated On-Policy Distillation via Sign-Consistency Gating and Phased Teacher Sampling

    Jun 8, 2026Haoran Xu, Hongyu Wang, Yifei Gao +3Language Model DistillationOn-Policy Distillation

  13. Trajectory-Refined Distillation

    Jun 7, 2026Li Jiang, Haoran Xu, Yichuan Ding +1On-Policy DistillationLLM Post-Training

  14. OPRD: On-Policy Representation Distillation

    Jun 4, 2026Shenzhi Yang, Guangcheng Zhu, Bowen Song +8Cross-Architecture Knowledge DistillationLanguage Model Distillation

  15. Compress-Distill: Reasoning Trace Compression for Efficient Knowledge Distillation

    Jun 4, 2026Maxime Griot, Paul Steven Scotti, Tanishq Mathew AbrahamCoT DistillationTeacher-Student Learning

  16. ViCuR: Visual Cues as Recoverable Privilege for Multimodal On-Policy Distillation

    Jun 4, 2026Kanghui Tian, Siyuan Liu, Ziang Yan +3VLM DistillationMultimodal Reasoning

  17. SocraticPO: Policy Optimization via Interactive Guidance

    Jun 3, 2026Zirui Liu, Jie Ouyang, Qi Liu +8RL for Language Model ReasoningLLM-Guided RL

  18. When Should the Teacher Move? Temporal Coupling and Stability in Self On-Policy Distillation

    Jun 2, 2026Haowei Guo, Baolong Bi, Ruicheng Zhang +2On-Policy Self-DistillationLanguage Model Distillation

  19. What Makes Interaction Trajectories Effective for Training Terminal Agents?

    Jun 2, 2026Sidi Yang, Chaofan Tao, Jierun Chen +11AI Agent BenchmarksLLM Agent Training

  20. What Do Students Learn? A Feature-Level Analysis of Dark Knowledge

    Jun 2, 2026Seungu Kang, Songkuk KimRepresentation LearningSelf-Distillation

  21. Why Are DMD Students Lazy? Understanding the Copying Behavior in Few-Step Distillation

    Jun 1, 2026Shucheng Li, Iolo Jones, Alexander Tong +1Diffusion Model DistillationFew-Step Diffusion Sampling

  22. What Makes a Strong Model? A Unified Spectral Analysis of Knowledge Transfer over High-dimensional Linear Regression

    May 31, 2026Wendao Wu, Fangqing Zhang, Haihan Zhang +1Weak-to-Strong GeneralizationTeacher-Student Learning

  23. OPD+: Rethinking the Advantage Design for On-Policy Distillation

    May 31, 2026Hanyang Zhao, Haoxian Chen, Han Lin +3Language Model DistillationPolicy Gradient Methods

  24. Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher

    May 31, 2026Arda Uzunoglu, Alvin Zhang, Daniel KhashabiData SelectionTraining Data Selection

  25. Subliminal Learning Is Steering Vector Distillation

    May 31, 2026Camila Blank, Agam Bhatia, Senthooran Rajamanoharan +2Language Model SteeringSubliminal Learning

  26. DASH: Dual-Branch Score Distillation for Guidance-Calibrated Compact Diffusion Models

    May 30, 2026Abdullah Al Shafi, Kazi Saeed Alam, Sk Imran Hossain +1Model CompressionClassifier-Free Guidance

  27. Logit Distillation on Manifolds: Mapping by Learning

    May 30, 2026Yiru Yang, Junling Wang, Nishant Kumar Singh +2Teacher-Student LearningRepresentation Alignment