Selective Knowledge Distillation

Latest papers 19

All topics
CardsList
  1. Learning What to Distill: Bilevel Top-K Token Selection for Self-Distillation in Large Language Models

    Oct 5, 2026Heng Liang, Xinwen Zhang, Hongchang GaoSelective Knowledge DistillationLanguage Model Distillation

  2. Transfer-Stratified On-Policy Distillation for RL-Improved Reasoning Teachers

    Oct 5, 2026Xiaoyu Chen, Bo Shao, Tiangang Zhu +5RL for Language Model ReasoningSelective Knowledge Distillation

  3. Know Thyself, Teach Thyself: Internal Information Flow for Selective Self-Distillation

    Sep 29, 2026Rui Wang, Ruijie Wang, Bo Chen +2On-Policy Self-DistillationSelective Knowledge Distillation

  4. EOPSA: Efficient On-Policy Self-Distilled Safety Alignment

    Sep 28, 2026Qirui Liu, Yichen Sun, Yan Wang +5On-Policy Self-DistillationSelective Knowledge Distillation

  5. Not Every Token Is Worth Distilling: Selective Supervision for Direct-OPD

    Sep 24, 2026Yibo Zhao, Zixuan Yang, Yunshi Lan +1Selective Knowledge DistillationOn-Policy Distillation

  6. BioKD: Selective Physiology-to-Video Knowledge Distillation via Reliability Gate for Emotion Recognition

    Aug 6, 2026Bojing Hou, Ruohao Li, Yitong Zhu +3Cross-Modal Knowledge DistillationSelective Knowledge Distillation

  7. Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance

    Aug 1, 2026Zhuowen Han, Jinwei Xiao, Zhengxi Lu +9Self-Training for Language ModelsRL for Language Model Reasoning

  8. When Does Knowledge Distillation Hurt? Reliability-Aware Distillation for Low-Resource Language Summarization

    Jul 22, 2026Dipto Sumit, Ankan Kumar Roy Srizon, Sadia Khair Rodela +4Selective Knowledge DistillationText Summarization

  9. UNIEGO: Proxies as Mediators for Unified Egocentric Video Representation Learning

    Jun 18, 2026Wenhao Chi, Arkaprava Sinha, Dominick Reilly +2Cross-Modal Knowledge DistillationMulti-View Learning

  10. FADA: Accessible fetal ultrasound interpretation and annotation with a selectively distilled unified vision-language model

    Jun 9, 2026Mahmood Alzubaidi, Uzair Shah, Raden Muaz +6Ultrasound Image SegmentationSelective Knowledge Distillation

  11. Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation

    May 26, 2026Yuanyi Wang, Su Lu, Yanggan Gu +6Selective Knowledge DistillationLanguage Model Distillation

  12. Not All Timesteps Matter Equally: Selective Alignment Knowledge Distillation for Spiking Neural Networks

    May 14, 2026Kai Sun, Peibo Duan, Yongsheng Huang +4Selective Knowledge DistillationSpiking Neural Networks

  13. Prefix Teach, Suffix Fade: Local Teachability Collapse in Strong-to-Weak On-Policy Distillation

    May 13, 2026Kaiyuan Liu, Ziyuan Zhuang, Yang Bai +3Selective Knowledge DistillationLanguage Model Distillation

  14. Distilling the Essence: Efficient Reasoning Distillation via Sequence Truncation

    Dec 24, 2025Wei-Rui Chen, Vignesh Kothapalli, Ata Fatahibaarzi +5CoT DistillationSelective Knowledge Distillation