Token-Level Entropy

Latest papers 43

All topics
CardsList
  1. No Model Required: Text Entropy Rate Filtering Mitigates Iterative Fine-Tuning Collapse

    Oct 1, 2026Lewis MitchellCollapseToken-Level Entropy

  2. Scaling Attention Head Analysis via Gradient-Based Attribution in Context-Aware Machine Translation

    Sep 23, 2026Paweł Mąka, Yusuf Can Semerci, Jan Scholtes +1Token-Level EntropyDisambiguation

  3. GMTS: Gradient Magnitude-based Token Selection Improves RLVR Training for LLM Reasoning

    Aug 31, 2026Outongyi Lv, Yuanwei Zhang, Xiaoqun ZhangToken-Level EntropyReinforcement Learning With Verifiable Reward

  4. Tokenizer-Generator Coupling in Medical Image Generation

    Aug 7, 2026Liam ChalcroftMedical Image GenerationVisual Tokenizers

  5. Predicting Task Difficulty Without Rollouts

    Aug 6, 2026Stefan Krsteski, Charlotte MeyerAgentic BenchmarksEvaluation Metrics

  6. Demystifying Entropy-based Selection for Chain-of-Thought Compression in Large Reasoning Models

    Jul 30, 2026Sara Candussio, Daniel Scalena, Luca Bortolussi +3Token-Level EntropyLarge Reasoning Models

  7. Group Entropy-Controlled Policy Optimization

    Jul 18, 2026Guangran Cheng, Chengqi Lyu, Songyang Gao +2Token-Level Entropy

  8. Which Tokens Matter? Adaptive Token Selection for RLVR with the Relative Surprisal Index

    Jun 30, 2026Outongyi Lv, Yanzhao Zheng, Yuanwei Zhang +5Token-Level EntropyReinforcement Learning With Verifiable Reward

  9. Defending Against Harmful Supervision Hidden in Benign Samples

    Jun 29, 2026Bang An, Yibo Yang, Dandan Guo +3Supervised FinetuningModel-Agnostic Defense

  10. Smooth Scaling Laws Hide Stepwise Token Learning

    Jun 29, 2026Pingjie Wang, Zechen Hu, Peiru Yang +2Scaling LawsLarge Language Model Training

  11. What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics

    Jun 23, 2026Sofiia Nikolenko, Michele Papucci, Mina Rezaei +1Large Language Model JailbreaksToken-Level Entropy

  12. On the Position Bias of On-Policy Distillation

    Jun 21, 2026Yan Xie, Sijie Zhu, Tiansheng Wen +2Efficient On-Policy DistillationToken-Level Entropy

  13. Beyond Entropy: Learning from Token-Level Distributional Deviations for LLM Reasoning

    Jun 18, 2026Xuanzhi Feng, Zhengyang Li, Zeyu Liu +6Token-Level EntropyLLM Reasoning Strategies

  14. STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability

    Jun 17, 2026Haipeng Luo, Qingfeng Sun, Songli Wu +4Token-Level EntropyOffline Reinforcement Learning

  15. Routing-Aware Expert Calibration for Machine Unlearning in Mixture-of-Experts Language Models

    Jun 9, 2026Jingyi Xie, Yijun Lin, Yinjiang Xiong +2Mixture-Of-Experts Large Language ModelsLarge Language Model Unlearning

  16. Sample Where You Struggle: Sharpening Base Model Reasoning via Entropy-Guided Power Sampling

    Jun 7, 2026Hong Guo, Nianhui Guo, Christoph Meinel +1Token-Level EntropyMarkov Chain Monte Carlo

  17. PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

    Jun 7, 2026Shumeng Yang, Yisu Liu, Jiayi Zheng +2Reinforcement Learning With Verifiable RewardToken-Level Entropy

  18. Trajectory-Refined Distillation

    Jun 7, 2026Li Jiang, Haoran Xu, Yichuan Ding +1Symbolic DistillationTeacher

  19. Entropy Gate: Entropy Quenching for Near-Lossless Token Compression in LLM Pipelines

    Jun 2, 2026Justice Owusu Agyemang, Jerry John Kponyo, Kwame Opuni-Boachie Obour Agyekum +3Token CompressionToken-Level Entropy

  20. Entropy-aware Masking for Masked Language Modeling

    May 27, 2026Gokul Srinivasagan, Kai Hartung, Munir GeorgesMasked Diffusion Language ModelsToken-Level Entropy

  21. Entropy Distribution as a Fingerprint for Hallucinations in Generative Models

    May 27, 2026Mattia J. Villani, Pranav Deshpande, Akshay Seshadri +2Large Language Model HallucinationHallucination Detection

  22. Cyclical Entropy Eruption: Entropy Dynamics in Agent Reinforcement Learning

    May 27, 2026Wendi Li, Shawn Im, Sharon LiAgentic Reinforcement LearningOffline Reinforcement Learning

  23. Detecting Fluent Optimization-Based Adversarial Prompts via Sequential Entropy Changes

    May 19, 2026Mohammed Alshaalan, Miguel R. D. RodriguesAdversarial PromptsToken-Level Entropy

  24. Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

    May 19, 2026Shuyu Wei, Jian Sun, Delai Qiu +6LLM Reasoning StrategiesToken-Level Entropy

  25. When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy

    May 12, 2026Xiaofeng Tan, Jun Liu, Bin-Bin Gao +5DiversityToken-Level Entropy

  26. Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control

    May 12, 2026Jiazheng Zhang, Ziche Fu, Junrui Shen +17Frictive Policy OptimizationToken-Level Entropy

  27. Taming Extreme Tokens: Covariance-Aware GRPO with Gaussian-Kernel Advantage Reweighting

    May 12, 2026Cheng Wang, Qin Liu, Wenxuan Zhou +1Repair-Based Group-Relative Policy OptimizationToken-Level Entropy

  28. Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

    May 12, 2026Huimin Xu, Shuai Zhao, Xiaobao Wu +1Reinforcement Learning With Verifiable RewardToken-Level Entropy

  29. Entropy-informed Decoding: Adaptive Information-Driven Branching

    May 10, 2026Benjamin Patrick Evans, Sumitra Ganesh, Leo ArdonConstrained DecodingToken-Level Entropy

  30. Scaling Categorical Flow Maps

    May 8, 2026Oscar Davis, Anastasiia Filippova, Pierre Ablin +4Latent FlowFlow Map

  31. Not All Tokens Learn Alike: Attention Entropy Reveals Heterogeneous Signals in RL Reasoning

    May 8, 2026Gengyang Li, Zheng-Fan Wu, Siqi Bao +1Token-Level EntropyPost-Training

  32. Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models

    May 5, 2026Wei-Ping Huang, Chee-En Yu, Guan-Ting Lin +1Stable Test-Time AdaptationAutoregressive Model

  33. Entropy Centroids as Intrinsic Rewards for Test-Time Scaling

    Apr 28, 2026Wenshuo Zhao, Qi Zhu, Xingshan Zeng +4Token-Level EntropyTest-Time Scaling

  34. Representational Curvature Modulates Behavioral Uncertainty in Large Language Models

    Apr 27, 2026Jack King, Evelina Fedorenko, Eghbal A. HosseiniToken-Level EntropyLarge Language Model Uncertainty

  35. Understanding Secret Leakage Risks in Code LLMs: A Tokenization Perspective

    Apr 20, 2026Meifang Chen, Zhe Yang, Huang Nianchen +4Data LeakageToken-Level Entropy

  36. Forget What Matters, Keep the Rest: Selective Unlearning of Informative Tokens

    Apr 20, 2026Seunghee Koh, Sunghyun Baek, Youngdong Kim +1Large Language Model UnlearningToken-Level Entropy

  37. Dissecting Failure Dynamics in Large Language Model Reasoning

    Apr 16, 2026Wei Zhu, Jian Zhang, Lixing Yu +2LLM Reasoning StrategiesToken-Level Entropy

  38. Entropy-Aware Token Rejection for Improving Speculative Decoding

    Dec 29, 2025Tiancheng Su, Meicong Zhang, Guoxiu HeSpeculative DecodingToken-Level Entropy