Neural Network Pruning

Momentum

22 papers in the last four weeks, up 83% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 161

All topics
CardsList
  1. When Can You Prune Your Network? A Study of Intermediate Neurons in Multilingual Speech Parsing

    Oct 8, 2026Minnie Kabra, Benjamin Lecouteux, Maximin CoavouxNeural Network Pruning

  2. CHASE: Channel-Aligned Structure Exploitation for Geometry-Aware Model Engineering

    Oct 7, 2026Wei Wang, Wei Jiang, Ziran LiuNeural Network CompressionKV-Cache Compression

  3. Towards Efficient Robotic Manipulation Models with Self-Recursive Pruning

    Oct 6, 2026Zijia Chen, Yuenan Hou, Yu Li +2Robot Policy LearningRobotic Manipulation

  4. Task-Aware Joint Pruning and Distillation for Efficient Audio Deepfake Detection

    Oct 4, 2026Miao He, Peng Cheng, Zhongjie Ba +4Audio Deepfake DetectionKnowledge Distillation

  5. MWOP: Modality-aware Width-wise Operation Pruning for Efficient MLLMs

    Oct 1, 2026Xudong Wang, Hao Wu, Haozhe Hu +5LLM PruningEfficient VLM Inference

  6. DIET: Deletion-response Expert Trimming for Video Diffusion Transformers

    Sep 29, 2026Jiachang Zhang, Teng Hu, Bohao Feng +4Mixture-of-Experts PruningVideo Diffusion Models

  7. Does the VGGT Family Need All Its Layers?

    Sep 29, 2026Fengyi Zhang, Holger Caesar, Xiangyu Sun +3Multi-View 3D ReconstructionFeed-Forward 3D Reconstruction

  8. Output-aware Residual Stream Pruning for Large Language Models

    Sep 28, 2026Chayne Thrash, Kevin Chen, Soheil KolouriLLM PruningLLM Compression

  9. GroupMask: Layer-Adaptive Group-wise Sparsity for Semi-Structured LLM Pruning

    Sep 27, 2026Zhengao Li, Shuoqiu Li, Xiaofang Zhang +7LLM PruningLLM Compression

  10. Fisher Simplicity in Kolmogorov-Arnold Networks and Multilayer Perceptrons

    Sep 26, 2026Ami Tavory, Meir FederNeural Network InterpretabilityKolmogorov-Arnold Networks

  11. AERIAL: Adversarial Evaluation of Robustness in Accuracy-Preserving Low-Precision EEG Decoders

    Sep 24, 2026Saim Rehman, Muhammad ShafiqueQuantization-Aware TrainingEEG Decoding

  12. Six Layers Less: Encoder Pruning for Whisper with Label-Free Recovery

    Sep 23, 2026Rasmus Aagaard, Nicki Skafte DetlefsenSpeech Foundation ModelsAutomatic Speech Recognition

  13. DTKDP: A Dual Teacher Knowledge Distillation and Pruning Framework for Lightweight Oriented SAR Ship Detection

    Sep 21, 2026Yuming Li, Fan Zhang, Alin M. AchimMulti-Teacher Knowledge DistillationObject Detection

  14. Artificial Structure Function Search: Preserving Artificial Functional Connectivity for Structured Pruning

    Sep 21, 2026Mindula Illeperuma, Rafael Pina, Charuka Herath +2Model CompressionFunctional Connectivity

  15. Prescriptive SVD-Inspired Attention via Spectral Energy Retention

    Sep 21, 2026Vasileios Arampatzakis, Vasileios Sevetlidis, George PavlidisTransformer InterpretabilityLow-Rank Attention

  16. Higher-order pruning of experts in mixture-of-experts language models

    Sep 16, 2026Alex M. Tseng, Prannay Kaul, Luca Zancato +2Mixture-of-Experts PruningMixture-of-Experts Language Models

  17. Theoretical Guarantees for One-Shot Magnitude Pruning and Compute-Adaptive Early Exit

    Sep 14, 2026Erdem KoyuncuNeural Network GeneralizationEarly-Exit Neural Networks

  18. X-RACE: XAI-assisted Recurrent neural network Attribution for Channel Estimation

    Sep 12, 2026Abdul Karim Gizzini, Yahia MedjahdiEfficient Neural Network InferenceExplainable Artificial Intelligence

  19. LILA: Calibration-Free Structured Pruning of Large Language Models via Latent Spectral Geometry

    Sep 11, 2026Sankar Behera, Dhruv Singh, Anshika Agnihotri +3LLM PruningLLM Compression

  20. One Loop, Two Gains: Can Active Learning win the Lottery for Free?

    Sep 9, 2026Benedikt Tscheschner, Eduardo Veas, Marc MasanaActive LearningNeural Network Pruning

  21. LinearMask-GS: Stable-Mask Importance Pruning for Compact 3D Gaussian Splatting

    Sep 9, 2026Donghun Ryu, Minhyeok Lee3DGS Compression3D Gaussian Splatting

  22. Forward-Free LLM Depth Pruning via Weight Redundancy

    Sep 9, 2026Vincent-Daniel Yun, Woosang LimLLM PruningLLM Inference Acceleration

  23. Understanding the Impact of Model Pruning on Long-Tail Forgetting and Explanation Reliability in Medical Imaging

    Sep 7, 2026Nazish Khalid, Tausifa Jan Saleem, Amal Saqib +2Model CompressionExplainable Medical Image Analysis

  24. TAP-Path: Task-Adaptive Structural and Token Pruning for Efficient and Trustworthy Pathology Foundation Models

    Sep 3, 2026Mehedi Hasan, Ashfak Yeafi, Md Khairul IslamComputational PathologyPathology Foundation Models

  25. Measurement-Driven Sub-Network Selection for On-Premise Retrieval-Augmented Factory Agents

    Sep 2, 2026Vasileios Rizeakos, Georgios Paisios, Alexandros Machairas +2Hardware-Aware NASEdge Computing

  26. Debias-SparseGPT: Bias-Aware Pruning for Large Language Models

    Sep 2, 2026Irina Proskurina, Guillaume Metzler, Antoine Gourru +1LLM PruningNeural Network Pruning