Structured Pruning

Momentum

3 papers in the last four weeks, down 25% on the four weeks before. 0.0% of all new papers.

Jul 13Week of Sep 28

Latest papers 37

All topics
CardsList
  1. LILA: Calibration-Free Structured Pruning of Large Language Models via Latent Spectral Geometry

    Sep 11, 2026Sankar Behera, Dhruv Singh, Anshika Agnihotri +3LLM PruningLLM Compression

  2. Dense Structural Compression of Transformers via Gauge-Correct Channel Removal

    Sep 7, 2026Jed A. Duersch, Naïm Es-Sebbani, Nathanaël Haas +1Efficient Transformer InferenceStructured Pruning

  3. Functional Degeneracy in Neural Networks: Measurement and Pruning

    Aug 31, 2026Maria Matveev, Pascal Esser, Ayush Bharadwaj +2Structured PruningNeural Network Compression

  4. PruneShift: A Framework for Evaluating Decision Reliability in Structured Pruning

    Aug 30, 2026Hao Ye, Gaopeng ZhangSurrogate ModelingStructured Pruning

  5. Unifying Depth and Width Pruning for LLMs via Binary Knapsack Optimization

    Aug 13, 2026Palaash Goel, Ayan Sengupta, Akshay Nambi +1LLM PruningStructured Pruning

  6. InterPruner: Interactive Structured Pruning via Taylor-Implicit Criterion and Language-Prior Modulator for Multimodal Object Detection

    Aug 11, 2026Qi Ming, Zihan Yang, Shaoguang Huang +6Multispectral Object DetectionInfrared Object Detection

  7. StaticSegFormer: An Efficient High-Performance Semantic Segmentation Based on Static Structured Pruning

    Aug 5, 2026Timo Bartels, Danish Nazir, Jan Piewek +2Efficient Neural Network InferenceStructured Pruning

  8. CoCurve: Cross-Module Co-Pruning Curvature for Training-Free Structured LLM Pruning

    Jul 20, 2026Zhiren Gong, Zihao Zeng, Zijie Wang +3LLM PruningLLM Compression

  9. Structured Pruning of Large Language Models via Power Transformation and Sign-Preserving Score Aggregation with Adaptive Feature Retention

    Jul 9, 2026Ryota Kobayashi, Tsubasa Hirakawa, Takayoshi Yamashita +4LLM PruningLLM Compression

  10. FlexMoE: One-for-All Nested Intra-Expert Pruning for MoE Language Models

    Jun 26, 2026Fan Mo, Yuxuan Han, Geng Zhang +2Mixture-of-Experts PruningLLM Pruning

  11. Drop-Then-Recovery: How Redundant Are Vision-Language-Action Models?

    Jun 26, 2026Guoheng Sun, Kaixi Feng, Shwai He +8Model CompressionEfficient VLA Models

  12. Cascaded Multi-Granularity Pruning for On-Device LLM Inference in Industrial IoT

    Jun 25, 2026Jinghan Wang, Yanjun Chen, Wei Zhang +3LLM PruningOn-Device Language Model Inference

  13. Attribution-Guided and Coverage-Maximized Pruning for Structural MoE Compression

    Jun 16, 2026Yifu Ding, Jiacheng Wang, Ge Yang +4Mixture-of-Experts PruningLLM Pruning

  14. Squeeze-Release: Iterative Pruning with Exact Structural Minimization

    Jun 12, 2026Roman Denkin, Ida Akerholm, Prashant Singh +1Structured PruningNeural Network Compression

  15. Small LLMs: Pruning vs. Training from Scratch

    Jun 12, 2026Yufeng Xu, Taiming Lu, Kunjun Li +3Language Model PretrainingLLM Pruning

  16. Beyond FLOPs: Benchmarking Real Inference Acceleration of LLM Pruning under a GEMM-Centric Taxonomy

    Jun 8, 2026Haozhe Hu, Hao Wu, Anhao Zhao +4LLM PruningLLM Inference Acceleration

  17. Less is MoE: Trimming Experts in Domain-Specialist Language Models

    Jun 4, 2026Haoze He, Xinkai Zou, Xuan Jiang +4Mixture-of-Experts PruningLLM Pruning

  18. TENP: Trapezoidal Expert Neuron Pruning For Mixture-of-Experts

    Jun 3, 2026Jiangyang He, Shaolin Zhu, Deyi XiongMixture-of-Experts PruningLLM Pruning

  19. PSViT: A Methodology for Structurally Pruning Spiking Vision Transformers

    Jun 2, 2026Rachmad Vidya Wicaksana Putra, Achyuta Muthuvelan, Alberto Marchisio +1Efficient ViTsSpiking Neural Networks

  20. PrunePath: Towards Highly Structured Sparse Language Models

    May 27, 2026Zhexuan Gu, Zixun Fu, Yancheng YuanMixture-of-Experts PruningLLM Pruning

  21. MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning

    May 25, 2026Aritra Dutta, Somak AdityaVision-Language ModelsEfficient VLM Inference

  22. Pruning Deep Neural Networks via the Marchenko--Pastur Distribution

    May 23, 2026Leonid Berlyand, Theo Bourdais, Houman Owhadi +1Structured PruningSparse Neural Networks

  23. Prune, Update and Trim: Robust Structured Pruning for Large Language Models

    May 18, 2026Diego Coello de Portugal Mecke, Tom Hanika, Lars Schmidt-ThiemeLLM PruningLLM Inference Acceleration

  24. MedCore: Boundary-Preserving Medical Core Pruning for MedSAM

    May 13, 2026Cenwei Zhang, Suncheng Xiang, Lei YouModel CompressionImage Segmentation

  25. Relative Kinetic Utility: Calibrating Cross-Layer Credit for Global Structured LLM Pruning

    May 9, 2026Tianhao Qian, Guilin Qi, Jiayu ChenLLM PruningStructured Sparsity

  26. Compact SO(3) Equivariant Atomistic Foundation Models via Structural Pruning

    May 9, 2026Chen Wang, Siyu Hu, Guangming Tan +1Equivariant GNNsStructured Pruning

  27. Task Relevance Is Not Local Replaceability: A Two-Axis View of Channel Information

    May 8, 2026Houman Safaai, Andrew T. Landau, Celia C. Beron +2Structured PruningNeural Network Pruning

  28. SlimDiffSR: Toward Lightweight and Efficient Remote Sensing Image Super-Resolution via Diffusion Model Distillation

    May 4, 2026Ce Wang, Zhenyu Hu, Wanjie SunDiffusion Model DistillationRemote Sensing Image Super-Resolution

  29. GETA-3DGS: Automatic Joint Structured Pruning and Quantization for 3D Gaussian Splatting

    May 3, 2026Baobing Zhang, Wanxin Sui3DGS CompressionMixed-Precision Quantization

  30. Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

    Apr 27, 2026Yiran Huang, Lukas Thede, Massimiliano Mancini +2Large Vision-Language ModelsStructured Pruning

  31. GRASPrune: Global Gating for Budgeted Structured Pruning of Large Language Models

    Apr 21, 2026Ziyang Wang, Jiangfeng Xiao, Chuan Xiao +3LLM PruningLLM Compression

  32. One Shot vs. Iterative: Rethinking Pruning Strategies for Model Compression

    Aug 19, 2025Mikołaj Janusz, Tomasz Wojnar, Yawei Li +2Model CompressionStructured Pruning

  33. DarwinLM: Evolutionary Structured Pruning of Large Language Models

    Feb 11, 2025Shengkun Tang, Oliver Sieberling, Eldar Kurtic +2LLM PruningStructured Pruning