Neural Network Pruning

Momentum

22 papers in the last four weeks, up 83% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 161

All topics
CardsList
  1. HiAP: A Multi-Granular Stochastic Auto-Pruning Framework for Vision Transformers

    Mar 12, 2026Andy Li, Aiden Durrant, Milan Markovic +1Efficient ViTsAttention Head Pruning

  2. Universal Redundancies in Time Series Foundation Models

    Feb 2, 2026Anthony Bao, Venkata Hasith Vattikuti, Jeffrey Lai +1Transformer InterpretabilityKernel Regression

  3. Garbage Attention in Large Language Models: BOS Sink Heads and Sink-aware Pruning

    Jan 11, 2026Jaewon Sok, Jewon Yeom, Seonghyeon Park +2LLM PruningAttention Head Analysis

  4. Data-Free Pruning of Self-Attention Layers in LLMs

    Dec 3, 2025Dhananjay Saikumar, Blesson VargheseLLM PruningLLM Compression

  5. Pruning as Regularization: Sensitivity-Aware One-Shot Pruning in ASR

    Nov 11, 2025Julian Irigoyen, Arthur Söhler, Andreas Søeborg KirkedalAutomatic Speech RecognitionImplicit Regularization

  6. PATCH: Learnable Tile-level Hybrid Sparsity for LLMs

    Sep 27, 2025Younes Hourri, Mohammad Mozaffari, Maryam Mehri DehnaviLLM PruningLLM Compression

  7. One Shot vs. Iterative: Rethinking Pruning Strategies for Model Compression

    Aug 19, 2025Mikołaj Janusz, Tomasz Wojnar, Yawei Li +2Model CompressionStructured Pruning

  8. Hyperflux: Pruning Reveals Importance

    Apr 6, 2025Eugen Barbulescu, Antonio Alexoaie, Lucian BusoniuSparse Neural NetworksNeural Network Training Dynamics

  9. Explainable Bayesian deep learning through input-skip Latent Binary Bayesian Neural Networks

    Mar 13, 2025Eirik Høyheim, Lars Skaaret-Lund, Solve Sæbø +1Bayesian Neural NetworksUncertainty Quantification

  10. LLM Compression by Block Removal with Constrained Binary Optimization

    Date pendingDavid Jansen, Roman Rausch, Ali Hashemi +2LLM PruningLLM Compression