Bottlenecks

Momentum

25 papers in the last four weeks, up 213% on the four weeks before. 0.2% of all new papers.

Jul 13Week of Sep 28

Latest papers 200

All topics
CardsList
  1. AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization

    Apr 20, 2026Kosuke Matsushima, Yasuyuki Okoshi, Masato Motomura +1LLM Inference OptimizationBottlenecks

  2. Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?

    Apr 20, 2026H S V N S Kowndinya Renduchintala, Sumit BhatiaLarge Language Models FailLinguistics

  3. Capacity-Controlled Global Attention for Graph Transformers

    Apr 19, 2026Yang Liu, Dongxin Guo, Tom Zheng +3Transformer AttentionLinear Attention

  4. KAIROS: Stateful, Context-Aware Power-Efficient Agentic Inference Serving

    Apr 17, 2026Yichao Yuan, Mosharaf Chowdhury, Nishil TalatiAgentic InferencePeak Memory

  5. Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows

    Apr 17, 2026Luay Gharzeddine, Samer SaabMulti-Agent WorkflowsAgent Loop

  6. Placing Puzzle Pieces Where They Matter: A Question Augmentation Framework for Reinforcement Learning

    Apr 17, 2026Yangyi Fang, Jiaye Lin, Xiaoliang Fu +2PuzzleQuestion

  7. Concept-wise Attention for Fine-grained Concept Bottleneck Models

    Apr 17, 2026Minghong Zhong, Guoshuai Zou, Kanghao Chen +2Fine-Grained PerceptionConcept Bottleneck Models

  8. Exploiting Correlations in Federated Learning: Opportunities and Practical Limitations

    Apr 16, 2026Adrian Edin, Michel Kieffer, Mikael Johansson +1Federated LearningGenerative Image Compression

  9. MAPLE: Metadata Augmented Private Language Evolution

    Feb 26, 2026Eli Chien, Yuzheng Hu, Ryan McKenna +3Synthetic DataAlphaevolve

  10. TADA! Tuning Audio Diffusion Models through Activation Steering

    Feb 12, 2026Łukasz Staniszewski, Katarzyna Zaleska, Mateusz Modrzejewski +1Audio EditingLinear Activation Steering

  11. CauScale: Neural Causal Discovery at Scale

    Feb 9, 2026Bo Peng, Sirui Chen, Jiaguo Tian +2Causal Discovery MethodsCausal

  12. Near-Oracle KV Selection via Pre-hoc Sparsity for Long-Context Inference

    Feb 9, 2026Yifei Gao, Lei Wang, Rong-Cheng Tu +3Dynamic Sparse AttentionLLM Inference Optimization

  13. A Tool Bottleneck Framework for Clinically-Informed and Interpretable Medical Image Understanding

    Dec 24, 2025Christina Liu, Alan Q. Wang, Joy Hsu +2Medical ImagesMultimodal Clinical Data

  14. Block Sparse Flash Attention

    Dec 7, 2025Daniel Ohayon, Itay Lamprecht, Itay Hubara +3Block Sparse Flash AttentionEfficient Long-Context Inference

  15. VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR

    May 23, 2025Shenghui Chen, Po-han Li, Sandeep Chinchali +1SummarizationBottlenecks