Attention Mechanisms

Momentum

42 papers in the last four weeks, up 100% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 378

All topics
CardsList
  1. Look Beyond Saliency: Low-Attention Guided Dual Encoding for Video Semantic Search

    May 7, 2026Faisal Aljehrai, Mohammed A. Alkhrashi, Alreem Almuhrij +6Visual Representation LearningInformation Retrieval

  2. Neuromorphic visual attention for Sign-language recognition on SpiNNaker

    May 7, 2026Sarka Liskova, Olha Vedmedenko, Mazdak Fatahi +3Event-Based VisionVisual Attention

  3. Retrieval from Within: An Intrinsic Capability of Attention-Based Models

    May 7, 2026Elad Hoffer, Yochai Blau, Edan Kinderman +3Retrieval-Augmented GenerationQuestion Answering

  4. Large Vision-Language Models Get Lost in Attention

    May 7, 2026Gongli Xi, Ye Tian, Mengyu Yang +5Transformer FFNsVision-Language Models

  5. Average Attention Transformers and Arithmetic Circuits

    May 6, 2026Lena Ehrmuth, Laura StriekerTransformer ExpressivityTransformer

  6. Angle-I2P: Angle-Consistent-Aware Hierarchical Attention for Cross-Modality Outlier Rejection

    May 6, 2026Muyao Peng, Shun Zou, Pei An +2Point Cloud RegistrationCross-Modal Feature Matching

  7. FLUID: Continuous-Time Hyperconnected Sparse Transformer for Sink-Free Learning

    May 6, 2026Waleed Razzaq, Yun-Bo ZhaoTransformerTransformer Attention

  8. How Language Models Process Negation

    May 4, 2026Zhejian Zhou, Tianyi Zhou, Robin Jia +1LLM InterpretabilityLanguage Modeling

  9. When Attention Collapses: Residual Evidence Modeling for Compositional Inference

    May 4, 2026Niklas HoubaGravitational-Wave AstronomySelf-Attention

  10. Projection-Free Transformers via Gaussian Kernel Attention

    May 4, 2026Debarshi Kundu, Archisman Ghosh, Swaroop Ghosh +1TransformerSelf-Attention

  11. Attention Is Where You Attack

    Apr 30, 2026Aviral Srivastava, Sourav PandaSelf-AttentionAdversarial Attacks

  12. Better Models, Faster Training: Sigmoid Attention for single-cell Foundation Models

    Apr 29, 2026Vijay Sadashivaiah, Georgios Dasoulas, Judith Mueller +1Softmax AttentionSelf-Attention

  13. From Architecture to Output: Structural Origins of Hallucination in Large Language Models and the Amplifying Role of Data

    Apr 29, 2026Md. Rejaul Korim Sadi, Toufiqur Rahman Tasin, Golam Mostofa NaeemHallucination in Language ModelsLLM Reliability

  14. GateMOT: Q-Gated Attention for Dense Object Tracking

    Apr 29, 2026Mingjin Lv, Zelin Liu, Feifei Shao +4Visual Object TrackingGated Attention

  15. AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving

    Apr 28, 2026Zhongkai Yu, Haotian Ye, Chenyang Zhou +9Compute-in-MemoryLLM Inference Acceleration

  16. Emergent Self-Attention from Astrocyte-Gated Associative Memory Dynamics

    Apr 28, 2026Arnau Vivet, Alex ArenasSoftmax AttentionDynamical Systems

  17. QFlash: Bridging Quantization and Memory Efficiency in Vision Transformer Attention

    Apr 28, 2026Sehyeon Oh, Yongin Kwon, Jemin LeeSoftmax AttentionGPU Acceleration

  18. Transformer Approximations from ReLUs

    Apr 27, 2026Jerry Yao-Chieh Hu, Mingcheng Lu, Yi-Chen Lee +1Softmax AttentionNeural Network Approximation Theory

  19. Learning to Rotate: Temporal and Semantic Rotary Encoding for Sequential Modeling

    Apr 27, 2026Hailing Cheng, Daqi Sun, Xinyu LuRotary Positional EmbeddingsSelf-Attention

  20. Learning to Route Queries to Heads for Attention-based Re-ranking with Large Language Models

    Apr 27, 2026Yuxing Tian, Fengran Mo, Zhiqi Huang +2Attention Head AnalysisLearning to Rank