9 papers in the last four weeks, up 50% on the four weeks before. 0.1% of all new papers.
May 31, 2026·Mingi KangSelf-AttentionConvolutional Neural Networks
Bowdoin College
May 29, 2026·Harry Jake Cunningham, Nicola Muca CironeTransformer InterpretabilityFeature Attribution
Department of Computer Science, University College London, London, UK. · Cartesia.AI, San Francisco, US.
May 28, 2026·Hidir Yesiltepe, Jiazhen Hu, Tuna Han Salih Meral +4Video Diffusion ModelsSelf-Attention
1Virginia Tech · 2fal
May 28, 2026·Masaaki Imaizumi, Masanori Koyama, Noboru Isobe +1TransformerSelf-Attention
The University of Tokyo, Tokyo, Japan · RIKEN Advanced Intelligene Project, Tokyo, Japan · Kyoto University, Kyoto, Japan
May 28, 2026·Matthew Smart, Soumya Ganguly, Nilava Metya +2Self-AttentionEmpirical Bayes
Lewis-Sigler Institute for Integrative Genomics, Princeton University, Princeton, NJ, USA · Department of Mathematics, Rutgers University, Piscataway, NJ, USA · Department of Physics and Astronomy, Rutgers University, Piscataway, NJ, USA +1
May 27, 2026·Yifei Zuo, Dhruv Pai, Zhichen Zeng +3Language Model PretrainingSelf-Attention
Northwestern University · Tilde Research · University of Washington
May 27, 2026·Xiuying Wei, Caglar GulcehreMemory-Augmented Language ModelsSelf-Attention
EPFL, CLAIRE Lausanne, Switzerland
May 27, 2026·Jiangsheng, YouRecursive Least SquaresSelf-Attention
Jason
May 27, 2026·Alan FerrariEfficient Transformer InferenceSelf-Attention
Knowledge Lab AG Zürich, Switzerland
May 26, 2026·Keqi Deng, Shaoshi Ling, Ruchao Fan +1Self-AttentionLong-Context Language Model Inference
Microsoft, USA
May 26, 2026·Hyunmin Cho, Woo Kyoung Han, Kyong Hwan JinSelf-AttentionTransformer Attention
Department of Electrical Engineering, Korea University, Seoul, South Korea.
May 26, 2026·Semi Lee, Hyejin Go, Hyesong ChoiToken MergingSelf-Attention
Electronic Engineering · Soongsil University · Seoul, South Korea
May 25, 2026·Athanasios ZerisSelf-AttentionGated Attention
Independent Researcher, Athens, Greece.
May 25, 2026·Tuna Tuncer, Felix Becker, Thomas PfeilVideo Diffusion ModelsSelf-Attention
1Technical University of Munich · 2Tensordyne
May 25, 2026·Xintong Yang, Hao Gu, Binxing Xu +6Memory-Augmented Language ModelsSelf-Attention
The Hong Kong University of Science and Technology · Zhejiang University
May 24, 2026·Shogo Yamauchi, Tohru Nitta, Hideaki TamoriSelf-AttentionAttention Mechanisms
The Asahi Shimbun Company, Tokyo, Japan · Tokyo Woman’s Christian University, Tokyo, Japan.
May 23, 2026·Naoki Kiyohara, Harrison Bo Hua Zhu, Riccardo El Hassanin +4Self-AttentionLong-Context Language Modeling
May 22, 2026·Pál András Papp, Aleksandros Sobczyk, Anastasios ZouziasSelf-AttentionEfficient Attention
Computing Systems Lab, Huawei Technologies, Zurich, Switzerland
May 22, 2026·Guoqiang ZhangVisual AttentionSelf-Attention
University of Exeter
May 22, 2026·Yongzhong XuTransformer InterpretabilityCircuit Discovery
May 21, 2026·Joe SharrattMixed-Precision QuantizationSelf-Attention
May 21, 2026·Hangyue Zhao, Paul Caillon, Erwan Fagnou +1Self-AttentionStructured Sparsity
ESPCI PSL, Paris, France · LAMSADE, Université Paris Dauphine - PSL, Paris, France
May 21, 2026·Jaehyuk Lee, Hanyoung Kim, Yanggee Kim +1Self-AttentionEfficient ViTs
Department of Mathematics Korea University Seoul, Republic of Korea · Korea University Seoul, Republic of Korea
May 21, 2026·Kabir Swain, Sijie Han, Daniel Karl I. Weidele +2Memory-Augmented Neural NetworksSelf-Attention
Massachusetts Institute of Technology, Cambridge, MA, USA · University of Toronto, Toronto, Canada · IBM Research, Cambridge, MA, USA
May 21, 2026·Athanasios ZerisSelf-AttentionGated Attention
May 20, 2026·Ian Li, Kapilesh Guruprasad, Raunak Sengupta +3Language Model SteeringSelf-Attention
University of California, San Diego
May 20, 2026·Gonçalo Duarte, Miguel Couceiro, Marcos V. TrevisoSelf-AttentionKV Caching
Instituto Superior Técnico, Universidade de Lisboa. · ELLIS Unit Lisbon. · INESC-ID. +1
May 20, 2026·Omar Coser, Loredana Zollo, Paolo Soda +1Self-AttentionMasked Language Modeling
May 20, 2026·Yanyi Lyu, Letian Chen, Futing Sun +3Self-AttentionDiffusion Language Model Inference
Harbin Institute of Technology (Shenzhen)
May 19, 2026·Wenhu Zhang, Yiming Wu, Huanyu Wang +6Self-AttentionLong-Context Language Modeling
The Hong Kong University of Science and Technology · The University of Hong Kong · Zhejiang University +1