7 papers in the last four weeks, up 133% on the four weeks before. 0.1% of all new papers.
May 7, 2026·Kuanwei Lin, Wenhao Zhang, Ge LiEfficient VLM InferenceLong-Video Understanding
School of Electronic and Computer Engineering, Peking University
May 6, 2026·Haibin He, Maoyuan Ye, Jing Zhang +2Video QAVideo Frame Selection
School of Computer Science, National Engineering Research Center for Multimedia Software, Institute of Artificial Intelligence, and Hubei Key Laboratory of Multimedia and Network Communication Engineering, Wuhan University, Wuhan, Hubei, China
May 3, 2026·Martin Q. Ma, Willis Guo, Aditya Agrawal +4Active PerceptionEfficient VLM Inference
Carnegie Mellon University · MIT
Apr 19, 2026·Shaoguang Wang, Weiyu Guo, Ziyang Chen +2Multimodal Large Language ModelsMultimodal QA
Thrust of Artificial Intelligence, HKUST (Guangzhou) Guangzhou, China · Department of CSE, HKUST Hong Kong SAR, China
Mar 19, 2026·Dan Ben-Ami, Gabriele Serussi, Kobi Cohen +1Multimodal Large Language ModelsMultimodal QA
INSIGHT Lab, Ben-Gurion University of the Negev, Israel · Ben-Gurion University of the Negev, Israel
Mar 4, 2026·Tatiana Zemskova, Solomon Andryushenko, Ilya Obrubov +4Efficient VLM InferenceEgocentric Video QA
AXXX · MIRAI · Yandex +1