Apr 27, 2026·Hongxin Li, Yuntao Chen, Zhaoxiang ZhangGraphical User InterfaceVision-Language Model Grounding
University of Chinese Academy of Sciences, Beijing, China. · New Laboratory of Pattern Recognition, State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences, Beijing, China. · Hong Kong Institute of Science & Innovation, Chinese Academy of Sciences, Hong Kong, China.