机构
*
AMAP, Alibaba Group(阿里集团AMAP)
;
University of California at Merced(加州大学默塞德分校)
;
University of Queensland(昆士兰大学)
;
Case Western Reserve University(凯斯西储大学)
Mema: Memory-Augmented Adapter for Enhanced Vision-Language Understanding
Mema:增强视觉-语言理解的内存增强适配器
Ying Liu, Yudong Han, Kean Shi, Liyuan Pan
机构
*
Beijing Institute of Technology(北京理工大学)
;
Peking University(北京大学)
;
Yangtze Delta Region Academy of Beijing Institude of Technology(北京理工大学长江三角洲地区研究院)
机构
*
Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)
;
Peng Cheng Laboratory(鹏城实验室)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
Lenovo Research(联想研究院)
Comments9 pages, 5 figures. Accepted to workshop on AI and Partial Differential Equations, Foundation Models for Science: Real-World Impact and Science-First Design, Machine Learning for Genomics Explorations, and Generative and Experimental Perspectives for Biomolecular Design at ICLR 2026
机构
*
University of California, Los Angeles(加州大学洛杉矶分校)
;
Tencent Hunyuan(腾讯混元)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
机构
*
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Southeast University(东南大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
National University of Singapore(新加坡国立大学)
;
Wuhan AI Research(武汉人工智能研究院)
;
Guangdong Provincial Key Laboratory of Intellectual Property and Big Data, Guangdong Polytechnic Normal University(广东技术师范大学广东省知识产权大数据重点实验室)
Bohan Jia, Wenxuan Huang, Yuntian Tang, Junbo Qiao, Jincheng Liao, Shaosheng Cao, Fei Zhao, Zhaopeng Feng, Zhouhong Gu, Zhenfei Yin, Lei Bai, Wanli Ouyang, Lin Chen, Fei Zhao, Yao Hu, Zihan Wang, Yuan Xie, Shaohui Lin
机构
*
East China Normal University(华东师范大学)
;
Xiaohongshu Inc.(小红书)
;
KLATASDS, MOE, China(中国教育部统计与数据科学重点实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Zhejiang University(浙江大学)
;
Fudan University(复旦大学)
;
University of Oxford(牛津大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Nanjing University(南京大学)
RVLM: Recursive Vision-Language Models with Adaptive Depth
具有自适应深度的递归视觉-语言模型
Nicanor Mayumu, Zeenath Khan, Melodena Stephens, Patrick Mukala, Farhad Oroumchian
机构
*
Department of Computer Science(计算机科学系)
;
University of Wollongong in Dubai(迪拜沃林戈大学)
;
Dubai Knowledge Park(迪拜知识园区)
;
Mohammed Bin Rashid School of Government(穆罕默德·本·拉希德政府学院)
CommentsAccepted at the 27th International Conference on Artificial Intelligence in Education (AIED 2026). The final authenticated version will appear in Springer LNAI/LNCS proceedings