机构
*
Peking University(北京大学)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
Chinese University of Hong Kong(香港中文大学)
;
University of California, Santa Barbara(加州大学圣芭芭拉分校)
机构
*
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University)(东南大学新一代人工智能技术及其交叉应用关键实验室)
;
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Wuhan University of Technology(武汉理工大学)
;
National University of Singapore(新加坡国立大学)
CoViPAL: Layer-wise Contextualized Visual Token Pruning for Large Vision-Language Models
Zicong Tang, Ziyang Ma, Suqing Wang, Zuchao Li, Lefei Zhang, Hai Zhao, Yun Li, Qianren Wang
机构
*
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机学院)
;
Cognitive AI Lab, Shanghai Huawei Technologies, China(上海华为技术有限公司认知人工智能实验室)
VisionUnite: A Vision-Language Foundation Model for Ophthalmology Enhanced with Clinical Knowledge
Zihan Li, Diping Song, Zefeng Yang, Deming Wang, Fei Li, Xiulan Zhang, Paul E. Kinahan, Yu Qiao
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
University of Washington(华盛顿大学)
;
Shenzhen Institutes of Advanced Technology(深圳先进技术研究所)
;
Chinese Academy of Sciences(中国科学院)
;
State Key Laboratory of Ophthalmology(眼科学国家重点实验室)
;
Zhongshan Ophthalmic Center(中山眼科中心)
;
Sun Yat-sen University(中山大学)
;
Guangdong Provincial Key Laboratory of Ophthalmology and Visual Science(广东省眼科学与视觉科学重点实验室)
;
Guangdong Provincial Clinical Research Center for Ocular Diseases(广东省眼科临床研究中心)
;
Department of Bioengineering(生物工程系)
;
Department of Radiology(放射科)
Robotic Perception with a Large Tactile-Vision-Language Model for Physical Property Inference
Zexiang Guo, Hengxiang Chen, Xinheng Mai, Qiusang Qiu, Gan Ma, Zhanat Kappassov, Qiang Li, Nutan Chen
机构
*
College of Big Data and Internet, Shenzhen Technology University, China(大数据与互联网学院,深圳科技大学,中国)
;
Sino-German College of Intelligent Manufacturing, Shenzhen Technology University, China(中德智能制造学院,深圳科技大学,中国)
;
Robotics Department, Institute of Smart Systems and Artificial Intelligence (ISSAI), Nazarbayev University, Kazakhstan(机器人系,智能系统与人工智能研究所(ISSAI),纳扎尔巴耶夫大学,哈萨克斯坦)
;
Foundation Robotics Labs, Germany(基础机器人实验室,德国)
CommentsThis paper has been accepted by the 2025 International Conference on Climbing and Walking Robots (CLAWAR). These authors contributed equally to this work: Zexiang Guo, Hengxiang Chen, Xinheng Mai
The Safety Reminder: A Soft Prompt to Reactivate Delayed Safety Awareness in Vision-Language Models
Peiyuan Tang, Haojie Xin, Xiaodong Zhang, Jun Sun, Qin Xia, Zijiang Yang
机构
*
School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院)
;
School of Computer Science and Technology, University of Science and Technology of China(中国科学技术大学计算机科学与技术学院)
;
School of Computing and Information Systems, Singapore Management University(新加坡管理学院计算与信息系统学院)
机构
*
Microsoft Research(微软研究院)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Language Technology Lab, University of Cambridge(剑桥大学语言技术实验室)
;
Nanjing University(南京大学)
ReadBench: Measuring the Dense Text Visual Reading Ability of Vision-Language Models
Benjamin Clavié, Florian Brand
机构
*
Artificial Intelligence and Intelligent Information Systems, University of Trier(人工智能与智能信息系统,特里尔大学)
;
German Research Center for Artificial Intelligence (DFKI), Trier(德国人工智能研究中心(DFKI),特里尔)
Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos
Markos Stamatakis, Joshua Berger, Christian Wartena, Ralph Ewerth, Anett Hoppe
机构
*
TIB – Leibniz Information Centre for Science and Technology(蒂宾根-莱比锡信息科学与技术研究中心)
;
Hochschule Hannover – Data(汉诺威高等学院-数据)
;
H Institute for Applied Data Science(汉诺威应用数据科学研究所)
;
L3S Research Center – Leibniz University Hannover(L3S研究中心-汉诺威莱比锡大学)