机构
*
School of Future Technology, South China University of Technology(未来技术学院,华南理工大学)
;
School of Software Engineering, South China University of Technology(软件工程学院,华南理工大学)
;
Shien-Ming Wu School of Intelligent Engineering, South China University of Technology(智能工程学院,华南理工大学)
专题命中
图文多模态
:multimodal(abstract);分类 cs.AI
Comments15 pages, 4 figures, under review of NeurIPS
Visual hallucination detection in large vision-language models via evidential conflict
Tao Huang, Zhekun Liu, Rui Wang, Yang Zhang, Liping Jing
机构
*
Beijing Key Lab of Traffic Data Mining(北京交通数据挖掘与具身智能重点实验室)
;
State Key Laboratory of Advanced Rail Autonomous Operation(先进轨道交通自主运行国家重点实验室)
;
School of Computer Science and Technology(计算机科学与技术学院)
;
Beijing Jiaotong University(北京交通大学)
;
School of Automation and Intelligence(自动化与智能学院)
;
School of Electronic and Information Engineering(电子与信息工程学院)
专题命中
图文多模态
:multimodal(abstract);分类 cs.CV
Journal refInternational Journal of Approximate Reasoning, Volume 186, November 2025, Article 109507
Continual Retinal Vision-Language Pre-training upon Incremental Imaging Modalities
Yuang Yao, Ruiqi Wu, Yi Zhou, Tao Zhou
机构
*
School of Computer Science and Engineering, Southeast University, China(计算机科学与工程学院,东南大学,中国)
;
Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications, Ministry of Education, China(新一代人工智能技术及其交叉应用重点实验室,教育部,中国)
;
Nanjing University of Science and Technology, China(南京理工大学,中国)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Institute of Information Science, Beijing Jiaotong University(北京交通大学信息科学研究院)
;
School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院)
A CLIP-Powered Framework for Robust and Generalizable Data Selection
Suorong Yang, Peng Ye, Wanli Ouyang, Dongzhan Zhou, Furao Shen
机构
*
National Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong(香港中文大学)
Evolution of ReID: From Early Methods to LLM Integration
Amran Bhuiyan, Mizanur Rahman, Md Tahmid Rahman Laskar, Aijun An, Jimmy Xiangji Huang
机构
*
Information Retrieval and Knowledge Management Research Lab, York University, Toronto, Canada(信息检索与知识管理研究实验室,约克大学,多伦多,加拿大)
;
Department of Electrical Engineering and Computer Science, York University, Toronto, Canada(电气工程与计算机科学系,约克大学,多伦多,加拿大)
机构
*
School of Computer Science and Engineering, Sun Yat-Sen University(中山大学计算机科学与工程学院)
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy(广东省人工智能与数字经济实验室)
;
Huawei Technologies Co., Ltd(华为技术有限公司)
机构
*
Defense Innovation Institute, Chinese Academy of Military Science(国防科技研究院,中国军事科学院)
;
Tianjin Artificial Intelligence Innovation Center(天津人工智能创新中心)
;
School of Mechanical Engineering, Tianjin University(天津大学机械工程学院)
Does Your 3D Encoder Really Work? When Pretrain-SFT from 2D VLMs Meets 3D VLMs
Haoyuan Li, Yanpeng Zhou, Yufei Gao, Tao Tang, Jianhua Han, Yujie Yuan, Dave Zhenyu Chen, Jiawang Bian, Hang Xu, Xiaodan Liang
机构
*
Shenzhen campus of Sun Yat-sen University(中山大学深圳校区)
;
Huawei Noah’s Ark Lab(华为诺亚实验室)
;
MBZUAI
;
Peng Cheng Laboratory(鹏城实验室)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东省大数据分析与处理重点实验室)