机构
*
Om AI Research(Om AI研究机构)
;
Binjiang Institute of Zhejiang University(浙江大学滨江研究院)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
机构
*
Centre for Responsible Autonomous Systems in Healthcare (CRASH) Lab, Koita Centre for Digital Health(负责任的自主医疗系统中心(CRASH)实验室、Koita数字健康中心)
;
Ashoka University(阿什oka大学)
;
Independent Researcher(独立研究者)
;
Koita Centre for Digital Health(Koita数字健康中心)
专题命中
视觉推理
:visual reasoning(title,abstract);vision language model(abstract);分类 cs.AI、cs.LG
Comments29 pages, 7 figures, 7 tables, includes Annexure (1). Part of the work accepted at RSNA 2025 (Cutting Edge Oral Presentation)
Seeing Before Reasoning: A Unified Framework for Generalizable and Explainable Fake Image Detection
Kaiqing Lin, Zhiyuan Yan, Ruoxin Chen, Junyan Ye, Ke-Yue Zhang, Yue Zhou, Peng Jin, Bin Li, Taiping Yao, Shouhong Ding
机构
*
Guangdong Provincial Key Laboratory of Intelligent Information Processing, Shenzhen Key Laboratory of Media Security, and SZU-AFS Joint Innovation Center for AI Technology(广东省智能信息处理重点实验室、深圳媒体安全重点实验室及深圳大学-AFS联合人工智能技术创新中心)
;
Tencent Youtu Lab(腾讯优图实验室)
;
Peking University(北京大学)
;
Sun Yat-Sen University(中山大学)
专题命中
视觉推理
:multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Nanjing University(南京大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Peking University(北京大学)