Simple Domain Generalization for Strong Pixel-Level Image Tampering Detection in Modern VLMs
现代视觉语言模型中用于强像素级图像篡改检测的简单域泛化
Yi Tang, Xinyi Shang, Jiacheng Cui, Sondos Mahmoud Bsharat, Jiacheng Liu, Xiaohan Zhao, Tran Dinh Tien, Ahmed Elhagry, Salwa K. Al Khatib, Tianjun Yao, Yonina C. Eldar, Jing-Hao Xue, Hao Li, Salman Khan, Zhiqiang Shen
机构
*
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
University College London(伦敦大学学院)
;
Weizmann Institute of Science(魏茨曼科学研究所)
MedLVR: Latent Visual Reasoning for Reliable Medical Visual Question Answering
MedLVR: 基于潜在视觉推理的可靠医学视觉问答
Suyang Xi, Songtao Hu, Yuxiang Lai, Wangyun Dan, Yaqi Liu, Shansong Wang, Xiaofeng Yang
机构
*
Department of Radiation Oncology and Winship Cancer Institute, Emory University School of Medicine(埃默里大学医学院放射肿瘤学系与温希普癌症研究所)
;
Department of Biostatistics and Bioinformatics, Emory University(埃默里大学生物统计学与生物信息学系)
Multi-Agent LLMs as Ethics Advocates for AI-Based Systems
多智能体大语言模型作为AI系统的伦理倡导者
Asma Yamani, Malak Baslyman, Moataz Ahmed
机构
*
Information and Computer Science Department, KFUPM(信息与计算机科学系,KFUPM)
;
IRC for finance and digital economy, KFUPM(金融与数字经济研究所,KFUPM)
;
SDAIA-KFUPM Joint Research Center for Artificial Intelligence, KFUPM(人工智能联合研究中心,KFUPM)
From Perception to Assistance: Open-Vocabulary Shared Autonomy for Robotic Manipulation
从感知到协助:用于机器人操作的开放词汇共享自主性
Murilo Vinicius da Silva, Ricardo V. Godoy, Juliano Negri, Gustavo J. G. Lahr, Ranulfo Bezerra, Marcelo Becker
机构
*
University of São Paulo(圣保罗大学)
;
Instituto Israelita de Ensino e Pesquisa Albert Einstein(以色列爱因斯坦教育与研究机构)
;
Graduate School of Information Sciences, Tohoku University(东北大学信息科学研究生院)
机构
*
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Khalifa University(哈里发大学)
;
Indian Institute of Technology Delhi(印度理工学院德里分校)
Probing the Difficulty Perception Mechanism of Large Language Models
探究大语言模型的难度感知机制
Sunbowen Lee, Qingyu Yin, Chak Tou Leong, Jialiang Zhang, Yicheng Gong, Shiwen Ni, Min Yang, Xiaoyu Shen
机构
*
Institute of Digital Twin, EIT(数字孪生研究所,EIT)
;
Wuhan University of Science and Technology(武汉科技大学)
;
Zhejiang University(浙江大学)
;
Hong Kong Polytechnic University(香港理工大学)
;
Shenzhen Institutes of Advanced Technology, CAS(深圳先进技术研究院,中国科学院)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
VOPE: Revisiting Hallucination of Vision-Language Models in Voluntary Imagination Task
VOPE:重新审视视觉语言模型在自愿想象任务中的幻觉现象
Xingming Long, Jie Zhang, Shiguang Shan, Xilin Chen
机构
*
Key Laboratory of AI Safety of CAS, Institute of Computing Technology, Chinese Academy of Sciences (CAS)(中国科学院人工智能安全重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Zhongguancun Academy(中关村学院)
CommentsAccepted for publication in the 30th Conference on Medical Image Understanding and Analysis (MIUA 2026), Dublin. To appear in Springer Lecture Notes in Computer Science (LNCS)
机构
*
Lero, the Research Ireland Centre for Software, University of Limerick(Lero,爱尔兰科学基金会软件研究中心,利默里克大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
University of Ottawa(渥太华大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL
CommentsThis is a preprint version. A shorter version of this paper has been accepted for presentation and publication in the post-workshop proceedings of the 8th International Workshop on eXplainable Knowledge Discovery in Data Mining (XKDD 2026), co-located with ECML PKDD 2026. The appendix is included only in this preprint and is not part of the peer-reviewed proceedings paper
机构
*
Namibia University of Science and Technology(纳米比亚科技大学)
;
Indian Institute of Technology Indore(印度理工学院印多尔分校)
;
Namdeb Diamond Corporation(纳米比亚钻石公司)