机构
*
The Grainger College of Engineering, Nuclear, Plasma & Radiological Engineering, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校格雷格学院工程学院、核等工程学院)
;
Department of Nuclear Engineering, Hanyang University(汉阳大学核工程系)
;
University of Texas - El Paso(德克萨斯大学埃尔帕索分校)
;
National Center for Supercomputing Applications(国家超级计算应用中心)
;
Department of Applied Mechanics, Indian Institute of Technology Delhi(印度德里理工学院应用力学系)
;
Yardi School of Artificial Intelligence, Indian Institute of Technology Delhi(印度德里理工学院亚里人工智能学院)
What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness
什么使LVLMs更少产生幻觉?揭示影响幻觉鲁棒性的架构因素
Yusheng He, Jizhe Zhou, Xia Du, Zheng Lin, Jun Luo, Jiancheng Lv
机构
*
School of Computer Science, Engineering Research Center of Machine Learning and Industry Intelligence, Sichuan University(计算机科学学院,机器学习与产业智能工程研究中心,四川大学)
;
School of Computer and Information Engineering, Xiamen University of Technology(计算机与信息工程学院,厦门理工大学)
;
Department of Electrical and Computer Engineering, University of Hong Kong(电气与计算机工程系,香港大学)
;
College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学)
Cross-Modal Attention Calibration for LVLM Hallucination Mitigation
跨模态注意力校准用于LVLM幻觉缓解
Jiaming Li, Jiacheng Zhang, Zequn Jie, Lin Ma, Guanbin Li
机构
*
Sun Yat-sen University(中山大学)
;
The University of Hong Kong(香港大学)
;
Meituan(美团)
;
Inspur Database Technology(Inspur数据库技术)
;
Guilin University of Electronic Technology(桂林电子科技大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东大数据分析与处理重点实验室)
机构
*
School of Intelligence Science and Technology(智能科学与技术学院)
;
State Key Laboratory for Novel Software Technology(新型软件技术国家重点实验室)
;
Nanjing University(南京大学)
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.AI、cs.LG
Seeing vs. Believing: Evaluating the Language Bias of Open-Source MLLMs in Counter-Intuitive Scenes
看见 vs. 相信:评估开源多模态大模型在反直觉场景中的语言偏见
Chen Ling, Tongwei Zhang, Hanqian Li, Nai Ding
机构
*
Zhejiang University(浙江大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.CV、cs.AI
CommentsThis paper has been accepted by the International Journal of Computer Vision (IJCV), 2026. The first two authors contributed equally to this work. 28 pages
机构
*
Institute of Big Data Science and Industry(大数据科学与产业研究院)
;
Key Laboratory of Evolutionary Science Intelligence of Shanxi Province(山西省进化智能科学重点实验室)
;
School of Artificial Intelligence, Shanxi University(山西大学人工智能学院)
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.CV、cs.LG
Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy
在不遗忘的情况下寻找正确的视觉证据:通过层间视觉注意力差异减轻LVLMs中的幻觉
Yutong Xie, Zhenglin Hua, Ran Wang, Wing W. Y. Ng, Xizhao Wang, Yuheng Jia
机构
*
School of Computer Science and Engineering, Southeast University, Nanjing, China(东南大学计算机科学与工程学院)
;
School of Artificial Intelligence, Shenzhen University, Shenzhen, China(深圳大学人工智能学院)
;
College of Computer Science and Software Engineering, Shenzhen University, Shenzhen, China(深圳大学计算机科学与软件工程学院)
;
Engineering, South China University of Technology, Guangzhou, China(华南理工大学工程学院)
;
National Engineering Laboratory for Big Data Systems Computing Technology, Shenzhen University, Shenzhen, China(深圳大学大数据系统计算技术国家工程实验室)
;
Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及其交叉应用重点实验室(东南大学),中华人民共和国教育部)
机构
*
State Key Lab. of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院人工智能安全国家重点实验室,计算技术研究所)
;
School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学先进交叉科学学院)
;
School of Computer Science and Technology, University of Chinese Academy of Sciences(中国科学院大学计算机科学与技术学院)
Jailbreaking on Text-to-Video Models via Scene Splitting Strategy
通过场景分割策略对文本到视频模型进行劫持
Wonjun Lee, Haon Park, Doehyeon Lee, Bumsub Ham, Suhyun Kim
机构
*
Yonsei University(延世大学)
;
Korea Institute of Science and Technology(韩国科学技术院)
;
AIM Intelligence(AIM智能)
;
Seoul National University(首尔国立大学)
;
Kyung Hee University(庆熙大学)
When and Why Adversarial Training Improves PINNs: A Neural Tangent Kernel Perspective
何时以及为何对抗训练能提升PINNs:神经 tangent 核视角
Yuan-dong Cao, Chi Chiu SO, Jun-Min Wang, He Wang
机构
*
School of Mathematics and Statistics, Beijing Institute of Technology, China(北京理工大学数学与统计学院,中国)
;
School of Professional Education and Executive Development The Hong Kong Polytechnic University, China(香港理工大学专业教育学院及管理发展学院,中国)
;
Department of Computer Science & UCL AI Centre, University College London, UK(伦敦大学学院计算机科学系及UCL人工智能中心,英国)