Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities
大型视觉语言模型(LVLMs)能否揭示视觉错觉背后的真相?感知与推理能力分析
机构 * Adelaide University(阿德莱德大学) ; Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) ; University of Chinese Academy of Sciences(中国科学院大学) ; Tsinghua University(清华大学) ; Amap, Alibaba Group(阿里巴巴集团高德地图) ; Beihang University(北京航空航天大学)
AI总结 本文利用IllusionReasoning基准,以视觉错觉为诊断工具评估LVLMs,发现多款LVLMs的推理能力不及宣称水平,为其优化提供了新方向。
Comments EMNLP2026 Findings