Do Images Speak Louder than Words? Investigating the Effect of Textual Misinformation in VLMs
图像胜过言语吗?探讨文本误导在VLMs中的影响
机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校) ; Pennsylvania State University(宾夕法尼亚州立大学) ; New York University(纽约大学) ; University of Chinese Academy of Sciences(中国科学院大学) ; AG2ai, Inc.(AG2ai公司)
专题命中 视觉问答 :vision-language model(abstract)
AI总结 研究探讨了文本误导对视觉-语言模型(VLMs)的影响,发现模型易受误导性文本提示影响,导致性能显著下降。
Comments 24 pages, 10 figures. Accepted at EACL 2026 (main conference)