Can Vision Language Models Understand Mimed Actions?
机构 * Information Sciences Institute(信息科学研究所) ; Institute for Creative Technologies(创意技术研究所) ; Department of Computer Science(计算机科学系) ; University of Southern California(南加州大学) ; University of California, Santa Barbara(加州大学圣巴巴拉分校) ; Aristotle University of Thessaloniki(希腊雅典纳大学)
专题命中 视觉问答 :vision language model(title);vision-language model(abstract);分类 cs.CV、cs.AI
Comments ACL 2025 Findings