Language Models Can Resolve Reference Compositionally, But It's Not Their Native Strength: The Case of the Personal Relation Task
语言模型可以组合性地解析指代,但这并非其天然优势:以个人关系任务为例
专题命中 视觉定位与Grounding :grounding(abstract)
AI总结 通过个人关系任务,比较人类与大型语言模型在外延任务(确定指称对象)和内涵任务(结构化表示意义)上的表现,发现人类更擅长外延任务而LLM更擅长内涵任务,表明缺乏指称基础是LLM模拟人类语言理解的关键缺失。
Comments A pre-MIT Press publication version. Paper accepted to Transactions of the Association for Computational Linguistics