ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
ExpliCa:评估大型语言模型中的显式因果推理
机构 * CoLing Lab, Department of Philology, Literature, and Linguistics, University of Pisa, Italy(皮尔森大学哲学、文学与语言学系协作语言实验室) ; Department of Informatics, University of Pisa, Italy(皮尔森大学信息学系) ; Department of Chinese and Bilingual Studies, The Hong Kong Polytechnic University(香港理工大学中文与双语研究系)
AI总结 ExpliCa数据集用于评估大型语言模型在显式因果推理中的能力,发现顶级模型在准确率上仍存在显著不足,且模型性能受语言顺序和大小影响明显。
Comments Accepted for publication in Findings of ACL 2025
Journal ref In Findings of the Association for Computational Linguistics: ACL 2025, pages 17335-17355, Vienna, Austria. Association for Computational Linguistics