Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models
探究大语言模型忠实自解释的训练与泛化
机构 * The University of Tokyo(东京大学) ; Riken(理化学研究所) ; Tohoku University(东北大学) ; NII LLMC(日本信息处理学会大语言模型委员会)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL
AI总结 本研究通过训练提升大语言模型的自解释忠实性,并验证其在不同任务和风格中的泛化能力。
Comments To appear in the Proceedings of the Asia-Pacific Chapter of the Association for Computational Linguistics: Student Research Workshop (AACL-SRW 2025)