Investigating the Multilingual Calibration Effects of Language Model Instruction-Tuning
探究语言模型指令微调的多语言校准效应
机构 * Mila - Quebec AI Institute(魁北克人工智能研究所) ; Université de Montréal(蒙特利尔大学) ; The University of Tokyo(东京大学) ; Western University(西方大学) ; Vector Institute(向量研究所) ; Polytechnique Montréal(蒙特利尔理工学院) ; CIFAR AI Chair(CIFAR人工智能主席) ; AIST(日本产业技术综合研究所)
AI总结 本研究探讨了多语言环境下语言模型指令微调对校准的影响,发现高资源语言SFT数据能显著提升模型置信度,但准确性提升有限,揭示了标准SFT在多语言中的局限性。
Comments Accepted to The 19th Conference of the European Chapter of the Association for Computational Linguistics (EACL)