机构
*
School of Computer Science and Technology, Dalian University of Technology(大连理工大学计算机科学与技术学院)
;
Key Laboratory of Social Computing and Cognitive Intelligence (Dalian University of Technology), Ministry of Education(社会计算与认知智能重点实验室(大连理工大学),教育部)
;
Institute of Image Processing and Understanding, North Minzu University(北华大学图像处理与理解研究所)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
从噪声到信号到自我目的:在NLP后训练时代的重新框架化人类标签变异
Shanshan Xu, Santosh T. Y. S. S, Barbara Plank
机构
*
Department of Computer Science, University of Copenhagen, Denmark(丹麦哥本哈根大学计算机科学系)
;
Faculty of Law, University of Copenhagen, Denmark(丹麦哥本哈根大学法学院)
;
Amazon(亚马逊)
;
LMU Munich & Munich Center for Machine Learning (MCML)(慕尼黑大学及慕尼黑机器学习中心(MCML))
机构
*
Ruhr University Bochum(波鸿鲁尔大学)
;
Eindhoven University of Technology(埃因霍温理工大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
University of Liverpool(利物浦大学)
机构
*
Chinese Information Processing Laboratory, Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所信息处理实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
机构
*
Ant Group(蚂蚁集团)
;
Renmin University of China(中国人民大学)
;
Zhejiang University(浙江大学)
;
Westlake University(西湖大学)
;
HongKong University of Science and Technology(香港科学与技术大学)
机构
*
National Centre for Text Mining, The University of Manchester(曼彻斯特大学文本挖掘中心)
;
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
Center for Language and Information Research, Wuhan University(武汉大学语言与信息研究中心)
;
The Fin AI(Fin AI公司)
;
Baidu Inc.(百度公司)
;
Archimedes Research(阿基米德研究)
Tracing the Representation Geometry of Language Models from Pretraining to Post-training
Melody Zixuan Li, Kumar Krishna Agrawal, Arna Ghosh, Komal Kumar Teru, Adam Santoro, Guillaume Lajoie, Blake A. Richards
机构
*
Computer Science, McGill University(麦吉尔大学计算机科学系)
;
Mila - Quebec AI Institute(魁北克人工智能研究所)
;
UC Berkeley(加州大学伯克利分校)
;
G o o g l e , Paradigms of Intelligence Team(谷歌,智能范式团队)
;
Mathematics and Statistics, Université de Montréal(蒙特利尔大学数学与统计学系)
;
Neurology & Neurosurgery and Montreal Neurological Institute, McGill University(神经病学与神经外科及蒙特利尔神经研究所,麦吉尔大学)
;
CIFAR Learning in Machines & Brains Program(CIFAR 机器与大脑学习计划)
ConfPO: Exploiting Policy Model Confidence for Critical Token Selection in Preference Optimization
Hee Suk Yoon, Eunseop Yoon, Mark Hasegawa-Johnson, Sungwoong Kim, Chang D. Yoo
机构
*
Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院)
;
University of Illinois Urbana-Champaign (UIUC)(伊利诺伊大学厄巴纳-香槟分校)
;
Korea University(韩国大学)