机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Pengcheng Laboratory(鹏城实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Wenge Technology(文生科技)
;
Fudan University(复旦大学)
专题命中
后训练与偏好优化
:LLM(title,abstract);large language model(abstract);language model(abstract);post-training(abstract)
More is Less: The Pitfalls of Multi-Model Synthetic Preference Data in DPO Safety Alignment
Yifan Wang, Runjin Chen, Bolian Li, David Cho, Yihe Deng, Ruqi Zhang, Tianlong Chen, Zhangyang Wang, Ananth Grama, Junyuan Hong
机构
*
Purdue University(普渡大学)
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
The University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
University of California, Los Angeles(加州大学洛杉矶分校)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);post-training(abstract);RLHF(abstract)
CommentsThis version includes updated results and expanded discussion
Cultivating Helpful, Personalized, and Creative AI Tutors: A Framework for Pedagogical Alignment using Reinforcement Learning
Siyu Song, Wentao Liu, Ye Lu, Ruohua Zhang, Tao Liu, Jinze Lv, Xinyun Wang, Aimin Zhou, Fei Tan, Bo Jiang, Hao Hao
机构
*
Shanghai Innavation Institute(上海创新研究院)
;
Shanghai Institute of AI for Education(上海人工智能教育研究院)
;
School of Computer Science and Technology(计算机科学与技术学院)
;
Department of Educational Information Technology(教育信息技术系)
专题命中
后训练与偏好优化
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Preference learning made easy: Everything should be understood through win rate
Lily H. Zhang, Rajesh Ranganath
机构
*
Center for Data Science, New York University, New York, USA(数据科学中心,纽约大学,纽约,美国)
;
Courant Institute, New York University, New York, USA(柯朗研究所,纽约大学,纽约,美国)