The Alignment Paradox of Medical Large Language Models in Infertility Care: Decoupling Algorithmic Improvement from Clinical Decision-making Quality
医学大语言模型在不孕症护理中的对齐悖论:解耦算法改进与临床决策质量
Dou Liu, Ying Long, Sophia Zuoqiu, Kaipeng Xie, Runze Yang, Di Liu, Kang Li, Yiting Lin, Hanyi Liu, Rong Yin, Tian Tang
机构
*
Department of Obstetrics and Gynecology, West China Second University Hospital(妇产科部门,西昌第二大学医院)
;
Reproductive Medical Center, Department of Obstetrics and Gynecology, West China Second University Hospital(生殖医学中心,妇产科部门,西昌第二大学医院)
Towards Reward Fairness in RLHF: From a Resource Allocation Perspective
Sheng Ouyang, Yulan Hu, Ge Chen, Qingyang Li, Fuzheng Zhang, Yong Liu
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院)
;
Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理研究重点实验室)
;
Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程技术研究中心,教育部)
;
Kuaishou Technology(快手科技)
;
University of Chinese Academy of Sciences(中国科学院大学)
Yu Pan, Zhongze Cai, Guanting Chen, Huaiyang Zhong, Chonghuan Wang
机构
*
University of Sydney(悉尼大学)
;
Imperial College London(伦敦帝国理工学院)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
Virginia Tech(弗吉尼亚理工大学)
;
University of Texas at Dallas(德克萨斯大学达拉斯分校)
Offline and Online KL-Regularized RLHF under Differential Privacy
Yulian Wu, Rushil Thareja, Praneeth Vepakomma, Francesco Orabona
机构
*
King Abdullah University of Science and Technology (KAUST)(卡布尔大学科学与技术学院)
;
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)
;
Massachusetts Institute of Technology (MIT)(麻省理工学院)
机构
*
Institute of Artificial Intelligence (TeleAI), China Telecom(人工智能研究院(TeleAI),中国电信)
;
RUC(中国人民大学)
;
USTC(University of Science and Technology of China)
;
NTU(National University of Technology)
;
NUS(National University of Singapore)
;
USC(University of Southern California)
;
SCU(Sichuan University)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
Avoiding $\mathbf{exp(R_{max})}$ scaling in RLHF through Preference-based Exploration
Mingyu Chen, Yiding Chen, Wen Sun, Xuezhou Zhang
机构
*
Department of Electrical & Computer Engineering(电气与计算机工程系)
;
Boston University(波士顿大学)
;
Department of Computer Science(计算机科学系)
;
Cornell University(康奈尔大学)
;
Faculty of Computing & Data Sciences(计算与数据科学学院)
InfAlign: Inference-aware language model alignment
Ananth Balashankar, Ziteng Sun, Jonathan Berant, Jacob Eisenstein, Michael Collins, Adrian Hutter, Jong Lee, Chirag Nagpal, Flavien Prost, Aradhana Sinha, Ananda Theertha Suresh, Ahmad Beirami
机构
*
Google DeepMind(谷歌DeepMind)
;
Google Research(谷歌研究)
机构
*
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Google Cloud AI Research(谷歌云人工智能研究)
;
Google DeepMind(谷歌DeepMind)
;
University of Virginia(弗吉尼亚大学)
机构
*
Department of Statistics, Rutgers University, New Brunswick, United States
;
Department of Computer Science, Rutgers University, New Brunswick, United States
;
College of Management of Technology, EPFL, Switzerland
;
Department of Computer Science, ETH Zurich, Switzerland
Yinuo Ren, Tesi Xiao, Michael Shavlovsky, Lexing Ying, Holakou Rahmanian
机构
*
Institute for Computational and Mathematical Engineering (ICME), Stanford University(计算与数学工程研究所(ICME),斯坦福大学)
;
Amazon(亚马逊)
;
Department of Mathematics, Stanford University(数学系,斯坦福大学)