Multi-Reward GRPO Fine-Tuning for De-biasing Large Language Models: A Study Based on Chinese-Context Discrimination Data
Deng Yixuan, Ji Xiaoqiang
机构
*
School of Science and Engineering, The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳)科学与工程学院)
;
School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳)人工智能学院)
;
Shenzhen Institute of Artificial Intelligence and Robotics for Society, China(深圳人工智能与机器人研究院)
专题命中
后训练与偏好优化
:large language model(title,abstract);language model(title,abstract);RLHF(abstract);分类 cs.CL
CommentsThe paper is withdrawn due to the need for further revision and verification of experimental results. A revised version will be resubmitted once the updates are completed
Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence
Sean McLeish, Ang Li, John Kirchenbauer, Dayal Singh Kalra, Brian R. Bartoldson, Bhavya Kailkhura, Avi Schwarzschild, Jonas Geiping, Tom Goldstein, Micah Goldblum
机构
*
University of Maryland(马里兰大学)
;
New York University(纽约大学)
;
Lawrence Livermore National Laboratory(劳伦斯利弗莫尔国家实验室)
;
University of North Carolina(北卡罗来纳大学)
;
ELLIS Institute Tübingen, Max Planck Institute for Intelligent Systems, Tübingen AI Center(图宾根ELLIS研究所、马克斯·普朗克智能系统研究所、图宾根人工智能中心)
;
Columbia University(哥伦比亚大学)
GRAPH-GRPO-LEX: Contract Graph Modeling and Reinforcement Learning with Group Relative Policy Optimization
Moriya Dechtiar, Daniel Martin Katz, Mari Sundaresan, Sylvain Jaume, Hongming Wang
机构
*
Harvard University(哈佛大学)
;
Illinois Tech - Chicago Kent College of Law(伊利诺伊理工学院-芝加哥肯特法学院)
;
CLTDS, Bucerius Law School(CLTDS,布塞里乌斯法学院)
;
Yong Pung How School of Law, Singapore Management University(永丰何法学院,新加坡管理大学)
;
CodeX - The Stanford Center for Legal Informatics, Stanford University(CodeX-斯坦福法律信息中心,斯坦福大学)
;
Georgetown University(乔治城大学)
;
Massachusetts Institute of Technology(麻省理工学院)
专题命中
后训练与偏好优化
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG