机构
*
Institute of Science Tokyo, Department of Computer Science(东京科学研究所,计算机科学系)
;
National Institute of Advanced Industrial Science and Technology(国家先进工业科学与技术研究院)
;
Institute of Science Tokyo, Institute of Integrated Research, Supercomputing Research Center(东京科学研究所,整合研究所,超级计算研究中心)
Humans and LLMs Diverge on Probabilistic Inferences
人类与大语言模型在概率推断上存在分歧
Gaurav Kamath, Sreenath Madathil, Sebastian Schuster, Marie-Catherine de Marneffe, Siva Reddy
机构
*
McGill University(麦吉尔大学)
;
Mila – Quebec AI Institute(魁北克人工智能研究所)
;
University of Vienna(维也纳大学)
;
FNRS – UCLouvain(佛罗伦萨-鲁汶大学)
;
Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)
Spurious Rewards: Rethinking Training Signals in RLVR
虚假奖励:重新思考强化学习中的训练信号
Rulin Shao, Shuyue Stella Li, Rui Xin, Scott Geng, Yiping Wang, Sewoong Oh, Simon Shaolei Du, Nathan Lambert, Sewon Min, Ranjay Krishna, Yulia Tsvetkov, Hannaneh Hajishirzi, Pang Wei Koh, Luke Zettlemoyer
机构
*
University of Washington, Seattle, WA, USA(华盛顿大学)
;
Allen Institute for Artificial Intelligence, Seattle, WA, USA(人工智能研究院)
;
University of California, Berkeley, Berkeley, CA, USA(加州大学伯克利分校)
Bidipta Sarkar, Mattie Fellows, Juan Agustin Duque, Alistair Letcher, Antonio León Villares, Anya Sims, Clarisse Wibault, Dmitry Samsonov, Dylan Cope, Jarek Liesen, Kang Li, Lukas Seier, Theo Wolf, Uljad Berdica, Valentin Mohl, Alexander David Goldie, Aaron Courville, Karin Sevegnani, Shimon Whiteson, Jakob Nicolaus Foerster
Joint Continual Learning of Local Language Models and Cloud Offloading Decisions with Budget Constraints
带有预算约束的本地语言模型与云卸载决策的联合持续学习
Evan Chen, Wenzhi Fang, Shiqiang Wang, Christopher Brinton
机构
*
Elmore Family School of Electrical and Computer Engineering, Purdue University, West Lafayette, IN(电子与计算机工程学院,普渡大学)
;
Department of Computer Science, University of Exeter, UK(计算机科学系,埃克塞特大学)
Chao Huang, Yujing Lu, Quangang Li, Shenghe Wang, Yan Wang, Yueyang Zhang, Long Xia, Jiashu Zhao, Zhiyuan Sun, Daiting Shi, Tingwen Liu
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
Baidu Inc.(百度公司)
;
Wilfrid Laurier University(威尔弗里德·劳里埃大学)
Optimizing Prompts for Large Language Models: A Causal Approach
为大型语言模型优化提示:一种因果方法
Wei Chen, Yanbin Fang, Shuran Fu, Fasheng Xu, Xuan Wei
机构
*
School of Business, University of Connecticut(康涅狄格大学商学院)
;
Antai College of Economics and Management, Shanghai Jiao Tong University(上海交通大学安泰经济管理学院)