Bolzano: Case Studies in LLM-Assisted Mathematical Research
Bolzano:大语言模型辅助数学研究的案例研究
Martin Balko, Jan Grebík, Pavel Hubáček, Martin Koutecký, Matěj Kripner, Václav Rozhoň, Robert Šámal, Adrián Zámečník
机构
*
Computer Science Institute, Charles University(查尔斯大学计算机科学研究所)
;
Institute of Mathematics, Czech Academy of Sciences(捷克科学院数学研究所)
;
Institute of Formal and Applied Linguistics, Charles University(查尔斯大学形式与应用语言学研究所)
;
Department of Applied Mathematics, Charles University(查尔斯大学应用数学系)
CommentsWe argue that when the goal is to intervene on model behavior, skill characterization should be *model-native*: grounded in the model's own representations rather than imposed through external ontologies
Revisiting Entropy Regularization: Adaptive Coefficient Unlocks Its Potential for LLM Reinforcement Learning
重新审视熵正则化:自适应系数解锁其在大语言模型强化学习中的潜力
Xiaoyun Zhang, Xiaojian Yuan, Di Huang, Wang You, Chen Hu, Jingqing Ruan, Ai Jian, Kejiang Chen, Xing Hu
机构
*
State Key Lab of Processors, Institute of Computing Technology, CAS(处理器国家重点实验室,计算技术研究所,中国科学院)
;
University of Science and Technology of China(中国科学技术大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
StepFun Inc(StepFun公司)
SAHOO: Safeguarded Alignment for High-Order Optimization Objectives in Recursive Self-Improvement
SAHOO:递归自我改进中高阶优化目标的安全保障
Subramanyam Sahoo, Aman Chadha, Vinija Jain, Divya Chaudhary
机构
*
MARS 4.0 Fellowship, Cambridge AI Safety Hub(CAISH), University of Cambridge(剑桥大学)
;
AWS Generative AI Innovation Center, Amazon Web Services, USA(亚马逊网络服务)
;
Google, USA(谷歌)
;
Stanford University(斯坦福大学)
;
Northeastern University, Seattle, WA, USA(东北大学)
机构
*
Zhejiang University(浙江大学)
;
Ant Group(蚂蚁集团)
;
National University of Singapore(新加坡国立大学)
;
Zhejiang University - Ant Group Joint Laboratory of Knowledge Graph(知识图谱联合实验室)
Menglin Yang, Ram Samarth B B, Aosong Feng, Bo Xiong, Jihong Liu, Irwin King, Rex Ying
机构
*
HKUST(GZ)(香港科技大学(广州))
;
HKUST(香港科技大学)
;
Indian Institute of Science(印度科学研究院)
;
Yale University(耶鲁大学)
;
Stanford University(斯坦福大学)
;
The Chinese University of Hong Kong(香港中文大学)
Hearing is Believing? Evaluating and Analyzing Audio Language Model Sycophancy with SYAUDIO
听信还是不信?评估和分析音频语言模型的趋炎附势行为 with SYAUDIO
Junchi Yao, Lokranjan Lakshmikanthan, Annie Zhao, Danielle Zhao, Shu Yang, Zikang Ding, Di Wang, Lijie Hu
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed人工智能大学)
;
King Abdullah University of Science and Technology(卡布斯大学)
;
Georgia Institute of Technology(佐治亚理工学院)