Scaling Large Reasoning Models beyond Human Supervision: A Path toward Superintelligence
超越人类监督扩展大型推理模型:通往超级智能的路径
机构 * The Hong Kong University of Science and Technology(香港科技大学) ; Zhongguancun Academy(中关村学院) ; Xi’an Jiaotong University(西安交通大学) ; The Chinese University of Hong Kong(香港中文大学) ; The University of Hong Kong(香港大学) ; Hong Kong Baptist University(香港浸会大学) ; Hunyuan Tencent(腾讯混元) ; National University of Singapore(新加坡国立大学) ; Xiamen University(厦门大学)
AI总结 本文研究人类监督逐步退出时大型推理模型(LRMs)的扩展路径,提出五级阶梯框架,分析自主化带来的风险,围绕策略能力等开展评估,为开发通往超级智能的自我维持学习系统提供结构化方案。
Comments 72pages