arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-08-18 至 2025-08-18 共收录 1 信号源:cs.CL, cs.AI, cs.LG

1. 测试时计算 1 篇

2505.10981 2025-08-18 cs.AI cs.CL cs.LG 80%

Rethinking the Role of Prompting Strategies in LLM Test-Time Scaling: A Perspective of Probability Theory

Yexiang Liu, Zekun Li, Zhi Fang, Nan Xu, Ran He, Tieniu Tan

机构 * MAIS, Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) University of California, Santa Barbara(加州大学圣巴巴拉分校) Beijing Wenge Technology Co., Ltd(北京文格科技有限公司)

专题命中 测试时计算 :reasoning(abstract);chain-of-thought(abstract);test-time compute(abstract);分类 cs.CL、cs.AI、cs.LG

Comments ACL 2025 Outstanding Paper Award, 33 pages, 51 figures

Journal ref ACL.Volume 1: Long Papers (2025) 27962-27994

详情

展开后加载摘要…

URL PDF HTML 收藏