arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-12-16 至 2025-12-16 共收录 6 信号源:cs.CL, cs.AI, cs.LG

1. 逻辑推理 6 篇

2512.12792 2025-12-16 cs.LG cs.AI 81%

Liquid Reasoning Transformers: A Sudoku-Based Prototype for Chess-Scale Algorithmic Tasks

液体推理变换器:基于数独的棋盘级算法任务原型

Shivansh Sahni, Wenzhi Zhang

机构 * Detroit Country Day School(底特律国家日学校)

专题命中 逻辑推理 :reasoning(title,abstract);分类 cs.AI、cs.LG

AI总结 液体推理变换器通过迭代更新和内部调整机制,在数独任务中实现了高准确率,展示了其在复杂算法任务中的潜力。

Comments 11 pages, 0 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12656 2025-12-16 cs.LO 78%

Argumentative Reasoning with Language Models on Non-factorized Case Bases

基于非因子化案例库的语言模型论证推理

Wachara Fungwacharakorn, May Myo Zin, Ha-Thanh Nguyen, Yuntao Kong, Ken Satoh

专题命中 逻辑推理 :reasoning(title,abstract)

AI总结 本文提出AAM-CBR框架,通过语言模型实现非因子化案例库的基于案例推理,提升灵活性和隐私性,并展示在因素数量增加时需结合符号推理以提高效果。

Comments Presented at NeLaMKRR@KR, 2025 (arXiv:2511.09575)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12888 2025-12-16 physics.optics cs.AI cs.CL cs.LG 71%

Meta-GPT: Decoding the Metasurface Genome with Generative Artificial Intelligence

Meta-GPT:用生成式人工智能解码元表面基因组

David Dang, Stuart Love, Meena Salib, Quynh Dang, Samuel Rothfarb, Mysk Alnatour, Andrew Salij, Hou-Tong Chen, Ho Wai, Lee, Wilton J. M. Kort-Kamp

专题命中 逻辑推理 :chain-of-thought(abstract,comments);分类 cs.CL、cs.AI、cs.LG;reasoning(comments)

AI总结 Meta-GPT通过METASTRINGS符号语言实现光子学设计,以生成式人工智能解码元表面基因组。

Comments Keywords: Physics-informed machine learning; Transformer models; Reinforcement learning; Chain-of-thought reasoning; Metasurfaces; Nanophotonics; Inverse design

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12817 2025-12-16 cs.HC cs.AI 70%

Decoding Human and AI Persuasion in National College Debate: Analyzing Prepared Arguments Through Aristotle's Rhetorical Principles

解码人类与AI的说服力:通过亚里士多德修辞原则分析辩论中的准备论点

Mengqian Wu, Jiayi Zhang, Raymond Z. Zhang

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

AI总结 本研究通过亚里士多德修辞原则分析人类与AI在辩论中的说服力,比较GPT-4与人类辩手生成的证据卡片质量,探讨AI在辩论训练中的应用与局限性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12487 2025-12-16 cs.CV 67%

More Than the Final Answer: Improving Visual Extraction and Logical Consistency in Vision-Language Models

不止于最终答案:提升视觉提取和逻辑一致性的视觉语言模型

Hoang Anh Just, Yifei Fan, Handong Zhao, Jiuxiang Gu, Ruiyi Zhang, Simon Jenni, Kushal Kafle, Ruoxi Jia, Jing Shi

机构 * Virginia Tech(弗吉尼亚理工大学) Adobe Research(Adobe研究院) Apple(苹果公司)

专题命中 逻辑推理 :reasoning(abstract);chain-of-thought(abstract)

AI总结 PeRL-VL通过解耦框架提升视觉语言模型的视觉提取和推理一致性,实现Pass@1准确率提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13505 2025-12-16 cs.AI 57%

Defending the Hierarchical Result Models of Precedential Constraint

为 precedential constraint 的分层结果模型辩护

Henry Prakken, Wijnand van Woerkom

机构 * Department of Information and Computing Sciences, Utrecht University, The Netherlands(信息与计算科学系,乌特勒支大学,荷兰) Max Planck Institute for Comparative and International Private Law, Hamburg, Germany(比较与国际私法Max Planck研究所,汉堡,德国)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

AI总结 本文针对 van Woerkom 的分层结果模型,回应了 Bench-Capon 对其在某些情况下可能产生错误结果的批评,并指出通过将中间因素视为维度可避免这些批评。

Comments This is the long version of a paper with the same title presented at the 38th International Conference on Legal Knowledge and Information Systems

详情

展开后加载摘要…

URL PDF HTML 收藏