arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2026-03-03 至 2026-03-03 共收录 5 信号源:cs.CL, cs.AI, cs.LG

1. 代码与定理证明 5 篇

2506.20666 2026-03-03 cs.CL cs.AI 62%

Cognitive models can reveal interpretable value trade-offs in language models

认知模型可以揭示语言模型中的可解释价值权衡

Sonia K. Murthy, Rosie Zhao, Jennifer Hu, Sham Kakade, Markus Wulfmeier, Peng Qian, Tomer Ullman

机构 * Kempner Institute for Natural and Artificial Intelligence, Harvard University(哈佛大学自然与人工智能研究所) Google DeepMind(谷歌DeepMind) Department of Psychology, Harvard University(哈佛大学心理学系)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 本文提出利用认知模型评估语言模型中的价值权衡,揭示其行为特征变化及对社会行为的影响。

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01407 2026-03-03 cs.AI cs.MA cs.SI 57%

The Observer-Situation Lattice: A Unified Formal Basis for Perspective-Aware Cognition

观察者-情境格:一种面向视角感知认知的统一基础

Saad Alqithami

机构 * Computer Science Department, Al-Baha University(阿尔巴大学计算机科学系)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

AI总结 本文提出观察者-情境格(OSL),一种统一的数学结构,用于支持面向视角的认知推理,通过两个关键算法实现信念管理和矛盾分解,提升多智能体环境中的推理效率和鲁棒性。

Comments Extended version of the AAMAS 2026 paper with the same title

Journal ref 25th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2026), Paphos, Cyprus, May 25-29, 2026, IFAAMAS, 18 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01170 2026-03-03 cs.CR cs.AI 57%

ATLAS: AI-Assisted Threat-to-Assertion Learning for System-on-Chip Security Verification

ATLAS:基于人工智能的威胁到断言学习用于片上系统安全验证

Ishraq Tashdid, Kimia Tasnia, Alexander Garcia, Jonathan Valamehr, Sazadur Rahman

机构 * Department of Electrical and Computer Engineering, University of Central Florida(电子与计算机工程系,中央佛罗里达大学) Intel Corporation(英特尔公司)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

AI总结 ATLAS通过结合人工智能和形式验证技术,实现了基于漏洞知识的自动化SoC安全验证,提高了安全设计的可靠性。

Comments Accepted at the 63rd Design Automation Conference (DAC 2026), Long Beach, CA, USA (July, 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23239 2026-03-03 cs.AI cs.CY 57%

Agency and Architectural Limits: Why Optimization-Based Systems Cannot Be Norm-Responsive

代理与架构限制:为何基于优化的系统无法回应规范

Radha Sarma

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

AI总结 本文指出基于优化的系统无法回应规范,提出真正的代理性需要不可比较的边界和否定性回应机制,揭示了优化与规范治理的不兼容性及由此引发的收敛危机。

Comments About 10,600 words in all (includes ~1000 words of literature and ~2400 words of Appendices). Under journal review

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00617 2026-03-03 cs.CY 50%

Dial E for Ethical Enforcement: institutional VETO power as a governance primitive

为伦理执行而 Dial E:作为治理原语的机构 veto 权力

Subramanyam Sahoo, Vinija Jain, Aman Chadha, Divya Chaudhary

专题命中 代码与定理证明 :reasoning(abstract)

AI总结 本文提出机构 veto 权力作为治理原语,旨在通过正式权威阻止潜在武器化风险,推动模型去武装和伦理执行。

Comments Accepted at the AI for Peace Workshop at ICLR 2026.32 Pages and 2 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏