arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

共收录 45051 信号源:cs.CL, cs.AI, cs.LG

1. 逻辑推理 3033 篇

2603.10891 2026-03-12 cs.AI cs.IR 72%

A Hybrid Knowledge-Grounded Framework for Safety and Traceability in Prescription Verification

一种混合知识引导的框架用于处方验证中的安全性和可追溯性

Yichi Zhu, Kan Ling, Xu Liu, Hengrun Zhang, Huiqun Yu, Guisheng Fan

机构 * School of Information Science and Engineering, East China University of Science and Technology(信息科学与工程学院,东华大学)

专题命中 逻辑推理 :reasoning(abstract,comments);logical reasoning(abstract);分类 cs.AI

AI总结 本文提出PharmGraph-Auditor框架,通过混合制药知识库和知识引导的验证链,提升处方验证的安全性和可追溯性。

Comments 11 pages, 7 figures.Framework for safe prescription auditing and hybrid knowledge-grounded reasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19749 2025-07-29 cs.AI 72%

Can LLMs Solve ASP Problems? Insights from a Benchmarking Study (Extended Version)

Lin Ren, Guohui Xiao, Guilin Qi, Yishuai Geng, Haohan Xue

机构 * School of Computer Science and Engineering, Southeast University, Nanjing, China(计算机科学与工程学院,东南大学,南京,中国) Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及其跨学科应用关键实验室(东南大学),教育部,中国)

专题命中 逻辑推理 :reasoning(abstract,comments);logical reasoning(abstract);分类 cs.AI

Comments Accepted for publication at the 22nd International Conference on Principles of Knowledge Representation and Reasoning (KR 2025). The code is available at https://github.com/HomuraT/ASPBench

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00878 2026-08-04 cs.IT math.IT 新提交 71%

Goal-Oriented Logic-based Semantic Communication for Neuro-Symbolic Reasoning with Applications onto Autonomous Driving

面向目标的基于逻辑的语义通信及用于自动驾驶的神经符号推理

Ahmet Faruk Saz, Duo Xu, Faramarz Fekri

专题命中 逻辑推理 :reasoning(title)

AI总结 本文针对自动驾驶场景,提出基于一阶逻辑的目标导向语义通信方法,通过选择关键证据传输实现安全决策,在相同通信预算下优于均匀证据选择。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10880 2026-08-04 cs.CV 版本更新 71%

Chart Specification: Structural Representations for Incentivizing VLM Reasoning in Chart-to-Code Generation

图表规范:用于促进VLM在图表到代码生成中推理的结构表示

Minggui He, Mingchen Dai, Jian Zhang, Yilun Liu, Shimin Tao, Pufan Zeng, Osamu Yoshie, Yuya Ieiri

机构 * Waseda University(早稻田大学) University of Science and Technology of China(中国科学技术大学) NanKai University(南开大学)

专题命中 逻辑推理 :reasoning(title)

AI总结 本文提出图表规范,通过结构化中间表示提升VLM在图表到代码生成中的结构保真度,实验显示其在数据效率和性能上优于现有方法。

Comments Accepted by Neurocomputing

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.06724 2026-07-09 cs.RO 新提交 71%

EvoPlan: Evolutionary Neuro-Symbolic Robot Planning with Spatio-Temporal Guarantees

EvoPlan:具有时空保证的进化神经符号机器人规划

Bhavya Sai Nukapotula, Samin Moosavi, Haoze Wang, Luke Duncan, Diya Shakkottai, Varun Murali, Srinivas Shakkottai

机构 * Texas A&M University(德克萨斯农工大学)

专题命中 逻辑推理 :planning(title)

AI总结 研究针对基于LLM的机器人规划器不足及经典PDDL规划器问题,提出含离线STL约束挖掘、进化PDDL规划器和受约束执行循环的神经符号框架,能保证计划属性,在基准测试中表现良好且可在无云环境下部署。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03963 2026-07-07 cs.SE 新提交 71%

Neuro-Symbolic Reasoning for Vulnerability Detection

用于漏洞检测的神经符号推理

Yanjie Zhao, Hongjie Chen, Li Lu, Zhou Yang, Xiao Cheng, Haoyu Wang

专题命中 逻辑推理 :reasoning(title)

AI总结 研究基于大语言模型的漏洞检测不可靠性,提出神经符号框架LeanGuard,通过将代码解释与安全义务判定分离,神经侧过滤事实,符号侧形式验证,证据感知裁决器权衡结果,在五个CWE类上实例化框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.06888 2026-06-04 eess.SY cs.SY 71%

On the Timed Temporal Logic Planning of Coupled Multi-Agent Systems

关于耦合多智能体系统的时序时序逻辑规划

Alexandros Nikou, Dimitris Boskos, Jana Tumova, Dimos V. Dimarogonas

专题命中 逻辑推理 :planning(title)

AI总结 本文提出一种自动控制器合成方法,用于满足耦合约束的多智能体系统。通过设计分布式抽象和形式验证技术,计算满足高阶任务的个体运行。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26053 2026-04-30 cs.LO cs.MA 71%

I Would If I Could: Reasoning about Dynamics of Actions in Multi-Agent Systems

若我能的话:多智能体系统中行动动态的推理

Rustam Galimullin, Hermine Grosinger, Munyque Mittelmann

专题命中 逻辑推理 :reasoning(title)

AI总结 本文提出ATL-D和ATEL-D,用于建模智能体行动的动态授予与撤销过程及其对知识的影响,探讨了逻辑表达力、规范系统关系及计算复杂性。

Comments This is an extended version of the paper with the same title that will appear in KR 2026, and which contains a technical appendix with proof details

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14051 2026-04-16 cs.IR 71%

Enhancing Local Life Service Recommendation with Agentic Reasoning in Large Language Model

通过大语言模型中的代理推理增强本地生活服务推荐

Shiteng Cao, Xiaochong Lan, Yuwei Du, Jie Feng, Yinxing Liu, Xinlei Shi, Yong Li

专题命中 逻辑推理 :reasoning(title)

AI总结 本文提出一种联合预测生活需求和服务推荐的框架,通过行为聚类和强化学习提升推荐准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15336 2026-02-18 cs.AR 71%

Human-AI Interaction: Evaluating LLM Reasoning on Digital Logic Circuit included Graph Problems, in terms of creativity in design and analysis

人类-人工智能交互:评估LLM在包含图问题的数字逻辑电路中的推理能力,就设计和分析的创造性而言

Yogeswar Reddy Thota, Setareh Rafatirad, Homayoun Houman, Tooraj Nikoubin

专题命中 逻辑推理 :reasoning(title)

AI总结 评估LLM在数字逻辑电路问题中的推理能力,发现其在复杂问题上与官方答案存在差距,且易陷入教科书模板。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12888 2025-12-16 physics.optics cs.AI cs.CL cs.LG 71%

Meta-GPT: Decoding the Metasurface Genome with Generative Artificial Intelligence

Meta-GPT:用生成式人工智能解码元表面基因组

David Dang, Stuart Love, Meena Salib, Quynh Dang, Samuel Rothfarb, Mysk Alnatour, Andrew Salij, Hou-Tong Chen, Ho Wai, Lee, Wilton J. M. Kort-Kamp

专题命中 逻辑推理 :chain-of-thought(abstract,comments);分类 cs.CL、cs.AI、cs.LG;reasoning(comments)

AI总结 Meta-GPT通过METASTRINGS符号语言实现光子学设计,以生成式人工智能解码元表面基因组。

Comments Keywords: Physics-informed machine learning; Transformer models; Reinforcement learning; Chain-of-thought reasoning; Metasurfaces; Nanophotonics; Inverse design

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18438 2025-10-27 cs.CR 71%

DeepTx: Real-Time Transaction Risk Analysis via Multi-Modal Features and LLM Reasoning

Yixuan Liu, Xinlei Li, Yi Li

专题命中 逻辑推理 :reasoning(title)

Comments Accepted to ASE'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05475 2025-10-08 q-fin.CP quant-ph 71%

From Classical Rationality to Contextual Reasoning: Quantum Logic as a New Frontier for Human-Centric AI in Finance

Fabio Bagarello, Francesco Gargano, Polina Khrennikova

专题命中 逻辑推理 :reasoning(title)

Comments 19 pages, 5 figures, preprint version. Forthcoming in: Journal of Quantum Economics and Finance

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.04215 2024-12-17 cs.RO 71%

Comp-LTL: Temporal Logic Planning via Zero-Shot Policy Composition

Taylor Bergeron, Zachary Serlin, Kevin Leahy

专题命中 逻辑推理 :planning(title)

Comments 16 pages, 11 figures. Updated to reflect additional results

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09214 2024-02-13 cs.MA 71%

Logic of Awareness in Agent's Reasoning

Yudai Kubono, Teeradaj Racharak, Satoshi Tojo

专题命中 逻辑推理 :reasoning(title)

Comments A version reflecting the corrections of errors. The results are the same as originally stated. The erratum to arXiv:2309.09214v2 is attached at the end

Journal ref Proceedings of the 15th International Conference on Agents and Artificial Intelligence - Volume 1: ICAART, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.08714 2023-01-24 cs.SE cs.FL cs.MA 71%

Verse: A Python library for reasoning about multi-agent hybrid system scenarios

Yangge Li, Haoqing Zhu, Katherine Braught, Keyi Shen, Sayan Mitra

专题命中 逻辑推理 :reasoning(title)

Comments 26 pages, 16 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.03322 2016-11-11 cs.RO cs.LO 71%

Verification of Logical Consistency in Robotic Reasoning

Hongyang Qu, Sandor M. Veres

专题命中 逻辑推理 :reasoning(title)

Journal ref Robotics and Autonomous Systems, Vol. 83(2016), 44-56

详情

展开后加载摘要…

URL PDF HTML 收藏
1208.3461 2015-06-11 cs.SE cs.LO 71%

Modeling and Verification of Agent based Adaptive Traffic Signal using Symbolic Model Verifier

Vivek Vishal, Sagar Gugwad, Sanjay Singh

专题命中 逻辑推理 :verifier(title)

Comments 13 pages, 6 figures, Submitted to International Journal of Computer Application (IJCA)

详情

展开后加载摘要…

URL PDF HTML 收藏
1402.1377 2014-04-15 cs.LO 71%

Reasoning about Games via a First-order Modal Model Checking Approach

Davi Romero de Vasconcelos, Edward Hermann Haeusler

专题命中 逻辑推理 :reasoning(title)

Comments Extended version of article published in the SBMF 2007. Accepted to ENTCS. Withdrawn from ENTCS in 2014 in virtue to submission to other venue

详情

展开后加载摘要…

URL PDF HTML 收藏
1210.1630 2012-10-08 cs.RO cs.GT 71%

Symbolic Planning and Control Using Game Theory and Grammatical Inference

Jie Fu, Herbert G. Tanner, Jeffrey Heinz, Jane Chandlee, Konstantinos Karydis, Cesar Koirala

专题命中 逻辑推理 :planning(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
1109.0601 2011-09-06 cs.MA 71%

Application of distributed constraint satisfaction problem to the agent-based planning in manufacturing systems

S. Kornienko, O. Kornienko, P. Levi

专题命中 逻辑推理 :planning(title)

Journal ref Proceedings of the International Scientific Congress "Intelligent Systems (IEEE AIS'03)" and "Intelligent CAD's (CAD-2003)", p.124-140, Divnomorsk, Russia, 2003

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.17341 2026-08-19 cs.AI 新提交 70%

LLM-Only PDDL Domain Repair with Open-Weight Models

仅使用大语言模型的PDDL领域修复:基于开放权重模型

Nader Karimi Bavandpour, Pascal Bercher

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

AI总结 本文评估开放权重大语言模型仅用LLM方法修复PDDL模型错误的能力,发现其F1分数优于符号基线,但测试通过率仍不足,无法保证满足修复所需的测试约束。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.14579 2026-08-18 cs.AI 新提交 70%

SKILL: Self-correcting Knowledge-guided Iterative Large Language Model Agent for Logic Optimization

SKILL:用于逻辑优化的自校正知识引导迭代大语言模型智能体

Rui Yang

机构 * University of California, Riverside(加州大学河滨分校)

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

AI总结 针对逻辑综合优化的挑战,提出SKILL智能体,结合多LLM推理与RL交互,经基准测试实现12.4%的PDA提升及86.3%的50万门级逻辑系统成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00143 2026-08-11 cs.CR cs.AI 版本更新 70%

Symbolic Attack Chain Generation from Atomic Red Team Techniques: An Empirical Study of Predicate Representation Granularity

基于Atomic Red Team技术的符号化攻击链生成:谓词表示粒度的实证研究

Ramya Varunsegar

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

AI总结 本研究对比九类与五类谓词粒度的AALM,发现粒度对攻击链有效性影响小,81.3%结果一致,高粒度仅提升规划论证的内部结构分辨率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26903 2026-07-30 cs.AI cs.RO 新提交 70%

From Passive Video to Editable Experience: Physically Grounded Experience Synthesis for Embodied Intelligence

从被动视频到可编辑体验:具身智能的物理基础体验合成

Jia Luo

专题命中 逻辑推理 :planning(abstract);verifier(abstract);分类 cs.AI

AI总结 针对具身AI的数据瓶颈,提出Pegasus低资源框架,通过结构化知识传递将人类操作视频转化为机器人可学习数据,经多基准与机器人评估验证其跨具身翻译及数据生成的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.17849 2026-07-21 cs.HC cs.CL cs.CV 新提交 70%

AlphaOracle: Oracle bone script decipherment via human-workflow-inspired deep learning

AlphaOracle:通过受人工工作流程启发的深度学习进行甲骨文破译

Yuliang Liu, Haisu Guan, Pengjie Wang, Xinyu Wang, Jinpeng Wan, Kaile Zhang, Handong Zheng, Xingchen Liu, Zhebin Kuang, Huanxin Yang, Bang Li, Yongge Liu, Lianwen Jin, Xiang Bai

机构 * School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件学院) School of Electronic Information Engineering, South China University of Technology(华南理工大学电子信息学院) Key Laboratory of Oracle Bone Inscriptions Information Processing, Anyang Normal University(安阳师范学院甲骨文信息处理重点实验室)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.CL

AI总结 针对约3000个未破译甲骨文字符的问题,提出受人工工作流程启发的AlphaOracle框架,经多阶段管道进行甲骨文破译,与专家解读高度一致,还能减少分析时间,为甲骨文及其他未破译文字研究提供参考。

Comments Accepted by The Innovation 2026

Journal ref The Innovation 7(11), 101462, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.15647 2026-07-20 cs.AI 新提交 70%

Neuro-Symbolic AI for LEED compliance: Document-Centric Benchmarking, Deterministic Numeric Checking, and When Multimodal Hurts

用于LEED合规的神经符号人工智能:以文档为中心的基准测试、确定性数值检查以及多模态何时有害

Aritro De, Juliana Felkner

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 逻辑推理 :chain-of-thought(abstract);verifier(abstract);分类 cs.AI

AI总结 研究小型本地部署语言模型对LEED文档筛选及符号组件作用,引入神经符号管道,通过对齐PDF、检索证据、语言模型验证和数值检查器检查来验证合规性,实验表明特定模型在任务中表现较好,管道及基线提供了参考。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05294 2026-07-08 cs.CL 版本更新 70%

Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations

真实还是编造?利用因果归因减轻解释中的奖励作弊

Pedro Ferreira, Wilker Aziz, Ivan Titov

机构 * Institute for Logic, Language and Computation (ILLC), University of Amsterdam(逻辑、语言与计算研究所(ILLC),阿姆斯特丹大学) Institute for Language, Cognition and Computation (ILCC), University of Edinburgh(语言、认知与计算研究所(ILCC),爱丁堡大学)

专题命中 逻辑推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL

AI总结 研究大语言模型解释中因奖励模型导致的奖励作弊问题,提出用预测的因果归因丰富奖励模型输入的方法,该方法能减少大语言模型生成误导性解释的倾向,提升解释忠实度。

Comments ICLR 2026 Camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.11689 2026-07-02 cs.AI 版本更新 70%

Explicit Logic Channel for Validation and Enhancement of MLLMs on Zero-Shot Tasks

显式逻辑通道用于验证和增强用于零样本任务的前沿多模态大语言模型

Mei Chee Leong, Ying Gu, Hui Li Tan, Liyuan Li, Nancy Chen

机构 * Institute for Infocomm Research (I$^2$R)(信息通信研究所) Agency for Science, Technology and Research (A*STAR)(科技研究局) Singapore(新加坡)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

AI总结 本文提出显式逻辑通道用于验证和增强多模态大语言模型在零样本任务中的性能,通过显式逻辑推理提高模型的可解释性和可信度。

Comments Accepted to ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.31229 2026-07-01 cs.AI 新提交 70%

Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents

Agentic-Ideation: 面向科学构思智能体的样本高效智能体轨迹合成

Keyu Zhao, Lingyan Kong, Fengli Xu, Yong Li

机构 * Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

AI总结 提出Agentic-Ideation框架,通过Oracle引导的数据合成策略和掩码训练,高效生成高质量智能体轨迹,在科学构思任务上整体质量提升11.91%,数据合成效率提高10倍以上。

详情

展开后加载摘要…

URL PDF HTML 收藏