arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

共收录 45259 信号源:cs.CL, cs.AI, cs.LG

1. 逻辑推理 3043 篇

2511.20540 2025-11-26 cs.LO cs.AI cs.GT cs.MA 57%

Proceedings Twentieth Conference on Theoretical Aspects of Rationality and Knowledge

第二十届理性与知识理论研讨会会议纪要

Adam Bjorndahl

机构 * Edited by: Adam Bjorndahl(编辑:Adam Bjorndahl)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

AI总结 TARK 2025会议收录了关于理性与知识理论的最新研究成果,涵盖知识、信念、博弈论等多个领域。

Journal ref EPTCS 437, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18718 2025-11-25 cs.RO cs.AI 57%

AIRHILT: A Human-in-the-Loop Testbed for Multimodal Conflict Detection in Aviation

AIRHILT:航空多模态冲突检测的人机交互测试平台

Omar Garib, Jayaprakash D. Kambhampaty, Olivia J. Pinon Fischer, Dimitri N. Mavris

机构 * Daniel Guggenheim School of Aerospace Engineering, Georgia Institute of Technology(丹尼尔·古根海姆航空航天工程学院,佐治亚理工学院)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

AI总结 AIRHILT是一款用于航空多模态冲突检测的人机交互测试平台,通过集成ASR、视觉检测和决策模型,实现冲突检测的高效评估与研究。

Comments 9 pages, 4 figures, 1 table, 1 algorithm

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18604 2025-11-25 cs.RO cs.AI cs.MA 57%

An Analysis of Constraint-Based Multi-Agent Pathfinding Algorithms

基于约束的多智能体路径寻找算法分析

Hannah Lee, James D. Motes, Marco Morales, Nancy M. Amato

机构 * Parasol Lab, School of Computer Science, University of Illinois at Urbana Champaign(帕索尔实验室,计算机科学学院,伊利诺伊大学厄巴纳-香槟分校) Department of Computer Science at Instituto Tecnológico Autónomo de México (ITAM)(墨西哥自治理工学院(ITAM)计算机科学系)

专题命中 逻辑推理 :planning(abstract);分类 cs.AI

AI总结 本文分析了基于约束的多智能体路径寻找算法,探讨了保守型与激进型约束在不同场景下的性能差异,并提供了决策流程图以指导约束选择,同时强调了拓扑特征在多机器人运动规划中的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17644 2025-11-25 cs.AI 57%

Hybrid Neuro-Symbolic Models for Ethical AI in Risk-Sensitive Domains

混合神经符号模型用于风险敏感领域的伦理AI

Chaitanya Kumar Kolli

机构 * Independent Researcher USA Email(独立研究者)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出混合神经符号模型,用于在风险敏感领域实现伦理AI,通过结合神经网络与符号推理,提升AI的可解释性和合规性。

Comments 6 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21651 2025-11-24 cs.AI 57%

Can AI Perceive Physical Danger and Intervene?

AI能否感知物理危险并干预?

Abhishek Jindal, Dmitry Kalashnikov, R. Alex Hofer, Oscar Chang, Divya Garikapati, Anirudha Majumdar, Pierre Sermanet, Vikas Sindhwani

机构 * Google DeepMind Robotics(谷歌深Mind机器人技术)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出了一种用于评估具身体验AI系统物理安全性的基准测试方法,通过生成逼真图像和视频来测试模型对安全风险的理解和干预能力,并开发了训练后范式以提升模型的安全推理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14098 2025-11-19 cs.AI cs.MA cs.SI cs.SY eess.SY 57%

Collaborative QA using Interacting LLMs. Impact of Network Structure, Node Capability and Distributed Data

Adit Jain, Vikram Krishnamurthy, Yiming Zhang

机构 * Department of Electrical and Computer Engineering, Cornell University(电气与计算机工程系,康奈尔大学)

专题命中 逻辑推理 :test-time compute(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12916 2025-11-18 cs.AI 57%

Fault2Flow: An AlphaEvolve-Optimized Human-in-the-Loop Multi-Agent System for Fault-to-Workflow Automation

Yafang Wang, Yangjie Tian, Xiaoyu Shen, Gaoyang Zhang, Jiaze Sun, He Zhang, Ruohua Xu, Feng Zhao

机构 * ISILC, Victoria University(维多利亚大学ISILC) Kexin Melbourne AI Research Center(墨尔本凯欣人工智能研究中心) Eastern Institute of Technology(东部技术研究所)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11954 2025-11-18 cs.AI 57%

LLM-Assisted Formalization Enables Deterministic Detection of Statutory Inconsistency in the Internal Revenue Code

Borchuluun Yadamsuren, Steven Keith Platt, Miguel Diaz

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

Comments 29 pages, 3 appendices with Prolog code and full codebase available at: https://github.com/borchuluun/section121-inconsistency-detection

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11600 2025-11-18 cs.AI cs.IR 57%

CausalGuard: A Smart System for Detecting and Preventing False Information in Large Language Models

Piyushkumar Patel

机构 * Microsoft(微软)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08715 2025-11-13 cs.AI 57%

Bridging Natural Language and ASP: A Hybrid Approach Using LLMs and AMR Parsing

Connar Hite, Sean Saud, Raef Taha, Nayim Rahman, Tanvir Atahary, Scott Douglass, Tarek Taha

机构 * United States Air Force(美国空军)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07584 2025-11-12 cs.SE cs.AI cs.DC 57%

SemanticForge: Repository-Level Code Generation through Semantic Knowledge Graphs and Constraint Satisfaction

Wuyang Zhang, Chenkai Zhang, Zhen Luo, Jianming Ma, Wangming Yuan, Chuqiao Gu, Chenwei Feng

机构 * Department of Elec.&Comp. Science, University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校电子与计算机科学系) Department of Computer Sys. Engineering, Northeastern University(东北大学计算机系统工程系) Department of Computer Science, George Mason University(乔治·梅森大学计算机科学系) Department of Info. Networking Institude, Carnegie Mellon University(卡内基梅隆大学信息网络研究所) Department of Computer & Mathematical Sciences, Auckland University of Technology(奥克兰理工大学计算机与数学科学系)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

Journal ref INNO-PRESS: Journal of Emerging Applied AI, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06260 2025-11-11 cs.GT cs.AI cs.SY eess.SY 57%

LLM-Guided Reinforcement Learning with Representative Agents for Traffic Modeling

Hanlin Sun, Jiayang Li

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06175 2025-11-11 cs.AI cs.GT 57%

CSP4SDG: Constraint and Information-Theory Based Role Identification in Social Deduction Games with LLM-Enhanced Inference

Kaijie Xu, Fandi Meng, Clark Verbrugge, Simon Lucas

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26603 2025-10-31 cs.AI cs.MA cs.SY eess.SY 57%

Agentic AI Home Energy Management System: A Large Language Model Framework for Residential Load Scheduling

Reda El Makroum, Sebastian Zwickl-Bernhard, Lukas Kranzl

机构 * Department of Industrial Economics and Technology Management, The Norwegian University of Science and Technology(工业经济学与技术管理系,挪威科学与技术大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

Comments 34 pages, 9 figures. Code available at https://github.com/RedaElMakroum/agentic-ai-hems

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26353 2025-10-31 cs.LG 57%

Towards Explainable and Reliable AI in Finance

Albi Isufaj, Pablo Mollá, Helmut Prendinger

机构 * National Institute of Informatics Graduate University for Advanced Studies(日本信息处理研究所高级研究大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15356 2025-10-30 cs.CL 57%

NL-Debugging: Exploiting Natural Language as an Intermediate Representation for Code Debugging

Weiming Zhang, Qingyao Li, Xinyi Dai, Jizheng Chen, Kounianhua Du, Weiwen Liu, Yasheng Wang, Ruiming Tang, Yong Yu, Weinan Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Huawei Noah’s Ark Lab Shanghai(华为诺亚实验室)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24051 2025-10-29 cs.CL 57%

Pie: A Programmable Serving System for Emerging LLM Applications

In Gim, Zhiyao Ma, Seung-seob Lee, Lin Zhong

机构 * Yale University(耶鲁大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.CL

Comments SOSP 2025. Source code available at https://github.com/pie-project/pie

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08216 2025-10-28 cs.AI 57%

Grounding Methods for Neural-Symbolic AI

Rodrigo Castellano Ontiveros, Francesco Giannini, Marco Gori, Giuseppe Marra, Michelangelo Diligenti

机构 * University of Siena(锡耶纳大学) Scuola Normale Superiore(正规大学) KU Leuven(卢森堡大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

Journal ref Proceedings of the Thirty-Fourth International Joint Conference on Artificial Intelligence (IJCAI-25), pp. 4806-4814, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20345 2025-10-24 cs.AI 57%

LLM-empowered knowledge graph construction: A survey

Haonan Bian

机构 * Xidian University(西安电子科技大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19263 2025-10-23 cs.AI 57%

An Argumentative Explanation Framework for Generalized Reason Model with Inconsistent Precedents

Wachara Fungwacharakorn, Gauvain Bourgne, Ken Satoh

机构 * Center for Juris-Informatics, ROIS-DS, Tokyo, Japan(法律信息中心,ROIS-DS,东京,日本) LIP6, Sorbonne University, CNRS, Paris, France(LIP6,索邦大学,CNRS,巴黎,法国)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

Comments 10 pages, extended version for JURIX 2025 submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17002 2025-10-21 cs.LG 57%

EEschematic: Multimodal-LLM Based AI Agent for Schematic Generation of Analog Circuit

Chang Liu, Danial Chitnis

机构 * School of Engineering The University of Edinburgh Edinburgh, UK(工程学院 苏格兰爱丁堡大学)

专题命中 逻辑推理 :chain-of-thought(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14676 2025-10-17 cs.AI 57%

NAEL: Non-Anthropocentric Ethical Logic

Bianca Maria Lerma, Rafael Peñaloza

机构 * University of Milano-Bicocca(米兰-比科卡大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

Comments Accepted to the FEAR workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.10742 2025-10-15 cs.AI 57%

The Philosophical Foundations of Growing AI Like A Child

Dezhi Luo, Yijiang Li, Hokin Deng

机构 * University of Michigan(密歇根大学) University of California San Diego(加州大学圣地亚哥分校) Carnegie Mellon University(卡内基梅隆大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12287 2025-10-15 cs.CV cs.CL 57%

Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector

Sifan Li, Hongkai Chen, Yujun Cai, Qingwen Ye, Liyang Chen, Junsong Yuan, Yiwei Wang

机构 * University of California, Merced(加州大学梅尔德分校) vivo Mobile Communication Co., Ltd.(vivo移动通信有限公司) University of Queensland(昆士兰大学) UCLA(加州大学洛杉矶分校) University at Buffalo(布法罗大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11389 2025-10-14 cs.CL 57%

Beyond Survival: Evaluating LLMs in Social Deduction Games with Human-Aligned Strategies

Zirui Song, Yuan Huang, Junchang Liu, Haozhe Luo, Chenxi Wang, Lang Gao, Zixiang Xu, Mingfei Han, Xiaojun Chang, Xiuying Chen

机构 * Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.CL

Comments 34 pages, 32figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11281 2025-10-14 cs.AI 57%

PADME: Procedure Aware DynaMic Execution

Deepeka Garg, Sihan Zeng, Annapoorani L. Narayanan, Sumitra Ganesh, Leo Ardon

机构 * J.P. Morgan AI Research(摩根大通人工智能研究)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10517 2025-10-14 cs.PL cs.AI cs.SE 57%

ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs

Su-Hyeon Kim, Joonghyuk Hahn, Sooyoung Cha, Yo-Sub Han

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18233 2025-10-14 cs.AI 57%

Beyond Parameters: Exploring Virtual Logic Depth for Scaling Laws

Ruike Zhu, Hanwen Zhang, Kevin Li, Tianyu Shi, Yiqun Duan, Chi Wang, Tianyi Zhou, Arindam Banerjee, Zengyi Qin

机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) University of Toronto(多伦多大学) University of Technology Sydney(悉尼技术大学) Google DeepMind(谷歌DeepMind) University of Maryland, College Park(马里兰大学学院市分校) Massachusetts Institute of Technology(麻省理工学院)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09970 2025-10-14 cs.AI 57%

Follow My Lead: Logical Fallacy Classification with Knowledge-Augmented LLMs

Olivia Peiyu Wang, Tashvi Bansal, Ryan Bai, Emily M. Chui, Leilani H. Gilpin

机构 * Monta Vista High School(蒙塔维斯高中) Canyon Crest Academy(卡耶恩峡谷学院) Durham Academy Upper School(达灵顿学院上校)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

Comments Accepted as a poster at the Twelfth Annual Conference on Advances in Cognitive Systems. 21 pages, 7 figures and 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08175 2025-10-10 cs.AI 57%

Prepared mind, fast response: A temporal decoupling framework for adaptive knowledge orchestration in open-domain dialogue

Jinling Gan, Churong Liang, Runnan Li

机构 * Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 逻辑推理 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏