arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

共收录 3046 信号源:cs.CL, cs.AI, cs.LG

1. 逻辑推理 3046 篇

2511.18964 2025-11-25 cs.AI 70%

Synthesizing Visual Concepts as Vision-Language Programs

合成视觉概念作为视觉-语言程序

Antonia Wüst, Wolfgang Stammer, Hikaru Shindo, Lukas Helff, Devendra Singh Dhami, Kristian Kersting

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

AI总结 本研究提出视觉-语言程序(VLP),结合视觉模型的感知灵活性与程序合成的系统推理能力,通过生成结构化视觉描述并编译成神经符号程序,提升复杂逻辑推理任务的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10975 2025-11-13 cs.IR cs.CL 70%

ReFineG: Synergizing Small Supervised Models and LLMs for Low-Resource Grounded Multimodal NER

Jielong Tang, Shuang Wang, Zhenxing Wang, Jianxing Yu, Jian Yin

机构 * School of Artificial Intelligence, Sun Yat-sen University(中山大学人工智能学院) Key Laboratory of Sustainable Tourism Smart Assessment Technology, Ministry of Culture and Tourism, Sun Yat-sen University(文化旅游可持续评估技术重点实验室,中华人民共和国文化和旅游部,中山大学) Beijing Normal University(北京师范大学) Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.CL

Comments CCKS 2025 Shared Task Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.15153 2025-11-13 cs.CL 70%

Evaluating Deep Unlearning in Large Language Models

Ruihan Wu, Chhavi Yadav, Russ Salakhutdinov, Kamalika Chaudhuri

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00758 2025-11-04 cs.AI 70%

Active Thinking Model: A Goal-Directed Self-Improving Framework for Real-World Adaptive Intelligence

Hong Su

机构 * School of Computer Science, Chengdu University of Information Technology(信息工程大学计算机学院)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23127 2025-10-31 cs.AI 70%

Lost in Tokenization: Context as the Key to Unlocking Biomolecular Understanding in Scientific LLMs

Kai Zhuang, Jiawei Zhang, Yumou Liu, Hanqun Cao, Chunbin Gu, Mengdi Liu, Zhangyang Gao, Zitong Jerry Wang, Xuanhe Zhou, Pheng-Ann Heng, Lijun Wu, Conghui He, Cheng Tan

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Westlake University(西交大学) Shanghai Innovation Institute(上海创新研究院) Shanghai Jiaotong University(上海交通大学) The Chinese University of Hong Kong(香港中文大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

Comments 38 pages, under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16720 2025-10-28 cs.AI 70%

Beyond Pipelines: A Survey of the Paradigm Shift toward Model-Native Agentic AI

Jitao Sang, Jinlin Xiao, Jiarun Han, Jilin Chen, Xiaoyi Chen, Shuyu Wei, Yongjie Sun, Yuhang Wang

机构 * Institute for Clarity in Documentation(清晰文档研究所) Inria Paris-Rocquencourt(巴黎- Rocquencourt 国家信息与自动化所) Rajiv Gandhi University(拉吉夫·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕勒尔研究实验室) Beijing Jiaotong University(北京交通大学)

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19493 2025-10-23 cs.CL 70%

What is the Best Sequence Length for BABYLM?

Suchir Salhan, Richard Diehl Martinez, Zébulon Goriely, Paula Buttery

机构 * Department of Computer Science & Technology, University of Cambridge, U.K.(计算机科学与技术系,剑桥大学,英国) ALTA Institute, University of Cambridge, U.K.(ALTA研究所,剑桥大学,英国)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.CL

Comments Paper Accepted at the 2025 BabyLM Workshop @ EMNLP (Suzhou, China)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07581 2025-10-20 cs.LG 70%

Expanding the Action Space of LLMs to Reason Beyond Language

Zhongqi Yue, Weishi Wang, Yundaichuan Zhan, Juncheng Li, Daniel Dahlmeier, Fredrik D. Johansson

机构 * Chalmers University of Technology and University of Gothenburg(查尔姆斯理工大学和哥德堡大学) SAP(SAP公司) Zhejiang University(浙江大学)

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10701 2025-10-14 cs.AI cs.LO 70%

Extended Triangular Method: A Generalized Algorithm for Contradiction Separation Based Automated Deduction

Yang Xu, Shuwei Chen, Jun Liu, Feng Cao, Xingxing He

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

Comments 38 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08086 2025-10-10 cs.AI 70%

From Ethical Declarations to Provable Independence: An Ontology-Driven Optimal-Transport Framework for Certifiably Fair AI Systems

Sukriti Bhattacharya, Chitro Majumdar

机构 * Senior Scientist, Trustworthy AI, Luxembourg Institute of Science \& Technology, Maison de l'innovation 5, L-4362, Luxembourg Chief Investment Risk Strategist for Sovereign Institutions Founder, RsRL, Jumeirah Beach Residence (JBR), P.O.Box 29215, Dubai, United Arab Emirates

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

Comments 19 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06078 2025-10-08 cs.AI 70%

Constraint-Aware Route Recommendation from Natural Language via Hierarchical LLM Agents

Tao Zhe, Rui Liu, Fateme Memar, Xiao Luo, Wei Fan, Xinyue Ye, Zhongren Peng, Dongjie Wang

机构 * University of Kansas(堪萨斯大学) UW–Madison(威斯康星大学麦迪逊分校) University of Auckland(奥克兰大学) University of Alabama(阿拉巴马大学) University of Florida(佛罗里达大学)

专题命中 逻辑推理 :reasoning(abstract);verifier(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00300 2025-10-02 cs.AI 70%

ICL Optimized Fragility

Serena Gomez Wannaz

机构 * Serena Gomez Wannaz(独立研究者)

专题命中 逻辑推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.01872 2025-10-01 cs.CL 70%

Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions

Rachneet Sachdeva, Rima Hazra, Iryna Gurevych

机构 * Ubiquitous Knowledge Processing Lab (UKP Lab), Department of Computer Science and Hessian Center for AI (hessian.AI), Technical University of Darmstadt(德累斯顿技术大学计算机科学系、海斯堡人工智能中心(hessian.AI)、通用知识处理实验室(UKP Lab))

专题命中 逻辑推理 :reasoning(abstract);CoT(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025 (Main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21874 2025-09-29 cs.LG 70%

Abductive Logical Rule Induction by Bridging Inductive Logic Programming and Multimodal Large Language Models

Yifei Peng, Yaoli Liu, Enbo Xia, Yu Jin, Wang-Zhou Dai, Zhong Ren, Yao-Xiang Ding, Kun Zhou

机构 * State Key Laboratory of CAD&CG(计算机辅助设计与图形学国家重点实验室) National Key Laboratory for Novel Software Technology(新型软件技术国家实验室)

专题命中 逻辑推理 :reasoning(abstract);CoT(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14359 2025-09-25 cs.CL 70%

Triangulating LLM Progress through Benchmarks, Games, and Cognitive Tests

Filippo Momentè, Alessandro Suglia, Mario Giulianelli, Ambra Ferrari, Alexander Koller, Oliver Lemon, David Schlangen, Raquel Fernández, Raffaella Bernardi

机构 * University of Trento(特伦托大学) ILCC University of Edinburgh(爱丁堡大学ILCC学院) UCL(伦敦大学学院) Saarland University(萨尔兰大学) Heriot-Watt University(赫瑞-瓦特大学) University of Potsdam(波茨坦大学) DFKI(德国达姆施塔特研究所) University of Amsterdam(阿姆斯特丹大学) Faculty of Engineering, Free University of Bozen-Bolzano(博兹纳-博尔扎诺自由大学工程学院)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13932 2025-09-23 cs.AI cs.MA 70%

XAgents: A Framework for Interpretable Rule-Based Multi-Agents Cooperation

Hailong Yang, Mingxian Gu, Renhuo Zhao, Fuping Hu, Zhaohong Deng, Yitang Chen

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

Comments We intend to substantially revise the problem statement and scope; therefore we withdraw the current version

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15635 2025-09-22 cs.AI 70%

MicroRCA-Agent: Microservice Root Cause Analysis Method Based on Large Language Model Agents

Pan Tang, Shixiang Tang, Huanqi Pu, Zhiqing Miao, Zhixing Wang

机构 * School of Communication and Information Engineering, Shanghai University, Shanghai, China(上海大学通信与信息工程学院) School of Communication and Electronic Engineering, East China Normal University, Shanghai, China(华东师范大学通信与电子工程学院) School of Information and Electronics, Beijing Institute of Technology, Beijing, China(北京理工大学信息与电子学院)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

Comments 18 pages, 22 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07665 2025-09-10 cs.AI 70%

DeepGraphLog for Layered Neurosymbolic AI

Adem Kikaj, Giuseppe Marra, Floris Geerts, Robin Manhaeve, Luc De Raedt

机构 * Örebro University, Sweden(奥雷布罗大学) University of Antwerp, Belgium(安特卫普大学)

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05764 2025-09-09 cs.AI 70%

DRF: LLM-AGENT Dynamic Reputation Filtering Framework

Yuwei Lou, Hao Hu, Shaocong Ma, Zongfei Zhang, Liang Wang, Jidong Ge, Xianping Tao

机构 * State Key Laboratory for Novel Software Technology(新型软件技术国家重点实验室) Nanjing University(南京大学) Amazon.com Services LLC(亚马逊公司)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

Comments This paper has been accepted by ICONIP 2025 but not published

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05199 2025-09-08 cs.CL 70%

Triadic Fusion of Cognitive, Functional, and Causal Dimensions for Explainable LLMs: The TAXAL Framework

David Herrera-Poyatos, Carlos Peláez-González, Cristina Zuheros, Virilo Tejedor, Rosana Montes, Francisco Herrera

机构 * Department of Computer Science and Artificial Intelligence(计算机科学与人工智能系) Andalusian Institute of Data Science and Computational Intelligence(安达卢西亚数据科学与计算智能研究所) University of Granada(格拉纳达大学)

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.CL

Comments 27 pages, 9 tables and 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08613 2025-09-08 cs.CL 70%

Assessing the Sensitivity and Alignment of FOL Closeness Metrics

Ramya Keerthy Thatikonda, Wray Buntine, Ehsan Shareghi

机构 * Department of Data Science & AI, Monash University(数据科学与人工智能系,莫纳什大学) College of Engineering and Computer Science, VinUniversity(工程与计算机科学学院,文大学)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.CL

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20425 2025-09-03 cs.AI cs.RO 70%

Perspective-Shifted Neuro-Symbolic World Models: A Framework for Socially-Aware Robot Navigation

Kevin Alcedo, Pedro U. Lima, Rachid Alami

机构 * Institute for Systems and Robotics, Instituto Superior Técnico, Universidade de Lisboa(系统机器人研究所,理工学院,里斯本大学) LAAS-CNRS, Artificial and Natural Intelligence Toulouse Institute (ANITI)(LAAS-CNRS,图卢兹人工智能研究所(ANITI))

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

Comments Accepted as a regular paper at the 2025 IEEE International Conference on Robot & Human Interactive Communication (RO-MAN). \c{opyright} 2025 IEEE. The final version will appear in IEEE Xplore

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16315 2025-08-26 cs.LG 70%

OwkinZero: Accelerating Biological Discovery with AI

Nathan Bigaud, Vincent Cabeli, Meltem Gürel, Arthur Pignet, John Klein, Gilles Wainrib, Eric Durand

机构 * Owkin

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.LG

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17380 2025-08-26 cs.AI 70%

Mimicking the Physicist's Eye:A VLM-centric Approach for Physics Formula Discovery

Jiaqi Liu, Songning Lai, Pengze Li, Di Yu, Wenjie Zhou, Yiyang Zhou, Peng Xia, Zijun Wang, Xi Chen, Shixiang Tang, Lei Bai, Wanli Ouyang, Mingyu Ding, Huaxiu Yao, Aoran Wang

机构 * UNC–Chapel Hill(北卡罗来纳大学教堂山分校) HKUST (Guangzhou)(香港科技大学(广州)) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Fudan University(复旦大学) Tsinghua University(清华大学) Nankai University(南开大学) UC Santa Cruz(圣塔克鲁兹大学) The Chinese University of Hong Kong(香港中文大学) Shanghai Innovation Institute(上海创新研究院)

专题命中 逻辑推理 :reasoning(abstract);CoT(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15204 2025-08-22 cs.AI 70%

R-ConstraintBench: Evaluating LLMs on NP-Complete Scheduling

Raj Jain, Marc Wetter

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13743 2025-08-20 cs.CL 70%

Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA

Kaiwei Zhang, Qi Jia, Zijian Chen, Wei Sun, Xiangyang Zhu, Chunyi Li, Dandan Zhu, Guangtao Zhai

专题命中 逻辑推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11599 2025-08-18 cs.CR cs.AI 70%

CryptoScope: Utilizing Large Language Models for Automated Cryptographic Logic Vulnerability Detection

Zhihao Li, Zimo Ji, Tao Zheng, Hao Ren, Xiao Lan

机构 * Sichuan University(四川大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 逻辑推理 :chain-of-thought(abstract);CoT(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01300 2025-08-05 cs.AI 70%

How Far Are LLMs from Symbolic Planners? An NLP-Based Perspective

Ma'ayan Armony, Albert Meroño-Peñuela, Gerard Canal

专题命中 逻辑推理 :reasoning(abstract);planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22478 2025-07-31 cs.CL 70%

SLM-SQL: An Exploration of Small Language Models for Text-to-SQL

Lei Sheng, Shuai-Shuai Xu

机构 * Wuhan University of Technology(武汉理工大学) University of Science and Technology of China(中国科学技术大学)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.CL

Comments 16 pages, 2 figures, work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.19442 2025-07-31 cs.AI cs.DC 70%

A Survey on Large Language Model Acceleration based on KV Cache Management

Haoyang Li, Yiming Li, Anxin Tian, Tianhao Tang, Zhanchao Xu, Xuejia Chen, Nicole Hu, Wei Dong, Qing Li, Lei Chen

机构 * Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学) Department of Computer Science and Engineering(计算机科学与工程系) Department of Computer Science and Technology(计算机科学与技术系) Department of Computing and Data Science(计算与数据科学系)

专题命中 逻辑推理 :reasoning(abstract);logical reasoning(abstract);分类 cs.AI

Comments Accepted to TMLR 2025. The revised version incorporates more papers and has been further polished

详情

展开后加载摘要…

URL PDF HTML 收藏