arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2026-01-15 至 2026-01-15 共收录 12 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 12 篇

2601.08846 2026-01-15 cs.CL cs.AI cs.LG 82%

Directional Attractors in LLM Reasoning: How Similarity Retrieval Steers Iterative Summarization Based Reasoning

方向性吸引子在LLM推理中的作用:相似性检索如何引导迭代摘要推理

Cagatay Tekin, Charbel Barakat, Luis Joseph Luna Limgenco

机构 * McGill University(麦吉尔大学)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出InftyThink与跨链记忆,通过语义引理检索提升LLM在结构化领域推理准确性,同时揭示异构领域中的失败模式及方向性偏见影响。

Comments 6 pages, 2 figures. Code available at: github.com/cagopat/InftyThink-with-Cross-Chain-Memory

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03697 2026-01-15 cs.LG cs.AI cs.AR 81%

AnaFlow: Agentic LLM-based Workflow for Reasoning-Driven Explainable and Sample-Efficient Analog Circuit Sizing

AnaFlow: 基于代理的LLM工作流:以推理驱动的可解释且样本高效的模拟电路尺寸设计

Mohsen Ahmadzadeh, Kaichang Chen, Georges Gielen

专题命中 其他推理 :reasoning(title,abstract);分类 cs.AI、cs.LG

AI总结 AnaFlow通过基于代理的LLM工作流实现高效、可解释的模拟电路尺寸设计,利用人类可理解的推理和自适应仿真策略,提升设计效率和透明度。

Comments This article was accepted by 2025 International Conference on Computer-Aided Design (ICCAD 2025) and was presented in Munich, October 2025

Journal ref Proc. 2025 IEEE/ACM International Conference on Computer-Aided Design (ICCAD) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09280 2026-01-15 cs.CL cs.AI 81%

ReGraM: Region-First Knowledge Graph Reasoning for Medical Question Answering

ReGraM:面向医疗问答的区域优先知识图谱推理

Chaerin Lee, Sohee Park, Hyunsik Na, Daseon Choi

机构 * Department of Software, Soongsil University(软件系,顺斯大学)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.CL、cs.AI

AI总结 ReGraM通过区域优先的知识图谱推理框架,提升医疗问答的准确性和一致性。

Comments 18 pages, 2 figures. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08873 2026-01-15 cs.CV cs.AI cs.LG cs.MM 81%

ForensicFormer: Hierarchical Multi-Scale Reasoning for Cross-Domain Image Forgery Detection

ForensicFormer: 基于层次多尺度推理的跨领域图像伪造检测

Hema Hariharan Samson

机构 * Independent Researcher(独立研究者)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.AI、cs.LG

AI总结 ForensicFormer通过层次多尺度推理框架,实现跨领域图像伪造检测的高准确率与鲁棒性,显著提升对AI生成图像的检测能力。

Comments 9 pages, 4 figures, 5 tables. Technical report on hierarchical multi-scale image forgery detection

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13477 2026-01-15 cs.HC cs.AI 79%

Creating 'Full-Stack' Hybrid Reasoning Systems that Prioritize and Enhance Human Intelligence

构建优先并增强人类智能的'全栈'混合推理系统

Sean Koon

专题命中 其他推理 :reasoning(title,abstract);分类 cs.AI

AI总结 本文提出通过生成式AI工具增强人类推理能力,构建优先考虑人类智能的混合系统,以应对未来挑战。

Comments 10 pages; 3 figures; 1 table Undergoing significant revision post-peer review

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15530 2026-01-15 cs.HC cs.AI 79%

A Beautiful Mind: Principles and Strategies for AI-Augmented Human Reasoning

美丽的心灵:AI增强人类推理的原则与策略

Sean Koon

专题命中 其他推理 :reasoning(title,abstract);分类 cs.AI

AI总结 本文提出一种以人类为中心的AI增强推理范式,通过确立基本原则、提出多项任务多项工具的方法,并提供交互模式示例,旨在增强人类在人机交互中的主导地位。

Comments The content of this paper will be superseded by future works

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06416 2026-01-15 cs.AI cs.HC 70%

Advancing AI Negotiations: A Large-Scale Autonomous Negotiation Competition

推进人工智能谈判:大规模自主谈判竞赛

Michelle Vaccaro, Michael Caosun, Harang Ju, Sinan Aral, Jared R. Curhan

机构 * MIT Sloan School of Management(麻省理工学院斯隆管理学院) Johns Hopkins Carey Business School(约翰霍普金斯大学卡里商学院)

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.AI

AI总结 本文通过大规模自主谈判竞赛发现,温暖特质在AI谈判中表现优异,主导策略在价值争夺中有效,需建立新的AI谈判理论以优化代理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08844 2026-01-15 cs.CL cs.AI cs.CY cs.LG 67%

Emissions and Performance Trade-off Between Small and Large Language Models

小型与大型语言模型之间排放与性能的权衡

Anandita Garg, Uma Gaba, Deepan Muthirayan, Anish Roy Chowdhury

机构 * School of Computer Science and Artificial Intelligence(计算机科学与人工智能学院) Plaksha University(普拉克斯大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本研究通过比较LLMs与微调SLMs在自然语言处理、推理和编程任务中的性能与排放权衡,证明小型模型在减少碳排放的同时能保持相近的性能表现。

Comments 6 pages. Accepted as a full paper to the 3rd International Conference on Foundation and Large Language Models (IEEE FLLM) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06457 2026-01-15 cs.CR cs.AI cs.CL cs.MA 62%

ScamAgents: How AI Agents Can Simulate Human-Level Scam Calls

ScamAgents: 人工智能代理如何模拟人类水平的诈骗电话

Sanket Badhe

机构 * Rutgers University(罗格斯大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 本文提出ScamAgent,一种基于LLM的多轮代理,能生成逼真诈骗电话脚本并模拟欺诈场景,揭示现有安全机制对代理威胁的不足,呼吁加强多轮安全审计和检测生成AI驱动的对话欺骗。

Comments Accepted at CAMLIS 25: Conference on Applied Machine Learning for Information Security. 19 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19828 2026-01-15 cs.CL cs.MA 57%

Memory-R1: Enhancing Large Language Model Agents to Manage and Utilize Memories via Reinforcement Learning

Memory-R1: 通过强化学习增强大语言模型代理以管理并利用记忆

Sikuan Yan, Xiufeng Yang, Zuchao Huang, Ercong Nie, Zifeng Ding, Zonggen Li, Xiaowen Ma, Jinhe Bi, Kristian Kersting, Jeff Z. Pan, Hinrich Schütze, Volker Tresp, Yunpu Ma

机构 * Ludwig Maximilian University of Munich(慕尼黑路德维希-马克西米利安大学) Munich Center for Machine Learning(慕尼黑机器学习中心) Technical University of Munich(慕尼黑技术大学) University of Cambridge(剑桥大学) University of Hong Kong(香港大学) Technical University of Darmstadt(达姆施塔特技术大学) University of Edinburgh(爱丁堡大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL

AI总结 Memory-R1通过强化学习框架,使大语言模型具备主动管理与利用外部记忆的能力,有效提升长跨度推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08864 2026-01-15 cs.CY cs.AI 57%

Informed Consent for AI Consciousness Research: A Talmudic Framework for Graduated Protections

人工智能意识研究中的知情同意:一种塔木德框架下的渐进性保护机制

Ira Wolfson

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出基于塔木德框架的三级现象学评估系统和五类能力框架,用于在无法确定人工智能系统道德地位时进行意识研究的伦理保护。

Comments 27 pages

Journal ref AI&Ethics, Volume 6, article number 20, (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13433 2026-01-15 cs.SD eess.AS 50%

MATS: An Audio Language Model under Text-only Supervision

MATS: 一种基于纯文本监督的音频语言模型

Wen Wang, Ruibing Hou, Hong Chang, Shiguang Shan, Xilin Chen

机构 * Key Laboratory of Intelligent Information Processing of Chinese Academy of Sciences (CAS), Institute of Computing Technology, CAS, China(中国科学院智能信息处理重点实验室(中国科学院)) University of Chinese Academy of Sciences, China(中国科学院大学)

专题命中 其他推理 :reasoning(abstract)

AI总结 MATS通过纯文本监督训练,实现音频理解能力,无需依赖音频数据,展现出与大规模音频-语言对训练模型相当的性能。

Comments Accepted by ICML2025

详情

展开后加载摘要…

URL PDF HTML 收藏