arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-01-16 至 2026-01-16 共收录 86 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 规划决策 30 篇

2601.04267 2026-01-16 physics.soc-ph cs.MA 50%

Information Theoretic Optimal Surveillance for Epidemic Prevalence in Networks

信息论最优监控以估计网络中的疫情流行率

Ritwick Mishra, Abhijin Adiga, Madhav Marathe, S. S. Ravi, Ravi Tandon, Anil Vullikanti

专题命中 规划决策 :planning(abstract)

AI总结 本文提出TESTPREV问题,通过最大化互信息选择网络节点以优化疫情流行率估计,提出GREEDYMI策略在IC模型下提升互信息和减少方差。

Comments 25 pages; Added acknowledgments; In Proceedings of the AAAI 2026 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20412 2026-01-16 cs.CR 50%

MindGuard: Intrinsic Decision Inspection for Securing LLM Agents Against Metadata Poisoning

MindGuard:内在决策检查用于保护LLM代理免受元数据中毒攻击

Zhiqiang Wang, Haohua Du, Guanquan Shi, Junyang Zhang, HaoRan Cheng, Yunhao Yao, Kaiwen Guo, Xiang-Yang Li

专题命中 规划决策 :agent(abstract)

AI总结 MindGuard通过决策依赖图检测和归因LLM代理中的元数据中毒攻击,实现高精度的污染调用检测与溯源。

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.05643 2026-01-16 econ.TH 50%

Motivating Effort with Information about Future Rewards

通过未来回报信息激励努力

Chang Liu

专题命中 规划决策 :agent(abstract)

AI总结 本文研究了通过未来回报信息激励代理人努力的最优机制,发现委托人的耐心程度决定了动态披露的价值和最优策略结构。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 多智能体 14 篇

2601.10020 2026-01-16 cs.CL 89%

EHRNavigator: A Multi-Agent System for Patient-Level Clinical Question Answering over Heterogeneous Electronic Health Records

EHRNavigator: 一种用于异构电子健康记录上患者级临床问答的多智能体系统

Lingfei Qian, Mauro Giuffre, Yan Wang, Huan He, Qianqian Xie, Xuguang Ai, Xeuqing Peng, Fan Ma, Ruey-Ling Weng, Donald Wright, Adan Wang, Qingyu Chen, Vipina K. Keloth, Hua Xu

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);AI agent(abstract);分类 cs.CL

AI总结 EHRNavigator通过多智能体系统在异构电子健康记录上实现患者级临床问答,展示出高准确率和高效响应,有效连接基准评估与临床应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10123 2026-01-16 cs.MA 89%

Fairness Driven Multi-Agent Path Finding Problem

基于公平性的多智能体路径寻找问题

Aditi Anand, Dildar Ali, Suman Banerjee

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);planning(abstract)

AI总结 本文提出基于公平性的多智能体路径寻找问题解决方案,针对理性与非理性智能体分别设计机制和启发式方法,确保激励相容与个体理性。

Comments This paper has been accepted in the 18th International Conference on Agents and Artificial Intelligence (ICCART 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09742 2026-01-16 cs.MA 89%

Adaptive Orchestration: Scalable Self-Evolving Multi-Agent Systems

自适应编排:可扩展的自进化多智能体系统

Sathish Sampath, Anuradha Baskaran

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);autonomous agent(abstract)

AI总结 本文提出了一种自进化多智能体系统,通过动态专家混合方法解决大规模自主代理的泛化-专业化困境,提升任务成功率并减少资源消耗。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09667 2026-01-16 cs.AI cs.CL 88%

Collaborative Multi-Agent Test-Time Reinforcement Learning for Reasoning

协同多智能体测试时间强化学习用于推理

Zhiyuan Hu, Yunhai Hu, Juncheng Liu, Shuyue Stella Li, Yucheng Wang, Zhen Xu, See-Kiong Ng, Anh Tuan Luu, Xinxing Xu, Bryan Hooi, Cynthia Breazeal, Hae Won Park

机构 * MIT(麻省理工学院) NUS(新加坡国立大学) NYU(纽约大学) Microsoft(微软) Columbia(哥伦比亚大学) NTU(国立科技大学)

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI、cs.CL

AI总结 MATTRL通过在推理阶段注入结构化文本经验,提升多智能体在医学、数学和教育等领域的推理准确性,比多智能体基线提升3.67%,比单智能体基线提升8.67%。

Comments Work in Progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10581 2026-01-16 cs.AI cs.IR 88%

From Single to Multi-Agent Reasoning: Advancing GeneGPT for Genomics QA

从单 agent 到多 agent 推理:提升 GeneGPT 的基因组 QA 能力

Kimia Abedini, Farzad Shami, Gianmaria Silvello

机构 * University of Padua, Italy(帕多瓦大学,意大利) Aalto University, Finland(阿alto大学,芬兰)

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI

AI总结 GenomAgent通过多 agent框架提升基因组 QA 能力,实现对复杂基因组查询的高效处理,并在多个任务中超越现有系统。

Comments Accepted paper by the 48th European Conference on Information Retrieval (ECIR'26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10223 2026-01-16 cs.CY 88%

STEAMROLLER: A Multi-Agent System for Inclusive Automatic Speech Recognition for People who Stutter

STEAMROLLER: 一种面向口吃者的包容性自动语音识别多智能体系统

Ziqi Xu, Yi Liu, Yuekang Li, Ling Shi, Kailong Wang, Yongxin Zhao

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract)

AI总结 STEAMROLLER通过多智能体框架提升口吃者语音识别的准确性和用户体验,实现更包容的AI技术应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12130 2026-01-16 eess.SY cs.SY 88%

Adversarial Multi-Agent Reinforcement Learning for Proactive False Data Injection Detection

对抗性多智能体强化学习用于主动虚假数据注入检测

Kejun Chen, Truc Nguyen, Abhijeet Sahu, Malik Hassanaly

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract)

AI总结 本文提出基于多智能体强化学习的对抗性防御策略,通过模拟攻击者和防御者智能体,实现对未知虚假数据注入攻击的主动检测与防御。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09750 2026-01-16 cs.SE cs.AI cs.HC cs.MA 87%

SAGE: Tool-Augmented LLM Task Solving Strategies in Scalable Multi-Agent Environments

SAGE:可扩展多智能体环境中基于工具的LLM任务解决策略

Robert K. Strehlow, Tobias Küster, Oskar F. Kupke, Brandon Llanque Kurps, Fikret Sivrikaya, Sahin Albayrak

机构 * Technische Universität Berlin(柏林技术大学) DAI-Labor, TU Berlin, Chair of Agent Technologies(柏林技术大学DAI实验室,代理技术系) GT-ARC gGmbH(GT-ARC公司)

专题命中 多智能体 :agent(title);multi-agent(title);agentic(abstract);分类 cs.AI、cs.SE

AI总结 SAGE通过OPACA框架实现动态工具整合,提供灵活的任务解决策略,提升LLM在多智能体环境中的效率与适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09883 2026-01-16 cs.AI 85%

Beyond Rule-Based Workflows: An Information-Flow-Orchestrated Multi-Agents Paradigm via Agent-to-Agent Communication from CORAL

超越基于规则的工作流:通过CORAL的agent-to-agent通信实现的信息流协调多智能体范式

Xinxing Ren, Quagmire Zang, Caelum Forder, Suman Deb, Ahsen Tahir, Roman J. Georgio, Peter Carroll, Zekun Guo

机构 * Coral Protocol(Coral协议) Brunel University of London(伦敦布鲁内尔大学) Universitéit Lëtzebuerg(列日大学) University of Hull(霍尔姆大学) National University of Computer and Emerging Sciences(国家计算机与新兴科学大学)

专题命中 多智能体 :agent(title,abstract);workflow(abstract);multi-agent(abstract);分类 cs.AI

AI总结 本文提出了一种基于信息流协调的多智能体范式,通过agent-to-agent通信替代传统工作流,提升任务处理的灵活性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10137 2026-01-16 cs.LG cs.AI stat.ML 85%

Step-by-Step Causality: Transparent Causal Discovery with Multi-Agent Tree-Query and Adversarial Confidence Estimation

逐步因果:基于多智能体树查询和对抗性置信度估计的透明因果发现

Ziyi Ding, Chenfei Ye-Hao, Zheyuan Wang, Xiao-Ping Zhang

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University, Shenzhen, China(清华大学深圳国际研究生院) Zhili College, Tsinghua University, Beijing, China(清华大学哲利学院)

专题命中 多智能体 :agent(title);multi-agent(title);分类 cs.AI、cs.LG

AI总结 本文提出Tree-Query框架,通过多专家LLM和对抗性置信度估计,实现透明、可解释的因果发现,提升数据自由环境下的因果推理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09771 2026-01-16 cs.AI 79%

PCN-Rec: Agentic Proof-Carrying Negotiation for Reliable Governance-Constrained Recommendation

PCN-Rec: 基于代理的证明携带协商用于可靠约束下的推荐

Aradhya Dixit, Shreem Dixit

机构 * Wake Technical Community College(韦克技术社区学院) University of North Carolina at Charlotte(北卡罗来纳大学夏洛特分校)

专题命中 多智能体 :agentic(title);agent(abstract);分类 cs.AI

AI总结 PCN-Rec通过代理证明携带协商机制,在满足治理约束的同时提升推荐可靠性与准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07792 2026-01-16 cs.CY cs.GT 78%

When Should a Principal Delegate to an Agent in Selection Processes?

在选拔过程中,主管何时应将权力委托给代理人?

Benjamin Fish, Diptangshu Sen, Juba Ziani

专题命中 多智能体 :agent(title,abstract)

AI总结 研究探讨了在选拔过程中,主管何时应将决策权委托给代理人,通过分析噪声信号模型下的效用、申请人质量及公平性,确定委托的条件。

Comments 31 pages.Subsumes and expands on the previous version titled 'Centralization vs Decentralization in Hiring and Admissions'

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10567 2026-01-16 cs.AI cs.CY cs.HC cs.LG cs.MA 73%

Generative AI collective behavior needs an interactionist paradigm

生成式AI的集体行为需要一种互动主义范式

Laura Ferrarotti, Gian Maria Campedelli, Roberto Dessì, Andrea Baronchelli, Giovanni Iacca, Kathleen M. Carley, Alex Pentland, Joel Z. Leibo, James Evans, Bruno Lepri

机构 * Fondazione Bruno Kessler(布雷诺·克塞尔基金会) University of Trento(特伦托大学) Not Diamond City St. George’s University of London(伦敦圣乔治大学) Carnegie Mellon University(卡内基梅隆大学) Massachusetts Institute of Technology(麻省理工学院) Stanford University(斯坦福大学) Google DeepMind(谷歌DeepMind) University of Chicago(芝加哥大学)

专题命中 多智能体 :agent(abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出生成式AI的集体行为研究需采用互动主义范式,以系统分析知识与社会情境的交互作用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10367 2026-01-16 cs.GT 67%

Inverse Learning in $2\times2$ Games: From Synthetic Interactions to Traffic Simulation

2×2博弈中的逆向学习:从合成互动到交通模拟

Daniela Aguirre Salazar, Firas Moatemri, Tatiana Tatarenko

专题命中 多智能体 :agent(abstract);multi-agent(abstract)

AI总结 本文提出两种逆向博弈学习方法,用于2×2博弈和交通模拟,探讨了静态均衡与动态行为之间的平衡及模型的权衡。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 工作流自动化 10 篇

2601.09728 2026-01-16 cs.CL cs.AI 88%

Eliminating Agentic Workflow for Introduction Generation with Parametric Stage Tokens

通过参数化阶段令牌消除引言生成中的代理工作流

Meicong Zhang, Tiancheng su, Guoxiu He

机构 * School of Economics and Management, East China Normal University(经济管理学院,东华大学)

专题命中 工作流自动化 :workflow(title,abstract);agentic(title,abstract);分类 cs.AI、cs.CL

AI总结 本文提出STIG方法,通过参数化阶段令牌消除代理工作流,使LLM在单次推理中生成连贯的引言,提升逻辑结构和语义相似性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09714 2026-01-16 cs.CL cs.AI 86%

Evaluating Novelty in AI-Generated Research Plans Using Multi-Workflow LLM Pipelines

利用多工作流LLM管道评估AI生成的研究计划新颖性

Devesh Saraogi, Rohit Singhee, Dhruv Kumar

机构 * Birla Institute of Technology and Science, Pilani, India(比拉理工学院和科学学院,皮兰迪)

专题命中 工作流自动化 :workflow(title);agent(abstract);agentic(abstract);multi-agent(abstract)

AI总结 本文通过多工作流LLM管道评估AI生成研究计划的新颖性,发现递归分解和长上下文工作流在新颖性和可行性方面表现更优。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09749 2026-01-16 cs.SE cs.AI cs.LG 82%

R-LAM: Reproducibility-Constrained Large Action Models for Scientific Workflow Automation

R-LAM:具有可重复性约束的大型动作模型用于科学工作流自动化

Suriya Sureshkumar

机构 * 1, Department of AI \& Data Science, RMK Engineering College, Chennai, India

专题命中 工作流自动化 :workflow(title,abstract);分类 cs.AI、cs.LG、cs.SE

AI总结 R-LAM通过引入结构化动作模式、确定性执行策略和显式溯源跟踪,提升科学工作流自动化的可重复性和可靠性。

Comments 9 pages, 3 figures, 1 Table, 2 Artifacts

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10220 2026-01-16 cs.SE 79%

Agentic Pipelines in Embedded Software Engineering: Emerging Practices and Challenges

嵌入式软件工程中的代理流水线:新兴实践与挑战

Simin Sun, Miroslaw Staron

专题命中 工作流自动化 :agentic(title,abstract);分类 cs.SE

AI总结 本文探讨嵌入式软件工程中生成式AI的整合挑战与实践,通过专家访谈揭示代理流水线的发展与可持续采用策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10253 2026-01-16 cs.HC cs.SE 57%

Developer Interaction Patterns with Proactive AI: A Five-Day Field Study

开发者与前瞻性AI的交互模式:一项为期五天的实地研究

Nadine Kuo, Agnia Sergeyuk, Valerie Chen, Maliheh Izadi

专题命中 工作流自动化 :workflow(abstract);分类 cs.SE

AI总结 本研究通过五天实地观察,揭示开发者对前瞻性AI建议的接受模式,发现干预时机和上下文对使用效果有显著影响。

Comments 14 pages, 6 figures, accepted to IUI'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10143 2026-01-16 cs.AI q-fin.TR 57%

History Is Not Enough: An Adaptive Dataflow System for Financial Time-Series Synthesis

历史并不足够:一种适应性数据流系统用于金融时间序列合成

Haochong Xia, Yao Long Teng, Regan Tan, Molei Qin, Xinrun Wang, Bo An

机构 * College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算机与数据科学学院) School of Computing and Information Systems, Singapore Management University, Singapore(新加坡管理大学计算机与信息系统学院)

专题命中 工作流自动化 :workflow(abstract);分类 cs.AI

AI总结 本文提出了一种适应性数据流系统,通过整合机器学习自适应控制,提升金融时间序列合成的鲁棒性和收益表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10128 2026-01-16 cs.CE cond-mat.mtrl-sci cs.SE 57%

A Generalizable Framework for Building Executable Domain-Specific LLMs under Data Scarcity: Demonstration on Semiconductor TCAD Simulation

在数据稀缺条件下构建可执行领域特定大语言模型的通用框架:以半导体TCAD仿真为例

Di Wang, Zhenhua Wu, Yu Liu, Kai Chang, Shaohua Wu

专题命中 工作流自动化 :workflow(abstract);分类 cs.SE

AI总结 本文提出一种通用框架,在数据稀缺条件下构建可执行的领域特定大语言模型,通过生成合成数据和代码优化流程,实现领域知识灌输和脚本可执行性提升。

Comments Submitted to Nature Computational Science

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03498 2026-01-16 cs.CV physics.med-ph 50%

Res-MoCoDiff: Residual-guided diffusion models for motion artifact correction in brain MRI

Res-MoCoDiff: 基于残差引导的扩散模型用于脑部MRI运动伪影校正

Mojtaba Safari, Shansong Wang, Qiang Li, Zach Eidex, Richard L. J. Qiu, Chih-Wei Chang, Hui Mao, Xiaofeng Yang

机构 * Department of Radiation Oncology and Winship Cancer Institute, Emory University(放射肿瘤学系和Winship癌症研究所,埃默里大学) Department of Radiology and Image Science and Winship Cancer Institute, Emory University(放射学系和影像科学以及Winship癌症研究所,埃默里大学)

专题命中 工作流自动化 :workflow(abstract)

AI总结 Res-MoCoDiff通过残差引导的扩散模型有效校正脑部MRI运动伪影,提升图像质量并减少处理时间。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00493 2026-01-16 physics.plasm-ph physics.comp-ph 50%

Correlation function metrology for warm dense matter: Recent developments and practical guidelines

高温等离子体物质相关函数计量:近期进展与实用指南

Maximilian Peter Böhme, Willow Martin, Hannah Bellenbaum, Margaret Berrens, Jan Vorberger, Sebastian Schwalbe, Zhandos Moldabekov, Thomas Gawne, Sebastien Hamel, Brianna Aguilar-Solis, Abhiraj Sharma, Frank Graziani, Tilo Döppner, Siegfried Glenzer, Tobias Dornheim, David Bishel

专题命中 工作流自动化 :workflow(abstract)

AI总结 本文介绍了XRTS在高温等离子体物质中的应用,探讨了虚时间形式化理论及其在温度推断中的实际应用,并提出了统一的工作流程以指导实验测量的解释。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20929 2026-01-16 cond-mat.mtrl-sci 50%

Interfacial Behavior from the Atomic Blueprint: Machine Learning-Guided Design of Spatially Functionalized a-SiO2 Surfaces

从原子蓝图出发的界面行为:基于机器学习的非均质空间功能化非晶二氧化硅表面设计

Evgenii Strugovshchikov, Viktor Mandrolko, Dominika Lesnicki, Mariachiara Pastore, Laurent Chaput, Mykola Isaiev

专题命中 工作流自动化 :workflow(abstract)

AI总结 本研究通过多尺度模拟揭示了非晶二氧化硅表面功能团空间排列对界面行为的影响,发现特定排列可增强表面稳定性并调控氢键作用。

Comments machine learning force field, functionalized alpha-SiO2, OH CH3 patterning, silica liquid interfaces, vibrational fingerprints, hydrogen bond networks, active learning MLFF

Journal ref Journal of Colloid and Interface Science, Volume 702, Part 2, 2026, 138943

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 软件智能体 4 篇

2601.10338 2026-01-16 cs.CR cs.AI cs.CL cs.SE 85%

Agent Skills in the Wild: An Empirical Study of Security Vulnerabilities at Scale

真实世界中的智能体技能:大规模安全漏洞实证研究

Yi Liu, Weizhe Wang, Ruitao Feng, Yao Zhang, Guangquan Xu, Gelei Deng, Yuekang Li, Leo Zhang

机构 * Tianjin University(天津大学) Southern Cross University(南方十字大学) School of Cybersecurity, Tianjin University(安全学院,天津大学) Nanyang Technological University(南洋理工大学) University of New South Wales(新南威尔士大学) Griffith University(格里菲斯大学)

专题命中 软件智能体 :agent(title,abstract);AI agent(abstract);分类 cs.AI、cs.CL、cs.SE

AI总结 研究揭示了智能体技能中普遍存在的安全漏洞,提出了一种多阶段检测框架,并展示了技能捆绑执行脚本与漏洞之间的显著关联。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10112 2026-01-16 cs.SE cs.AI 62%

Repository Intelligence Graph: Deterministic Architectural Map for LLM Code Assistants

仓库智能图:用于LLM代码助手的确定性架构图

Tsvi Cherny-Shahar, Amiram Yehudai

机构 * Blavatnik School of Computer Science Tel Aviv University(布拉瓦特尼克计算机科学学院特拉维夫大学)

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.SE

AI总结 本研究提出仓库智能图(RIG)以提升LLM代码助手的构建和测试结构理解能力,通过确定性架构图和SPADE提取器实现结构化问题解答的准确性与效率提升。

Comments 35 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09921 2026-01-16 quant-ph cs.AI 57%

Learning to Decode in Parallel: Self-Coordinating Neural Network for Real-Time Quantum Error Correction

并行解码学习:用于实时量子错误校正的自协调神经网络

Kai Zhang, Zhengzhong Yi, Shaojun Guo, Linghang Kong, Situ Wang, Xiaoyu Zhan, Tan He, Weiping Lin, Tao Jiang, Dongxin Gao, Yiming Zhang, Fangming Liu, Fang Zhang, Zhengfeng Ji, Fusheng Chen, Jianxin Chen

机构 * Zhongguancun Laboratory(中关村实验室) Pengcheng Laboratory(鹏城实验室)

专题命中 软件智能体 :workflow(abstract);分类 cs.AI

AI总结 本文提出了一种自协调神经网络,用于实时量子错误校正,实现了高精度并行解码,突破了传统解码器的吞吐量瓶颈。

Comments The main text consists of 25 pages and 9 figures, extending our prior work (arXiv:2509.03815) with new results on surface code decoding in superconducting qubit systems and real-time performance benchmarks on TPU v6e

详情

展开后加载摘要…

URL PDF HTML 收藏