arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-01-13 至 2026-01-13 共收录 64 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 规划决策 64 篇

2601.07577 2026-01-13 cs.AI 85%

Beyond Entangled Planning: Task-Decoupled Planning for Long-Horizon Agents

超越纠缠规划:面向长时间 horizon 的任务解耦规划

Yunfan Li, Bingbing Xu, Xueyun Tian, Xiucheng Xu, Huawei Shen

机构 * State Key Laboratory of AI Safety, Institute of Computing Technology, CAS(人工智能安全国家重点实验室、计算技术研究所、中国科学院) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 规划决策 :planning(title,abstract);agent(abstract);workflow(abstract);分类 cs.AI

AI总结 本文提出任务解耦规划(TDP)方法,通过将任务分解为子目标图,减少长horizon智能体的规划误差传播,提升鲁棒性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11890 2026-01-13 cs.RO cs.AI 85%

Integrating Symbolic RL Planning into a BDI-based Autonomous UAV Framework: System Integration and SIL Validation

将符号RL规划整合到基于BDI的自主无人机框架中:系统集成与SIL验证

Sangwoo Jeon, Juchul Shin, YeonJe Cho, Gyeong-Tae Kim, Seongwoo Kim

专题命中 规划决策 :planning(title,abstract);agent(abstract);multi-agent(abstract);分类 cs.AI

AI总结 本研究提出AMAD-SRL框架,整合符号RL与BDI方法,通过SIL验证提升无人机任务效率75%。

Comments This submission has been withdrawn by the authors due to institutional and contractual requirements related to security and export-control review

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07782 2026-01-13 cs.CL cs.AI cs.IR 84%

Beyond Single-Shot: Multi-step Tool Retrieval via Query Planning

超越单次检索:通过查询规划实现多步工具检索

Wei Fang, James Glass

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 规划决策 :planning(title,abstract);agentic(abstract);分类 cs.AI、cs.CL

AI总结 本文提出TOOLQP框架,通过迭代查询规划解决多步工具检索问题,实现更高效的检索和执行能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06064 2026-01-13 cs.CY cs.AI cs.MA 84%

Socio-technical aspects of Agentic AI

群体技术视角下的代理AI

Praveen Kumar Donta, Alaa Saleh, Ying Li, Shubham Vaishnav, Kai Fang, Hailin Feng, Yuchao Xia, Thippa Reddy Gadekallu, Qiyang Zhang, Xiaodan Shi, Ali Beikmohammadi, Sindri Magnússon, Ilir Murturi, Chinmaya Kumar Dehury, Marcin Paprzycki, Lauri Loven, Sasu Tarkoma, Schahram Dustdar

机构 * Department of Computer and Systems Sciences, Stockholm University(斯德哥尔摩大学计算机与系统科学系) Center for Ubiquitous Computing, University of Oulu(奥卢大学无处不在计算中心) College of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Zhejiang A\&F University, Hangzhou(浙江工业大学之江学院) School of Computer Science, Peking University(北京大学计算机科学学院) Department of Mechatronics, University of Prishtina(普里什蒂纳大学机电系) Department of Computer Science, IISER Berhampur(伯尔哈普尔IISER计算机科学系) Systems Research Institute Polish Academy of Sciences(波兰科学院系统研究所) Department of Computer Science, University of Helsinki(赫尔辛基大学计算机科学系)

专题命中 规划决策 :agentic(title,abstract);planning(abstract);分类 cs.AI

AI总结 本文从社会技术视角探讨代理AI,分析其技术组件与社会背景的关联,揭示伦理挑战及未来研究方向。

Comments Dear Reviewer, please note that this is not survey/review or position paper. This paper introduced new framework (MAD-BAD-SAD Framework) for Socio-technical aspects of Agentic AI, Ethical considerations, which is very important to consider beside technical development

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07342 2026-01-13 cs.AI 83%

Agentic Diagnostic Reasoning over Telecom and Datacenter Infrastructure

基于电信和数据中心基础设施的代理诊断推理

Nicolas Tacheny

机构 * Ni2 Innovation Lab(Ni2创新实验室)

专题命中 规划决策 :agentic(title,abstract);agent(abstract);分类 cs.AI

AI总结 本文提出基于代理的诊断框架,利用LLM通过MCP协议自主导航基础设施模型,实现故障诊断与影响发现,为自主事件解决和风险预测提供基础。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07232 2026-01-13 cs.AI 83%

Yes FLoReNce, I Will Do Better Next Time! Agentic Feedback Reasoning for Humorous Meme Detection

是的,Florne,我下次会做得更好!用于幽默表情包检测的代理反馈推理

Olivia Shanhong Liu, Pai Chet Ng, De Wen Soh, Konstantinos N. Plataniotis

专题命中 规划决策 :agentic(title,abstract);agent(abstract);分类 cs.AI

AI总结 Florne通过闭环反馈机制提升表情包幽默检测的适应性和解释质量。

Comments LaMAS@AAAI 2026 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06373 2026-01-13 cs.MA 82%

DemMA: Dementia Multi-Turn Dialogue Agent with Expert-Guided Reasoning and Action Simulation

DemMA:具有专家引导推理和行动模拟的痴呆多轮对话代理

Yutong Song, Jiang Wu, Kazi Sharif, Honghui Xu, Nikil Dutt, Amir Rahmani

专题命中 规划决策 :agent(title,abstract);multi-agent(abstract)

AI总结 DemMA通过专家引导的推理和行动模拟,实现了高保真的痴呆症患者多轮对话模拟,优于现有基线模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06102 2026-01-13 cs.AI cs.LG 82%

Dynamic Intelligence Ceilings: Measuring Long-Horizon Limits of Planning and Creativity in Artificial Systems

动态智能上限:测量人工智能系统长期规划和创造力的长期限制

Truong Xuan Khanh, Truong Quynh Hoa

机构 * H&K Research Studio, Clevix LLC(Clevix LLC 研究室)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出动态智能上限概念,通过轨迹导向评估框架量化人工智能系统长期规划与创造力的限制,揭示智能上限的动态性与轨迹依赖性。

Comments This paper introduces a trajectory-centric evaluation framework for analyzing long-horizon intelligence limits in artificial systems, focusing on developmental behavior, planning, and structural creativity rather than proposing new learning algorithms. 11 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06818 2026-01-13 cs.CL 81%

AgentHallu: Benchmarking Automated Hallucination Attribution of LLM-based Agents

AgentHallu: 评估基于大语言模型的代理的自动幻觉归因

Xuannan Liu, Xiao Yang, Zekun Li, Peipei Li, Ran He

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Department of Computer Science & Technology, Tsinghua University(清华大学计算机科学与技术系) University of California, Santa Barbara(加州大学圣芭芭拉分校) Center for Research on Intelligent Perception and Computing, NLPR, CASIA(智能感知与计算研究中心,国家工程实验室)

专题命中 规划决策 :agent(abstract);tool-use(abstract);planning(abstract);agentic(abstract)

AI总结 AgentHallu提出一个评估基于大语言模型的代理自动识别幻觉来源的任务,通过高质量轨迹和多级注释评估13个模型,发现顶级模型在定位幻觉步骤上表现有限,工具使用幻觉最难识别。

Comments Project page: https://liuxuannan.github.io/AgentHallu.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08521 2026-01-13 cs.AI 81%

FlowSearch: Advancing deep research with dynamic structured knowledge flow

FlowSearch: 通过动态结构化知识流推进深度研究

Yusong Hu, Runmin Ma, Yue Fan, Jinxin Shi, Zongsheng Cao, Yuhao Zhou, Jiakang Yuan, Shuaiyu Zhang, Shiyang Feng, Xiangchao Yan, Shufei Zhang, Wenlong Zhang, Lei Bai, Bo Zhang

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

专题命中 规划决策 :agent(abstract);planning(abstract);agentic(abstract);multi-agent(abstract)

AI总结 FlowSearch通过动态结构化知识流提升多智能体系统在复杂任务中的推理与执行能力,实现跨学科研究的有效推进。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06164 2026-01-13 cs.SE cs.AI 81%

Contract2Plan: Verified Contract-Grounded Retrieval-Augmented Optimization for BOM-Aware Procurement and Multi-Echelon Inventory Planning

Contract2Plan: 一种基于合同的验证性检索增强优化方法用于BOM感知采购和多层级库存规划

Sahil Agarwal

机构 * Independent Researcher(独立研究者)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI、cs.SE

AI总结 Contract2Plan通过验证性检索增强优化方法,解决合同条款与BOM耦合导致的采购和库存规划问题,确保计划的可行性与合规性。

Comments 22 pages, 5 figures, 4 tables, 1 algorithm

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06054 2026-01-13 cs.CL cs.AI 81%

A Multi-Stage Workflow for the Review of Marketing Content with Reasoning Large Language Models

一种基于推理大语言模型的多阶段营销内容审查工作流程

Alberto Purpura, Emily Chen, Swapnil Shinde

机构 * AI Foundations, Capital One(人工智能基础,Capital One)

专题命中 规划决策 :workflow(title,abstract);分类 cs.AI、cs.CL

AI总结 本文提出一种基于推理大语言模型的多阶段工作流程,用于自动审查营销内容的合规性,并评估不同微调策略和奖励函数对模型性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12626 2026-01-13 cs.AI cs.DC 80%

Automated Planning for Optimal Data Pipeline Instantiation

最优数据管道实例化的自动化规划

Leonardo Rosa Amado, Adriano Vogel, Dalvan Griebler, Gabriel Paludo Licks, Eric Simon, Felipe Meneguzzi

机构 * Pontifical Catholic University of Rio Grande do Sul, Brazil(里约格朗德杜斯鲁斯天主教大学) Johannes Kepler University Linz, Austria(林茨约翰·凯撒大学) Sapienza University of Rome, Italy(罗马萨皮恩扎大学) SAP Labs, France(SAP实验室) University of Aberdeen, Scotland(阿伯丁大学)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI

AI总结 本文提出了一种基于动作成本的规划方法,用于优化数据管道实例化,通过启发式算法减少总执行时间,并在实验中验证了其有效性。

Journal ref Proceedings of the ECAI Workshop on AI-based Planning for Complex Real-World Applications (CAIPI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07597 2026-01-13 cs.NE cs.AI 79%

Pheromone-Focused Ant Colony Optimization algorithm for path planning

基于信息素聚焦的蚁群优化算法用于路径规划

Yi Liu, Hongda Zhang, Zhongxue Gan, Yuning Chen, Ziqing Zhou, Chunlei Meng, Chun Ouyang

机构 * College of Intelligent Robotics and Advanced Manufacturing, Fudan University(智能机器人与先进制造学院,复旦大学)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI

AI总结 本文提出基于信息素聚焦的蚁群优化算法,通过三种策略提升路径规划中的收敛速度和解质量。

Comments Accepted to 2025 IEEE International Conference on Systems, Man, and Cybernetics (SMC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14656 2026-01-13 cs.AI 79%

Cost-Awareness in Tree-Search LLM Planning: A Systematic Study

树搜索LLM规划中的成本意识:系统研究

Zihao Zhang, Hui Wei, Kenan Jiang, Shijia Pan, Shu Kai, Fei Liu

机构 * Emory University(埃默里大学) University of California, Merced(加州大学默塞德分校)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI

AI总结 本文研究了树搜索LLM规划器在资源受限下的成本意识问题,发现现有方法难以找到最优计划,双向搜索表现最佳,MCTS在短时间任务中最优,表明需新算法而非单纯增加计算量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06562 2026-01-13 cs.LG 79%

Mosaic: Unlocking Long-Context Inference for Diffusion LLMs via Global Memory Planning and Dynamic Peak Taming

Mosaic: 通过全局内存规划和动态峰值镇压解锁扩散语言模型的长上下文推理

Liang Zheng, Bowen Shi, Yitao Hu, Jiawei Zhang, Ruofan Li, Sheng Chen, Wenxin Li, Keqiu Li

机构 * Tianjin University, China(天津大学)

专题命中 规划决策 :planning(title,abstract);分类 cs.LG

AI总结 Mosaic通过全局内存规划和动态峰值镇压,提升扩散语言模型的长上下文推理能力,实现内存效率和推理长度的显著提升。

Comments 11 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06062 2026-01-13 cs.CY cs.AI cs.HC 79%

From Values to Frameworks: A Qualitative Study of Ethical Reasoning in Agentic AI Practitioners

从价值到框架:关于代理AI从业者伦理推理的定性研究

Theodore Roberts, Bahram Zarrin

专题命中 规划决策 :agentic(title,abstract);分类 cs.AI

AI总结 本文通过定性研究揭示代理AI从业者在伦理推理中的三种框架:用户、设计和伦理,强调需管理这些框架以确保伦理结果。

Comments 10 pages, 2 charts, 1 heatmap

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07186 2026-01-13 cs.RO 78%

PROTEA: Securing Robot Task Planning and Execution

PROTEA:保障机器人任务规划与执行的安全性

Zainab Altaweel, Mohaiminul Al Nahian, Jake Juettner, Adnan Siraj Rakin, Shiqi Zhang

专题命中 规划决策 :planning(title,abstract)

AI总结 PROTEA通过LLM-as-a-Judge机制评估机器人任务计划的安全性,解决维度和历史挑战,提升任务规划系统的鲁棒性和安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19329 2026-01-13 physics.med-ph 78%

A primer on treatment planning aspects for temporally modulated pulsed radiation therapy

脉冲时间调制放射治疗治疗计划方面的入门指南

Christian Velten, Adam Bayliss, Jiayi Huang, Wolfgang A. Tomé

专题命中 规划决策 :planning(title,abstract)

AI总结 本文介绍了时间调制脉冲放射治疗的治疗计划方法,强调了VMAT在优化剂量均匀性和适应性方面的优势,以及3D-CRT在提高治疗可及性方面的贡献。

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.05213 2026-01-13 math.OC 78%

A multi-objective mixed integer linear programming model for supply chain planning of 3D printing

面向3D打印供应链规划的多目标混合整数线性规划模型

Amirreza Talebi

专题命中 规划决策 :planning(title,abstract)

AI总结 本文提出了一种多目标混合整数线性规划模型,用于优化3D打印供应链规划,旨在最小化生产提前和延迟,同时最大化机器利用率,并通过数值示例分析了最佳边选择对结果的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16041 2026-01-13 math.AP 78%

Transporting a Dirac mass in a mean field planning problem

在均场规划问题中运输一个狄拉克质量

Pierre Cardaliaguet, Sebastian Munoz, Alessio Porretta

专题命中 规划决策 :planning(title,abstract)

AI总结 本文研究了初始密度为狄拉克质量的均场规划问题,证明了解的唯一性及自相似收敛性,并通过李雅普诺夫泛函分析解在初始时间附近的行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06122 2026-01-13 cs.CV cs.AI cs.LG 76%

COVR:Collaborative Optimization of VLMs and RL Agent for Visual-Based Control

COVR:视觉控制中视觉语言模型与强化学习代理的协同优化

Canming Xia, Peixi Peng, Guang Tan, Zhan Su, Haoran Xu, Zhenxian Liu, Luntong Li

专题命中 规划决策 :agent(title);分类 cs.AI、cs.LG

AI总结 COVR通过协同优化视觉语言模型与强化学习代理,利用RL生成数据提升VLM语义推理能力,并通过动作先验引导策略学习,实现视觉控制任务中的高效优化。

Comments The paper was accepted by the Fortieth AAAI Conference on Artificial Intelligence (AAAI-26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07690 2026-01-13 cs.LO 75%

On Angels and Demons: Strategic (De)Construction of Dynamic Models

关于天使与恶魔:动态模型的战略(破坏)构建

Davide Catta, Rustam Galimullin, Munyque Mittelmann

专题命中 规划决策 :agent(abstract);planning(abstract);multi-agent(abstract)

AI总结 本文提出三种逻辑用于研究动态图拓扑中策略的构建与破坏,探讨了其表达能力和模型检查复杂性。

Comments This is an extended version of the paper with the same title that will appear in the proceedings of AAMAS 2026. This version contains a technical appendix with proof details

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07038 2026-01-13 cs.AI 74%

Tool-Augmented Policy Optimization: Synergizing Reasoning and Adaptive Tool Use with Reinforcement Learning

工具增强的策略优化:通过强化学习协同推理与自适应工具使用

Wenxun Wu, Yuanyang Li, Guhan Chen, Linyue Wang, Hongyang Chen

专题命中 规划决策 :tool use(title);分类 cs.AI

AI总结 TAPO通过强化学习整合多跳推理与自适应工具调用,提升模型在知识密集型和计算密集型任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08812 2026-01-13 cs.RO cs.AI 74%

Adaptive Science Operations in Deep Space Missions Using Offline Belief State Planning

利用离线信念状态规划实现深空任务的自适应科学操作

Grace Ra Kim, Hailey Warner, Duncan Eddy, Evan Astle, Zachary Booth, Edward Balaban, Mykel J. Kochenderfer

机构 * Department of Aeronautics and Astronautics, Stanford University, Stanford, CA, 94305, USA(斯坦福大学航空航天系) Intelligent Systems Division, NASA Ames Research Center, Moffett Field, CA, 94035, USA(美国航空航天局阿姆斯研究中心智能系统分部)

专题命中 规划决策 :planning(title);分类 cs.AI

AI总结 本文提出基于POMDP的离线信念状态规划方法,用于深空任务中自适应科学仪器调度,通过整合贝叶斯网络提升数据可解释性与计算效率,并在Enceladus Orbilander任务中验证了其有效性。

Comments 7 pages, 4 tables, 5 figures, accepted in IEEE ISPARO 2025 (V2 - grammatical edits, also mispelled conference year)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07226 2026-01-13 cs.AI cs.CL 73%

Lost in the Noise: How Reasoning Models Fail with Contextual Distractors

陷入噪声中:推理模型在上下文干扰中的失败

Seongyun Lee, Yongrae Jo, Minju Seo, Moontae Lee, Minjoon Seo

机构 * University of Illinois Chicago(伊利诺伊大学芝加哥分校)

专题命中 规划决策 :tool-use(abstract);agentic(abstract);分类 cs.AI、cs.CL

AI总结 本文提出RARE方法,通过激励识别噪声中的有用信息,提升模型在上下文干扰下的鲁棒性,揭示增加计算量反而导致噪声环境下性能下降的反比例趋势。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06191 2026-01-13 cs.LG cs.AI 73%

TimeGNN-Augmented Hybrid-Action MARL for Fine-Grained Task Partitioning and Energy-Aware Offloading in MEC

时间图神经网络增强的混合动作多智能体强化学习用于MEC中的细粒度任务划分和能耗感知卸载

Wei Ai, Yun Peng, Yuntao Shou, Tao Meng, Keqin Li

机构 * College of Computer and Mathematics(计算机与数学学院) Central South University of Forestry and Technology(林业与技术中央南大学) Department of Computer Science(计算机科学系) State University of New York(纽约州立大学)

专题命中 规划决策 :agent(abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出TG-DCMADDPG算法,结合时间图神经网络和多智能体强化学习,优化MEC中的细粒度任务划分与能耗感知卸载。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07375 2026-01-13 cs.CL 70%

GROKE: Vision-Free Navigation Instruction Evaluation via Graph Reasoning on OpenStreetMap

GROKE: 通过OpenStreetMap上的图推理进行无视觉导航指令评估

Farzad Shami, Subhrasankha Dey, Nico Van de Weghe, Henrikki Tenkanen

机构 * Aalto University(阿alto大学) Ghent University(根特大学)

专题命中 规划决策 :agent(abstract);planning(abstract);分类 cs.CL

AI总结 GROKE提出一种基于OpenStreetMap的无视觉导航指令评估框架,通过图推理和拓扑导航提升评估精度与可扩展性。

Comments Under Review for ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06126 2026-01-13 cs.AI 70%

NL2Dashboard: A Lightweight and Controllable Framework for Generating Dashboards with LLMs

NL2Dashboard: 一种轻量且可控的基于LLM生成仪表盘的框架

Boshen Shi, Kexin Yang, Yuanbo Yang, Guanguang Chang, Ce Chi, Zhendong Wang, Xing Wang, Junlan Feng

机构 * Jiutian Research, China Mobile(九天研究,中国移动)

专题命中 规划决策 :agent(abstract);multi-agent(abstract);分类 cs.AI

AI总结 NL2Dashboard通过分析-呈现解耦原理,提出轻量可控的仪表盘生成框架,实现高效视觉质量和高可控性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06098 2026-01-13 cs.AI 70%

Automatic Question Generation for Intuitive Learning Utilizing Causal Graph Guided Chain of Thought Reasoning

利用因果图引导的推理链生成自动问题以促进直观学习

Nicholas X. Wang, Neel V. Parpia, Aaryan D. Parikh, Aggelos K. Katsaggelos

专题命中 规划决策 :agent(abstract);multi-agent(abstract);分类 cs.AI

AI总结 本文提出利用因果图引导的推理链和多智能体架构,生成准确且符合课程要求的问题,以提高直观学习的效果。

详情

展开后加载摘要…

URL PDF HTML 收藏