arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-02-17 至 2026-02-17 共收录 86 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 规划决策 86 篇

2510.11608 2026-02-17 cs.AI 92%

ParaCook: On Time-Efficient Planning for Multi-Agent Systems

ParaCook: 多智能体系统中高效时间规划的研究

Shiqi Zhang, Xinbei Ma, Yunqing Xu, Zouying Cao, Pengrui Lu, Haobo Yuan, Tiancheng Shen, Zhuosheng Zhang, Hai Zhao, Ming-Hsuan Yang

机构 * Shanghai Jiao Tong University(上海交通大学) University of California, Merced(加州大学梅尔德分校)

专题命中 规划决策 :agent(title,abstract);planning(title,abstract);multi-agent(title,abstract);分类 cs.AI

AI总结 ParaCook通过简化动作空间和烹饪任务实例化,为多智能体系统高效时间规划提供基准测试框架,评估LLMs在并行协调和高层次优化中的表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07443 2026-02-17 cs.AI cs.GT 89%

Approximating Human Strategic Reasoning with LLM-Enhanced Recursive Reasoners Leveraging Multi-agent Hypergames

用LLM增强的递归推理器近似人类战略推理

Vince Trencsenyi, Agnieszka Mensfelt, Kostas Stathis

机构 * Royal Holloway University of London(伦敦皇家霍洛威大学)

专题命中 规划决策 :agent(title,abstract);multi-agent(title,abstract);agentic(abstract);分类 cs.AI

AI总结 本文提出了一种基于LLM增强的递归推理器框架,通过多智能体超游戏模型评估LLM的递归推理能力,并展示了其在近似人类行为和最优解达成方面的优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01848 2026-02-17 cs.AI cs.MA 89%

ROMA: Recursive Open Meta-Agent Framework for Long-Horizon Multi-Agent Systems

ROMA:递归开放元代理框架用于长周期多代理系统

Salaheddin Alzu'bi, Baran Nama, Arda Kaz, Anushri Eswaran, Weiyuan Chen, Sarvesh Khetan, Rishab Bala, Tu Vu, Sewoong Oh

机构 * Sentient Virginia Tech(弗吉尼亚理工大学) UC Berkeley(加州大学伯克利分校) UC San Diego(加州大学圣地亚哥分校) University of Maryland(马里兰大学)

专题命中 规划决策 :agent(title,abstract);multi-agent(title,abstract);agentic(abstract);分类 cs.AI

AI总结 ROMA通过递归任务分解和结构化聚合,实现长周期多代理系统的高效推理和生成性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14160 2026-02-17 cs.AI 88%

Process-Supervised Multi-Agent Reinforcement Learning for Reliable Clinical Reasoning

过程监督多智能体强化学习用于可靠的临床推理

Chaeeun Lee, T. Michael Yates, Pasquale Minervini, T. Ian Simpson

机构 * School of Informatics, University of Edinburgh, UK(信息学院,爱丁堡大学)

专题命中 规划决策 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI

AI总结 本文提出了一种过程监督的多智能体强化学习框架,用于提升基因-疾病有效性校准的临床推理准确性与过程忠实度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20430 2026-02-17 cs.CL cs.AI cs.CV cs.MA 86%

An Agentic System for Rare Disease Diagnosis with Traceable Reasoning

一种具有可追溯推理的罕见病诊断代理系统

Weike Zhao, Chaoyi Wu, Yanjie Fan, Xiaoman Zhang, Pengcheng Qiu, Yuze Sun, Xiao Zhou, Yanfeng Wang, Xin Sun, Ya Zhang, Yongguo Yu, Kun Sun, Weidi Xie

专题命中 规划决策 :agentic(title,abstract);agent(abstract);multi-agent(abstract);分类 cs.AI、cs.CL

AI总结 DeepRare是一种基于大型语言模型的多代理系统,通过整合40多种专用工具和最新知识源,为罕见病诊断提供决策支持,实现了透明可追溯的推理链。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14606 2026-02-17 cs.MA cs.AI cs.CE 85%

Towards Selection as Power: Bounding Decision Authority in Autonomous Agents

迈向选择作为权力:自主代理中决策权威的边界

Jose Manuel de la Chica Rodriguez, Juan Manuel Vera Díaz

机构 * AI Lab, Grupo Santander Madrid, Spain(AI实验室,西班牙 Grupo Santander 马德里)

专题命中 规划决策 :autonomous agent(title,abstract);agent(abstract);agentic(abstract);分类 cs.AI

AI总结 本文提出一种治理架构,通过机械机制限制自主代理的选择权,以防止确定性结果并提升可审计性,重新定义治理为受限制的因果权力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13691 2026-02-17 cs.AI 85%

PhGPO: Pheromone-Guided Policy Optimization for Long-Horizon Tool Planning

PhGPO:基于信息素的策略优化用于长周期工具规划

Yu Li, Guangfeng Cai, Shengtian Yang, Han Luo, Shuo Han, Xu He, Dong Li, Lei Feng

专题命中 规划决策 :planning(title,abstract);tool use(abstract);tool-use(abstract);分类 cs.AI

AI总结 PhGPO通过学习历史轨迹中的信息素模式,指导策略优化以提升长周期工具规划能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14799 2026-02-17 cs.RO quant-ph 85%

Scalable Multi-Robot Path Planning via Quadratic Unconstrained Binary Optimization

通过二次无约束二元优化实现可扩展的多机器人路径规划

Javier González Villasmil

专题命中 规划决策 :planning(title,abstract);agent(abstract);multi-agent(abstract)

AI总结 本文提出了一种基于二次无约束二元优化的多机器人路径规划方法,通过逻辑预处理和时间窗口分解策略实现了高效且可扩展的路径规划。

Comments 21 pages, 9 figures, 1 table. Accompanying open-source implementation at https://github.com/JavideuS/Spooky

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14589 2026-02-17 cs.AI cs.CL cs.LG 85%

MATEO: A Multimodal Benchmark for Temporal Reasoning and Planning in LVLMs

MATEO:一种多模态基准,用于LVLMs中的时间推理和规划

Gabriel Roccabruna, Olha Khomyn, Giuseppe Riccardi

机构 * Signals and Interactive Systems Lab, University of Trento, Italy(特伦托大学信号与交互系统实验室) University of Trento(特伦托大学) Amazon(亚马逊)

专题命中 规划决策 :planning(title,abstract);AI agent(abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 MATEO是一个多模态基准,用于评估和提升大型视觉语言模型在时间推理和规划方面的能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14955 2026-02-17 cs.CL cs.SE 84%

Tool-Aware Planning in Contact Center AI: Evaluating LLMs through Lineage-Guided Query Decomposition

接触中心AI中的工具感知规划:通过 lineage 引导的查询分解评估LLMs

Varun Nathan, Shreyas Guha, Ayush Kumar

专题命中 规划决策 :planning(title,abstract);agentic(abstract);分类 cs.CL、cs.SE

AI总结 本文提出接触中心AI中的工具感知规划框架,通过lineage引导查询分解评估LLMs,揭示工具理解的不足及简计划的优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13313 2026-02-17 cs.CV cs.AI 83%

Agentic Spatio-Temporal Grounding via Collaborative Reasoning

基于协作推理的代理时空 grounding

Heng Zhao, Yew-Soon Ong, Joey Tianyi Zhou

机构 * CFAR, IHPC, Agency for Science, Technology and Research(A*STAR)(CFAR、IHPC、新加坡科技研究局) CCDS, Nanyang Technological University(CCDS、南洋理工大学)

专题命中 规划决策 :agentic(title,abstract);agent(abstract);分类 cs.AI

AI总结 本文提出ASTG框架,通过协作推理实现开放世界下的时空视频grounding,提升检索效率并优于现有弱监督和零样本方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13868 2026-02-17 cs.NI cs.IR 82%

Agentic Assistant for 6G: Turn-based Conversations for AI-RAN Hierarchical Co-Management

6G代理助手:基于回合的对话用于AI-RAN分层协同管理

Udhaya Srinivasan, Weisi Guo

专题命中 规划决策 :agentic(title,abstract);planning(abstract)

AI总结 本文提出了一种基于回合的对话助手,用于AI-RAN的分层协同管理,通过三层架构实现高效网络管理和降低运营成本。

Comments submitted to IEEE conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25260 2026-02-17 cs.AI cs.CL cs.LG 82%

Internal Planning in Language Models: Characterizing Horizon and Branch Awareness

语言模型中的内部规划:刻画视野与分支意识

Muhammed Ustaomeroglu, Baris Askin, Gauri Joshi, Carlee Joe-Wong, Guannan Qu

机构 * Department of Electrical and Computer Engineering(电气与计算机工程系) Carnegie Mellon University(卡内基梅隆大学)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 研究通过分析语言模型内部计算结构,揭示规划视野与分支意识的特性,为理解模型内部动态提供通用工具。

Comments Accepted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13622 2026-02-17 cs.HC 82%

The Shadow Boss: Identifying Atomized Manipulations in Agentic Employment of XR Users using Scenario Constructions

影子老板:通过场景构建识别代理雇佣中XR用户原子化操控

Lik-Hang Lee

专题命中 规划决策 :agentic(title,abstract);AI agent(abstract)

AI总结 本文通过场景构建方法,识别代理雇佣中XR用户面临的七个关键风险,揭示AI代理对人类劳动的原子化操控问题,呼吁需以用户为中心的XR和政策干预。

Comments 28 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13577 2026-02-17 cs.RO 82%

ONRAP: Occupancy-driven Noise-Resilient Autonomous Path Planning

ONRAP:基于占用度的抗噪声自主路径规划

Faizan M. Tariq, Avinash Singh, Vipul Ramtekkar, Jovin D'sa, David Isele, Yosuke Sakamoto, Sangjae Bae

机构 * Honda Research Institute, CA, USA(本田研究院,美国加利福尼亚州) Honda Research and Development(本田研发)

专题命中 规划决策 :planning(title,abstract);agent(abstract)

AI总结 ONRAP 提出了一种基于占用网格的抗噪声路径规划方法,结合占用流预测以生成安全可行的路径,实现实时运行并有效应对不确定环境。

Comments 8 pages, 9 figures - Presented at 2026 IEEE Intelligent Vehicles Symposium (IV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13666 2026-02-17 cs.LG cs.AI 81%

ALMo: Interactive Aim-Limit-Defined, Multi-Objective System for Personalized High-Dose-Rate Brachytherapy Treatment Planning and Visualization for Cervical Cancer

ALMo:交互式目标-限制定义的多目标系统,用于宫颈癌高剂量率近距离治疗计划与可视化

Edward Chen, Natalie Dullerud, Pang Wei Koh, Thomas Niedermayr, Elizabeth Kidd, Sanmi Koyejo, Carlos Guestrin

机构 * Stanford University(斯坦福大学) University of Washington(华盛顿大学) Stanford University School of Medicine(斯坦福大学医学院) Paul G. Allen School of Computer Science & Engineering(保罗·G·艾伦计算机科学与工程学院)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI、cs.LG

AI总结 ALMo是一种用于宫颈癌高剂量率近距离治疗的交互式多目标系统,通过自动化参数设置和直观的剂量学权衡控制,提高治疗计划质量和效率。

Comments Abstract accepted at Symposium on Artificial Intelligence in Learning Health Systems (SAIL) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13616 2026-02-17 cs.AI cs.LG 81%

DiffusionRollout: Uncertainty-Aware Rollout Planning in Long-Horizon PDE Solving

DiffusionRollout:长时间尺度PDE求解中的不确定性感知 rollout 计划

Seungwoo Yoo, Juil Koo, Daehyeon Choi, Minhyuk Sung

专题命中 规划决策 :planning(title,abstract);分类 cs.AI、cs.LG

AI总结 DiffusionRollout通过自适应选择步长策略,提升长时间尺度PDE求解的预测可靠性与准确性。

Comments TMLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04911 2026-02-17 cs.AI 79%

From Stories to Cities to Games: A Qualitative Evaluation of Behaviour Planning

从故事到城市到游戏:行为规划的定性评估

Mustafa F. Abdelwahed, Joan Espasa, Alice Toniolo, Ian P. Gent

专题命中 规划决策 :planning(title,abstract);分类 cs.AI

AI总结 本文通过三个案例研究展示行为规划在现实世界中的应用,包括叙事、城市规划和游戏评估。

Journal ref PlanSig 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14225 2026-02-17 cs.AI 79%

Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding

文本优先于视觉:针对超高清遥感理解的代理强化学习在超高清遥感理解中的知识注入至关重要

Fengxiang Wang, Mingshuo Chen, Yueying Li, Yajie Yang, Yuhao Zhou, Di Wang, Yifan Zhang, Haoyu Wang, Haiyan Zhao, Hongda Sun, Long Lan, Jun Song, Yulin Wang, Jing Zhang, Wenlong Zhang, Bo Du

机构 * National University of Defense Technology, China(国防科技大学) Beijing University of Posts and Telecommunications, China(北京邮电大学) University of the Chinese Academy of Sciences, China(中国科学院大学) Sichuan University, China(四川大学) Wuhan University, China(武汉大学) Chinese Academy of Science, China(中国科学院) Tsinghua University, China(清华大学) Shanghai Artificial Intelligence Laboratory, China(上海人工智能实验室) Renmin University of China, China(中国人民大学)

专题命中 规划决策 :agentic(title,abstract);分类 cs.AI

AI总结 本文提出分阶段知识注入方法,利用文本引导提升超高清遥感理解的视觉推理能力,实现XLRS-Bench上的高准确率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13222 2026-02-17 cs.CC cs.AI cs.FL 79%

Computability of Agentic Systems

代理系统可计算性

Chatavut Viriyasuthee

专题命中 规划决策 :agentic(title,abstract);分类 cs.AI

AI总结 本文提出Quest图框架,分析代理系统在有限上下文下的计算能力,揭示不同模型的计算复杂度及性能差异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14229 2026-02-17 cs.AI cs.ET cs.LG 79%

CORPGEN: Simulating Corporate Environments with Autonomous Digital Employees in Multi-Horizon Task Environments

CORPGEN:在多时间尺度任务环境中通过自主数字员工模拟企业环境

Abubakarr Jaye, Nigel Boachie Kumankumah, Chidera Biringa, Anjel Shaileshbhai Patel, Sulaiman Vesal, Dayquan Julienne, Charlotte Siska, Manuel Raúl Meléndez Luján, Anthony Twum-Barimah, Mauricio Velazco, Tianwei Chen

机构 * Microsoft Corporation(微软公司)

专题命中 规划决策 :agent(abstract);autonomous agent(abstract);planning(abstract);分类 cs.AI、cs.LG

AI总结 CorpGen通过分层规划、子代理隔离和分层内存等机制,在多时间尺度任务环境中实现自主数字员工的高效模拟,提升企业环境模拟的完成率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13286 2026-02-17 cs.RO 78%

BoundPlanner: A convex-set-based approach to bounded manipulator trajectory planning

BoundPlanner: 一种基于凸集的有界机械臂轨迹规划方法

Thies Oelerich, Christian Hartl-Nesic, Florian Beck, Andreas Kugi

机构 * Automation and Control Institute (ACIN), TU Wien(自动化与控制研究所(ACIN),维也纳技术大学)

专题命中 规划决策 :planning(title,abstract)

AI总结 BoundPlanner通过基于凸集的轨迹规划方法,实现机械臂在考虑运动学和碰撞约束下的高效在线轨迹规划。

Comments Published at RA-L

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17880 2026-02-17 cs.RO 78%

Autonomous Navigation of Quadrupeds Using Coverage Path Planning with Morphological Skeleton Map

四足机器人自主导航的覆盖路径规划与形态学骨架地图

Alexander James Becoy, Kseniia Khomenko, Luka Peternel, Raj Thilak Rajan

机构 * Department of Cognitive Robotics, ME, Delft University of Technology(认知机器人系,代尔夫特理工大学) Department of Microelectronics, EEMCS, Delft University of Technology(微电子系,代尔夫特理工大学)

专题命中 规划决策 :planning(title,abstract)

AI总结 本文提出利用形态学骨架地图进行四足机器人自主导航的覆盖路径规划方法,通过有限状态机控制导航与扫描模式,实现了高效路径规划和环境扫描。

Comments 15 pages, published to Fronters In Robotics (currently in production), major revision: title change, abstract revised, grammar fixed, mathematical notations fixed and made consistent, conclusion revised, related works extended, Algorithm 1-3 revised

Journal ref Frontiers in Robotics and AI, Volume 31, 1601862, July 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13474 2026-02-17 cs.RO cs.SY eess.SY 78%

Planning Human-Robot Co-manipulation with Human Motor Control Objectives and Multi-component Reaching Strategies

基于人类运动控制目标和多组件抓取策略的人机协同规划

Kevin Haninger, Luka Peternel

机构 * Department of Automation, Fraunhofer IPK(自动化系,弗劳恩霍夫IPK研究所)

专题命中 规划决策 :planning(title,abstract)

AI总结 本文提出基于人类运动控制模型和多组件抓取策略的人机协同规划方法,以实现更自然的人机交互。

Comments 10 Pages

Journal ref IEEE Robotics and Automation Letters, Volume 10, Issue 2, February 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00866 2026-02-17 stat.ME math.ST stat.AP stat.TH 78%

Planning for gold: Hypothesis screening with split samples for valid powerful testing in matched observational studies

为黄金规划:通过分割样本进行假设筛查,以在匹配观察研究中实现有效且有力的检验

William Bekerman, Abhinandan Dalal, Carlo del Ninno, Dylan S. Small

专题命中 规划决策 :planning(title,abstract)

AI总结 本文提出了一种通过分割样本进行假设筛选的方法,以在匹配观察研究中实现有效且有力的检验,以应对未测量混杂因素的影响。

Comments To be published in Biometrika

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13457 2026-02-17 cs.RO 78%

Inferring Turn-Rate-Limited Engagement Zones with Sacrificial Agents for Safe Trajectory Planning

利用牺牲代理推断受限转弯率的接触区域以实现安全轨迹规划

Grant Stagg, Cameron K. Peterson

机构 * Brigham Young University

专题命中 规划决策 :planning(title);agent(abstract)

AI总结 本文通过牺牲代理推断受限转弯率的接触区域,利用几何启发式和贝叶斯方法优化轨迹选择,实现安全高效的轨迹规划。

Comments Submitted to the Journal of Aerospace Information Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13802 2026-02-17 cs.LG 77%

Cast-R1: Learning Tool-Augmented Sequential Decision Policies for Time Series Forecasting

Cast-R1:学习工具增强的序列决策策略用于时间序列预测

Xiaoyu Tao, Mingyue Cheng, Chuang Jiang, Tian Gao, Huanjian Zhang, Yaguo Liu

机构 * State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(认知智能国家重点实验室,中国科学技术大学)

专题命中 规划决策 :agent(abstract);workflow(abstract);agentic(abstract);分类 cs.LG

AI总结 Cast-R1通过工具增强的代理工作流程,学习序列决策策略以提升时间序列预测的准确性与适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14526 2026-02-17 cs.RO cs.AI cs.LG 73%

TWISTED-RL: Hierarchical Skilled Agents for Knot-Tying without Human Demonstrations

TWISTED-RL:无人类示范的分层技能代理用于打结

Guy Freund, Tom Jurgenson, Matan Sudry, Erez Karpas

机构 * Reichman University(雷赫曼大学) Technion – Israel Institute of Technology(技术学院——以色列理工学院) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 规划决策 :agent(abstract);planning(abstract);分类 cs.AI、cs.LG

AI总结 TWISTED-RL通过强化学习和拓扑动作实现无示范的复杂打结任务,提升泛化能力和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13210 2026-02-17 cs.NI cs.AI cs.LG 73%

Large Language Model (LLM)-enabled Reinforcement Learning for Wireless Network Optimization

基于大语言模型的强化学习用于无线网络优化

Jie Zheng, Ruichen Zhang, Dusit Niyato, Haijun Zhang, Jiacheng Wang, Hongyang Du, Jiawen Kang, Zehui Xiong

机构 * State-Province Joint Engineering and Research Center of Advanced Networking and Intelligent Information Services, College of Computer Science, Northwest University(高级网络与智能信息服务省-市联合工程与研究中心,计算机科学学院,西北大学) College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) Institute of Artificial Intelligence, University of Science and Technology Beijing(人工智能研究院,北京科技大学) Department of Electrical and Electronic Engineering, the University of Hong Kong(电子与电气工程系,香港大学) Automation of School, Guangdong University of Technology(自动化学院,广东工业大学) Queen’s University Belfast(贝尔法斯特女王大学)

专题命中 规划决策 :agent(abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出利用大语言模型增强强化学习,以优化6G无线网络,通过多智能体强化学习框架提升网络性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05152 2026-02-17 math.OC 71%

Extreme-Scale EV Charging Infrastructure Planning for Last-Mile Delivery Using High-Performance Parallel Computing

面向最后一公里配送的极端规模电动汽车充电基础设施规划:利用高性能并行计算

Waquar Kaleem, Taner Cokyasar, Jeffrey Larson, Omer Verbas, Tanveer Hossain Bhuiyan, Anirudh Subramanyam

专题命中 规划决策 :planning(title)

AI总结 本文提出了一种基于高性能并行计算的框架,用于解决极端规模电动汽车充电基础设施规划问题,通过分解和并行化方法高效处理大规模优化问题。

Journal ref Transportation Research Part B: Methodological, 205, 103403, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏