arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-01-14 至 2026-01-14 共收录 16 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 多智能体 16 篇

2601.08288 2026-01-14 cs.AI 89%

OpenMic: A Multi-Agent-Based Stand-Up Comedy Generation System

OpenMic: 基于多智能体的站立喜剧生成系统

Yuyang Wu, Hanzhong Cao, Jianhao Chen, Yufei Li

机构 * School of Electronics Engineering and Computer Science, Peking University(电子工程与计算机科学学院,北京大学)

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);planning(abstract);分类 cs.AI

AI总结 OpenMic通过多智能体系统生成基于文化背景的站立喜剧,结合检索增强生成和专用JokeWriter提升幽默与节奏表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23742 2026-01-14 cs.SE cs.AI 88%

AgenticTCAD: A LLM-based Multi-Agent Framework for Automated TCAD Code Generation and Device Optimization

AgenticTCAD: 基于LLM的多智能体框架用于自动TCAD代码生成与器件优化

Guangxi Fan, Tianliang Ma, Xuguang Sun, Xun Wang, Kain Lu Low, Leilai Shao

机构 * State Key Laboratory of Micro-nano Engineering Science(微纳工程科学国家重点实验室) Micro-nano Engineering Sciences Research Center(微纳工程科学研究中心) School of Mechanical Engineering, Shanghai Jiao Tong University(上海交通大学机械工程学院)

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI、cs.SE

AI总结 AgenticTCAD通过基于LLM的多智能体框架实现自动TCAD代码生成与器件优化,显著提升设计效率。

Comments Accepted by DATE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08343 2026-01-14 cs.MA cs.CL 88%

When KV Cache Reuse Fails in Multi-Agent Systems: Cross-Candidate Interaction is Crucial for LLM Judges

当KV缓存复用在多智能体系统中失效:跨候选者交互对LLM判断者至关重要

Sichu Liang, Zhenglin Wang, Jiajia Chu, Pengfei Xia, Hui Zang, Deyu Zhou

机构 * Southeast University(东南大学) Huawei Technologies Ltd(华为技术有限公司)

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);分类 cs.CL

AI总结 本研究揭示了在多智能体系统中KV缓存复用的失效模式,指出跨候选者交互对LLM判断者的重要性,强调了以判断者为中心的推理需要专门设计。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08327 2026-01-14 cs.RO cs.AI 88%

Safe Heterogeneous Multi-Agent RL with Communication Regularization for Coordinated Target Acquisition

安全的异构多智能体强化学习与通信正则化用于协同目标获取

Gabriele Calzolari, Vidya Sumathy, Christoforos Kanellakis, George Nikolakopoulos

机构 * Department of Computer Science, Electrical and Space Engineering, Luleå University of Technology, Luleå, Sweden.(计算机科学与空间工程系,卢勒奥技术大学,卢勒奥,瑞典)

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI

AI总结 本文提出了一种安全的异构多智能体强化学习框架,通过通信正则化和安全过滤器实现协同目标获取任务中的安全与稳定执行。

Comments 7 pages, 4 figures, submitted to the IFAC World Congress 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08237 2026-01-14 cs.AI 88%

The End of Reward Engineering: How LLMs Are Redefining Multi-Agent Coordination

奖励工程的终结:大型语言模型如何重新定义多智能体协调

Haoran Su, Yandong Sun, Congjia Yu

机构 * New York University(纽约大学) Lerna AI

专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI

AI总结 本文探讨了大型语言模型如何通过语义奖励规范和动态适应替代传统奖励工程,重新定义多智能体协调。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00551 2026-01-14 cs.AI cs.LG 84%

Single-agent Reinforcement Learning Model for Regional Adaptive Traffic Signal Control

区域自适应交通信号控制的单代理强化学习模型

Qiang Li, Ningjing Zeng, Lina Yu

专题命中 多智能体 :agent(title,abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出一种基于单代理强化学习的区域自适应交通信号控制模型,利用探测车辆数据有效缓解区域拥堵。

Comments A critical error in the methodology. The reported congestion control effects were not caused by the proposed signal timing optimization, but by an incorrect traffic volume scaling factor during evaluation. The traffic demand was not properly amplified, resulting in misleading performance gains. Due to the substantial nature of the error, completion of revisions is not feasible in the short term

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00549 2026-01-14 cs.LG cs.AI 84%

Robust Single-Agent Reinforcement Learning for Regional Traffic Signal Control Under Demand Fluctuations

鲁棒的单智能体强化学习用于应对需求波动的区域交通信号控制

Qiang Li, Jin Niu, Lina Yu

专题命中 多智能体 :agent(title,abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出一种鲁棒的单智能体强化学习框架,用于应对交通需求波动的区域交通信号控制,通过集中决策和高效学习模型有效减少交通队列长度。

Comments A critical error in the methodology. The reported congestion control effects were not caused by the proposed signal timing optimization, but by an incorrect traffic volume scaling factor during evaluation. The traffic demand was not properly amplified, resulting in misleading performance gains. Due to the substantial nature of the error, completion of revisions is not feasible in the short term

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08219 2026-01-14 cs.LG 83%

A Preliminary Agentic Framework for Matrix Deflation

一个初步的代理框架用于矩阵消去

Paimon Goulart, Evangelos E. Papalexakis

机构 * University of California, Riverside(加州大学河滨分校)

专题命中 多智能体 :agentic(title,abstract);agent(abstract);分类 cs.LG

AI总结 本文提出了一种基于代理的矩阵消去方法,利用LLM和VLM实现无阈值的消去,通过上下文学习和排列优化,在不同数据集上取得有竞争力的结果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08259 2026-01-14 cs.NI 82%

Unleashing Tool Engineering and Intelligence for Agentic AI in Next-Generation Communication Networks

释放工具工程与智能以推动下一代通信网络中的智能体AI

Yinqiu Liu, Ruichen Zhang, Dusit Niyato, Abbas Jamalipour, Trung Q. Duong, Dong In Kim

专题命中 多智能体 :agentic(title,abstract);planning(abstract)

AI总结 本文提出通过工具工程和智能提升下一代通信网络中的智能体AI,展示工具智能在无人飞行器轨迹规划中的应用,并提供6G时代的智能体构建路线图。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10326 2026-01-14 cs.AI cs.GT cs.LG cs.MA 79%

VGC-Bench: Towards Mastering Diverse Team Strategies in Competitive Pokémon

VGC-Bench: 向掌握多样化团队策略的竞技宝可梦迈进

Cameron Angliss, Jiaxun Cui, Jiaheng Hu, Arrasy Rahman, Peter Stone

机构 * University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 多智能体 :agent(abstract);AI agent(abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 VGC-Bench通过提供标准化评估和人类对战数据集,研究如何让AI代理在多样化团队策略的竞技宝可梦中实现稳健适应与泛化。

Comments AAMAS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07901 2026-01-14 stat.ML cs.AI cs.LG 73%

Decentralized Online Convex Optimization with Unknown Feedback Delays

去中心化在线凸优化与未知反馈延迟

Hao Qiu, Mengxiao Zhang, Juliette Achddou

专题命中 多智能体 :agent(abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种去中心化在线凸优化算法,解决了未知反馈延迟问题,通过自适应学习率机制实现了改进的后悔界。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08481 2026-01-14 cs.CR 67%

Baiting AI: Deceptive Adversary Against AI-Protected Industrial Infrastructures

诱饵AI:针对AI保护工业基础设施的欺骗性对手

Aryan Pasikhani, Prosanta Gope, Yang Yang, Shagufta Mehnaz, Biplab Sikdar

专题命中 多智能体 :agent(abstract);multi-agent(abstract)

AI总结 本文提出了一种基于深度强化学习的欺骗性攻击方法,用于针对AI保护的工业基础设施,通过定制策略实现隐蔽攻击并降低系统可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07997 2026-01-14 eess.SY cs.SY 67%

Can Inherent Communication Noise Guarantee Privacy in Distributed Cooperative Control ?

内在通信噪声能否保证分布式协同控制中的隐私?

Yuwen Ma, Sarah K. Spurgeon, Tao Li, Boli Chen

专题命中 多智能体 :agent(abstract);multi-agent(abstract)

AI总结 本文提出了一种基于LQR的分布式协同控制框架,利用内在通信噪声实现差分隐私保护,无需额外隐私噪声,同时保证系统跟踪误差的收敛性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17247 2026-01-14 eess.SY cs.SY math.OC 67%

Inverse Optimal Control for Linear Quadratic Tracking with Unknown Target States

线性二次跟踪中未知目标状态的逆最优控制

Yao Li, Chengpu Yu, Hao Fang, Jie Chen

专题命中 多智能体 :agent(abstract);multi-agent(abstract)

AI总结 本文提出了一种用于多智能体系统联合聚类协调和意图识别的逆最优控制算法,通过线性方程和结构约束提升算法效率和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08210 2026-01-14 cs.LG 57%

Scalable Multiagent Reinforcement Learning with Collective Influence Estimation

可扩展的多智能体强化学习与集体影响估计

Zhenglong Luo, Zhiyong Chen, Aoxiang Liu, Ke Pan

机构 * School of Engineering, The University of Newcastle(工程学院) School of Automation, Central South University(自动化学院)

专题命中 多智能体 :agent(abstract);分类 cs.LG

AI总结 本文提出了一种可扩展的多智能体强化学习框架,通过集体影响估计网络实现高效协作,避免网络扩展并提升鲁棒性和部署可行性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00634 2026-01-14 cs.SE cs.HC 57%

Does GenAI Make Usability Testing Obsolete?

生成式AI会使可用性测试过时吗?

Ali Ebrahimi Pourasad, Walid Maalej

专题命中 多智能体 :workflow(abstract);分类 cs.SE

AI总结 本文提出UX-LLM,一种基于大视觉语言模型的可用性问题预测工具,虽无法完全替代传统测试,但可作为补充,尤其适用于资源有限的小型团队。

Comments Accepted for publication at The 47th IEEE/ACM International Conference on Software Engineering ICSE 2025

详情

展开后加载摘要…

URL PDF HTML 收藏