arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-02-18 至 2026-02-18 共收录 11 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 11 篇

2507.06134 2026-02-18 cs.AI 89%

OpenAgentSafety: A Comprehensive Framework for Evaluating Real-World AI Agent Safety

OpenAgentSafety: 一个全面评估现实世界AI代理安全性的框架

Sanidhya Vijayvargiya, Aditya Bharat Soni, Xuhui Zhou, Zora Zhiruo Wang, Nouha Dziri, Graham Neubig, Maarten Sap

机构 * Language Technologies Institute, Carnegie Mellon University(语言技术研究所,卡内基梅隆大学) Allen Institute for Artificial Intelligence(人工智能研究院)

专题命中 Agent评测 :agent(title,abstract);AI agent(title,abstract);agentic(abstract);分类 cs.AI

AI总结 OpenAgentSafety提出一个全面评估AI代理安全性的框架,通过真实工具和多任务测试揭示代理在现实世界中的安全漏洞,强调需要加强安全防护。

Comments 26 pages, 10 figures, Accepted at ICLR 2026 and IASEAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15212 2026-02-18 cs.AI 87%

Secure and Energy-Efficient Wireless Agentic AI Networks

安全且节能的无线代理人工智能网络

Yuanyan Song, Kezhi Wang, Xinmian Xu

机构 * Department of Computer Science, Brunel University of London(布鲁内尔大学计算机科学系) Nanjing University of Posts and Telecommunications(南京邮电大学)

专题命中 Agent评测 :agentic(title,abstract);agent(abstract);AI agent(abstract);workflow(abstract)

AI总结 本文提出了一种安全且节能的无线代理人工智能网络,通过动态分配代理和优化资源分配,提升服务质量并降低能耗。

Comments Submitted to journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15039 2026-02-18 hep-ex cs.AI 83%

GRACE: an Agentic AI for Particle Physics Experiment Design and Simulation

GRACE:一个用于粒子物理实验设计和模拟的智能体

Justin Hill, Hong Joo Ryoo

机构 * Data Science Institute, Columbia Engineering(数据科学研究院,哥伦比亚工程学院) Department of Physics and Astronomy, Johns Hopkins University(物理与天文学系,约翰霍普金斯大学)

专题命中 Agent评测 :agentic(title,abstract);agent(abstract);分类 cs.AI

AI总结 GRACE是一个用于粒子物理实验设计和模拟的智能体,通过自主探索设计修改提升物理性能,引入了新的基准测试和约束搜索方法。

Comments Both authors contributed equally. 43 pages, 12 figures, 6 tables, data can be found in https://github.com/just5034/GRACE_whitepaper_data

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11661 2026-02-18 cs.AI 79%

SR-Scientist: Scientific Equation Discovery With Agentic AI

SR-Scientist:基于代理AI的科学方程发现

Shijie Xia, Yuhan Sun, Pengfei Liu

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Innovation Institute(上海创新研究院) GAIR

专题命中 Agent评测 :agentic(title);agent(abstract);分类 cs.AI

AI总结 SR-Scientist通过代理AI实现科学方程发现,具备自主编写代码、分析数据、优化方程的能力,实验表明其在多个科学领域表现优异。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14296 2026-02-18 cs.MA 79%

From Agent Simulation to Social Simulator: A Comprehensive Review (Part 2)

从智能体模拟到社会模拟:全面综述(第二部分)

Xiao Xue, Deyu Zhou, Ming Zhang, Xiangning Yu, Fei-Yue Wang

专题命中 Agent评测 :agent(title,abstract)

AI总结 本文综述了从智能体模拟到社会模拟的方法,强调计算实验在探索系统复杂性因果机制中的作用,指出ABM的局限性及计算实验的优越性。

Comments This paper is Part II of a planned multi-part review series on "From Agent Simulation to Social Simulator". It is a self-contained article and can be read independently of Part I. Although the authors have previously published related work, this submission is not a revision or updated version of any earlier paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17812 2026-02-18 cs.CL cs.LG 79%

Can Multimodal LLMs Perform Time Series Anomaly Detection?

多模态大语言模型能否进行时间序列异常检测?

Xiongxiao Xu, Haoran Wang, Yueqing Liang, Philip S. Yu, Yue Zhao, Kai Shu

机构 * Illinois Institute of Technology(伊利诺伊理工学院) Emory University(埃默里大学) University of Illinois Chicago(伊利诺伊大学香槟分校) University of Southern California(南加州大学)

专题命中 Agent评测 :agent(abstract);planning(abstract);multi-agent(abstract);分类 cs.CL、cs.LG

AI总结 本研究提出基于多代理框架的TSAD-Agents,利用多模态大语言模型实现自动时间序列异常检测。

Comments ACM Web Conference 2026 (WWW'26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15270 2026-02-18 cs.AI 70%

Enhancing Diversity and Feasibility: Joint Population Synthesis from Multi-source Data Using Generative Models

增强多样性和可行性:利用生成模型从多源数据中联合合成人口

Farbod Abbasi, Zachary Patterson, Bilal Farooq

机构 * Concordia University(康科迪亚大学) Concordia Institute for Information Systems Engineering(康科迪亚信息系统工程研究所) Toronto Metropolitan University(多伦多 Metropolitan 大学)

专题命中 Agent评测 :agent(abstract);planning(abstract);分类 cs.AI

AI总结 本文提出利用WGAN与梯度惩罚联合生成多源数据,提升合成人口的多样性和可行性,实验显示其在召回率和精确率上均优于传统方法。

Comments 12 pages, 8 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08968 2026-02-18 cs.AI 57%

stable-worldmodel-v1: Reproducible World Modeling Research and Evaluation

stable-worldmodel-v1: 可复现的世界建模研究与评估

Lucas Maes, Quentin Le Lidec, Dan Haramati, Nassim Massaudi, Damien Scieur, Yann LeCun, Randall Balestriero

机构 * Mila & Université de Montréal(Mila与蒙特利尔大学) New York University(纽约大学) Brown University(布朗大学) Samsung SAIL(三星SAIL)

专题命中 Agent评测 :planning(abstract);分类 cs.AI

AI总结 stable-worldmodel-v1 提供了一个模块化、可复现的世界模型研究生态系统,支持标准化环境和持续学习研究,并用于评估DINO-WM的零样本鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15273 2026-02-18 cs.CY cs.CL 57%

FrameRef: A Framing Dataset and Simulation Testbed for Modeling Bounded Rational Information Health

FrameRef: 一个用于建模有限理性信息健康的框架数据集和仿真测试平台

Victor De Lima, Jiqun Liu, Grace Hui Yang

机构 * University of Oklahoma(俄克拉荷马大学)

专题命中 Agent评测 :agent(abstract);分类 cs.CL

AI总结 FrameRef通过仿真框架研究有限理性信息健康的动态,提供系统性数据集和方法论,影响人类判断和信息健康轨迹。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09424 2026-02-18 econ.TH 50%

Posterior-Separable Costs and Menu Preferences

后验可分离成本与菜单偏好

Henrique de Oliveira, Jeffrey Mensch

专题命中 Agent评测 :agent(abstract)

AI总结 本文研究了在贝叶斯说服框架下,通过理性不注意偏好和无知等价性公理,实现后验可分离成本的唯一超平面求解方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15062 2026-02-18 physics.soc-ph cs.MA 50%

Cooperative Game Theory Model for Sustainable UN Financing: Addressing Global Public Goods Provision

面向可持续联合国融资的协同博弈理论模型:解决全球公共物品供给问题

Labib Shami, Teddy Lazebnik

专题命中 Agent评测 :agent(abstract)

AI总结 本文提出一种协同博弈理论模型,通过个性化定价优化联合国融资,提升全球公共物品供给的可持续性和公平性。

详情

展开后加载摘要…

URL PDF HTML 收藏