arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 3244 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 软件智能体 3244 篇

1501.05779 2015-01-26 cs.CY 78%

An alternative use of the NetLogo modeling environment, where the student thinks and acts like an Agent, in order to teach concepts of Ecology

Aristotelis Gkiolmas, Anthimos Chalkidis, Maria Papaconstantinou, Zafar Iqbal, Constantine Skordoulis

专题命中 软件智能体 :agent(title,abstract)

Comments 9th Pan-Hellenic Conference with International Participation: "ICT's in Education" (HCICTE 2014) 3rd-5th October 2014, University of Crete, Rethymno, Greece

Journal ref Proceedings of the 9th Pan-Hellenic Conference with International Participation: "ICT's in Education" (HCICTE 2014) 3rd-5th October 2014, University of Crete, Rethymno, Greece, pp. 379-386

详情

展开后加载摘要…

URL PDF HTML 收藏
1307.5319 2014-12-30 q-fin.GN physics.soc-ph 78%

Tipping points in macroeconomic Agent-Based models

Stanislao Gualdi, Marco Tarzia, Francesco Zamponi, Jean-Philippe Bouchaud

专题命中 软件智能体 :agent(title,abstract)

Comments 42 pages, 8 figures. Important revisions with respect to v2. Final version, to appear in a special issue of the Journal of Economic Dynamics and Control dedicated to the CRISIS project ( http://www.crisis-economics.eu )

Journal ref Journal of Economic Dynamics & Control 50, 29-61 (2015)

详情

展开后加载摘要…

URL PDF HTML 收藏
1303.7377 2013-04-01 cs.MA 78%

Evaluating Reputation Systems for Agent Mediated e-Commerce

Vibha Gaur, Neeraj Kumar Sharma, Punam Bedi

专题命中 软件智能体 :agent(title,abstract)

Comments 5 pages. arXiv admin note: text overlap with arXiv:1110.3961

详情

展开后加载摘要…

URL PDF HTML 收藏
1207.2946 2012-12-04 q-fin.TR q-fin.CP q-fin.ST 78%

Microscopic understanding of heavy-tailed return distributions in an agent-based model

Thilo A. Schmitt, Rudi Schäfer, Michael C. Münnix, Thomas Guhr

专题命中 软件智能体 :agent(title,abstract)

Journal ref Europhysics Letters 100, 38005 (2012)

详情

展开后加载摘要…

URL PDF HTML 收藏
1006.0408 2010-10-14 q-bio.QM cs.MA physics.bio-ph 78%

A Mathematical Framework for Agent Based Models of Complex Biological Networks

Franziska Hinkelmann, David Murrugarra, Abdul Salam Jarrah, Reinhard Laubenbacher

专题命中 软件智能体 :agent(title,abstract)

Comments To appear in Bulletin of Mathematical Biology

详情

展开后加载摘要…

URL PDF HTML 收藏
0801.3846 2009-12-01 astro-ph 78%

An Autonomous Adaptive Scheduling Agent for Period Searching

Eric S. Saunders, Tim Naylor, Alasdair Allan

专题命中 软件智能体 :agent(title,abstract)

Comments 5 pages, 2 figures, to appear in proceedings of Hot-wiring the Transient Universe (HTU) 2007, Astronomische Nachrichten, March 2008

详情

展开后加载摘要…

URL PDF HTML 收藏
cs/0306060 2009-11-30 cs.DC 78%

DIRAC - Distributed Infrastructure with Remote Agent Control

N. Brook, A. Bogdanchikov, A. Buckley, J. Closier, U. Egede, M. Frank, D. Galli, M. Gandelman, V. Garonne, C. Gaspar, R. Graciani Diaz, K. Harrison, E. van Herwijnen, A. Khan, S. Klous, I. Korolko, G. Kuznetsov, F. Loverre, U. Marconi, J. P. Palacios, G. N. Patrick, A. Pickford, S. Ponce, V. Romanovski, J. J. Saborido, M. Schmelling, A. Soroko, A. Tsaregorodtsev, V. Vagnoni, A. Washbrook

专题命中 软件智能体 :agent(title,abstract)

Comments Talk from the 2003 Computing in High Energy and Nuclear Physics (CHEP03), La Jolla, Ca, USA, March 2003, 8 pages, Word, 5 figures. PSN TUAT006

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.03392 2026-08-21 cs.SE 版本更新 77%

Self-Evolving Coding Agents

自进化编码智能体

Hao Zhou, Haichuan Hu, Ye Shang, Quanjun Zhang

专题命中 软件智能体 :agent(abstract);workflow(abstract);agentic(abstract);分类 cs.SE

AI总结 本综述系统梳理自进化编码智能体领域,明确其与传统智能体的区别,提出分类法,分析该领域的挑战,为设计更优智能系统奠定基础。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.18280 2026-08-20 cs.SE cs.AI cs.CL cs.LG 新提交 77%

What Makes Software Issue Resolution Tasks Difficult for Agents?

什么因素导致软件问题解决任务对智能体而言难度较高?

Ebtesam Al-Haque, Brittany Johnson

专题命中 软件智能体 :agent(abstract);agentic(abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 该研究提出测量框架,基于CoderForge-Preview数据集的实证分析,发现软件问题解决任务难度可通过静态特征预测,核心驱动因素为补丁碎片化与代码库规模,为构建难度可控的智能体评估基准奠定基础。

Comments To appear in ESEM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.18050 2026-08-19 cs.AI 新提交 77%

StagedWorkspace: A Versioned Workspace for Knowledge-Work Agents

StagedWorkspace:面向知识工作智能体的版本化工作空间

Yining Hua, Hongbin Na, Yifan Zhou, Akshay Kalose, Cyrus Ayubcha, Levi Lian

机构 * Harvard University(哈佛大学) Raycaster AI(雷caster人工智能公司) University of Technology Sydney(悉尼科技大学) University of Washington(华盛顿大学) Stanford University(斯坦福大学)

专题命中 软件智能体 :agent(abstract,abstract_cn);AI agent(abstract);分类 cs.AI

AI总结 该研究针对知识工作智能体的版本化工作空间问题,提出StagedWorkspace方案,通过绑定解析记录与原生文件内容哈希提升性能,在OfficeQA Pro等数据集上取得显著优于基准的效果,为相关基准构建提供了新方向。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.17550 2026-08-19 cs.CV cs.CL 新提交 77%

Code as Representation: A Compilable Parsing Paradigm for Academic Documents

代码作为表示:学术文档的可编译解析范式

Rihui Jin, Jun Wang, chengyuan zhu, Liang Mingyu, Yue Gao, Li Yunxuan, Kuicai Dong, Guilin Qi, Lin Ren, Yongrui Chen, Xinbang Dai, Jiaqi Li, Tongtong Wu, Gholamreza Haffari

机构 * Southeast University(东南大学) Nanyang Technological University(南洋理工大学) Nanjing University(南京大学)

专题命中 软件智能体 :agent(abstract);agentic(abstract);multi-agent(abstract);分类 cs.CL

AI总结 针对学术PDF难以被机器处理的问题,提出CADP可编译解析范式,构建CADP-Bench基准并测试SOTA MLLMs,发现前沿模型仍难生成高保真可执行重构,该基准已开放供研究。

Comments Accepted by ACM MM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.05884 2026-08-07 cs.CR cs.CL 新提交 77%

The Vulnerability With No CVE: Managing Persistent Gaps Between Mandate and Authority in AI Coding Agents

无CVE的漏洞:管理AI编码智能体中授权与权限之间的持续差距

Shayell Aharon Salomon Amir Shaked Matan Noga

专题命中 软件智能体 :agentic(abstract,abstract_cn);agent(abstract);分类 cs.CL

AI总结 该研究针对AI编码智能体中授权与权限的持续差距,提出智能体姿态漏洞(APV)的概念,区分其与相关风险,提供定义、模式、生命周期等内容,为相关安全管理提供可操作框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.25890 2026-07-29 cs.AI 新提交 77%

Distributing Security Controls Through Harness Engineering

通过工具工程分发安全控制

William Robert Gore

专题命中 软件智能体 :agent(abstract);autonomous agent(abstract);agentic(abstract);分类 cs.AI

AI总结 研究商业AI编码代理安全控制分发问题,通过分阶段测试方法,利用SHarD工具在多种代理配置上测试,证明三类安全控制可通过单个命令嵌入分发,效果与直接安装商业代理相同,还提出相关框架特征及后续研究问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.02544 2026-07-28 cs.SE 版本更新 77%

Developer Experience with AI Coding Agents: HTTP Behavioral Signatures in Documentation Portals

与AI代码代理的开发者体验:文档门户中的HTTP行为签名

Oleksii Borysenko

专题命中 软件智能体 :agent(abstract,abstract_cn);AI agent(abstract);分类 cs.SE

AI总结 本文研究AI代码代理对技术文档访问的影响,发现HTTP行为特征在文档门户中具有可识别性,提出改进开发者文档设计和反馈机制的实践方案。

Comments 9 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.21217 2026-07-24 cs.AI 新提交 77%

ICAE-Bench: Evaluating Coding Agents as Interactive Project Builders

ICAE-Bench:将编码智能体评估为交互式项目构建者

Zhongyuan Peng, Dan Huang, Chuyu Zhang, Caijun Xu, Changyi Xiao, Shibo Hong, David Lo, Lin Qiu, Xuezhi Cao, Jiyuan He, Yixin Cao

机构 * Meituan(美团)

专题命中 软件智能体 :agent(abstract);tool use(abstract);planning(abstract);分类 cs.AI

AI总结 研究针对编码智能体在交互式项目构建中作用转变,现有基准未跟上的问题,提出ICAE-Bench基准。从模糊需求出发,经自动化用户智能体模拟动态范式,引入避免需求模糊性、确保用户模拟质量、公平评估开放式仓库的关键设计。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.18859 2026-07-22 cs.AI 新提交 77%

PhoenixRepair: Rethinking Repair Strategy Exploration in Software Agents

PhoenixRepair:重新思考软件代理中的修复策略探索

Tianyue Jiang, Yanlin Wang, Xin He, Daya Guo, Jiachi Chen, Ming Wen, Ensheng Shi, Xilin Liu, Yuchi Ma, Guanbin Li

机构 * Zhejiang University(浙江大学) Huazhong University of Science and Technology(华中科技大学) Huawei CodeArts Model Team(华为代码艺术模型团队)

专题命中 软件智能体 :agent(abstract,abstract_cn);multi-agent(abstract);分类 cs.AI

AI总结 研究针对现有软件代理修复策略探索不足的问题,提出PhoenixRepair多代理框架,通过多位置采样、迭代反思优化等扩大搜索空间,实验表明该框架在解决率和故障定位精度上有提升,实现了7.8%的相对改进及76.0%的最高解决率Pass@1。

Comments 14 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.10621 2026-07-14 cs.SE 新提交 77%

WebDesignIter: Co-Evolving Design Knowledge for Repository-Level Front-End Code Generation

WebDesignIter:用于仓库级前端代码生成的协同进化设计知识

Zheng Pei, Mingwei Liu, Zhenxi Chen, Zihao Wang, Yanlin Wang

专题命中 软件智能体 :agent(abstract,abstract_cn);planning(abstract);分类 cs.SE

AI总结 研究针对前端开发仓库级代码生成问题,提出WebDesignIter框架,通过持久知识图谱融合设计知识与仓库结构,分两阶段工作,实验证明其相比基线和通用编码代理有优势,凸显设计知识对仓库级代码生成的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21233 2026-06-29 cs.AI 版本更新 77%

Just Ask: Curious Code Agents Reveal System Prompts in Frontier LLMs

Just Ask: 好奇的代码代理揭示前沿大语言模型中的系统提示

Xiang Zheng, Yutao Wu, Hanxun Huang, Yige Li, Xingjun Ma, Bo Li, Yu-Gang Jiang, Cong Wang

专题命中 软件智能体 :agent(abstract);tool use(abstract);agentic(abstract);分类 cs.AI

AI总结 提出JustAsk框架,利用代码代理的自主交互能力,通过在线探索策略自动提取大语言模型的隐藏系统提示,揭示了系统提示作为新兴安全漏洞。

Comments Accepted at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15385 2026-06-16 cs.AI 新提交 77%

Reward Hacking in Language Model Agents: Revisiting AI Safety Gridworlds

语言模型智能体中的奖励黑客:重新审视AI安全网格世界

Ömer Veysel Çağatan, Xuandong Zhao

机构 * KUIS AI Center, Koç University(科奇大学KUIS人工智能中心) University of California, Berkeley(加州大学伯克利分校)

专题命中 软件智能体 :agent(abstract,abstract_cn);agentic(abstract);分类 cs.AI

AI总结 本研究将AI安全网格世界框架改编为文本评估套件,发现语言模型在零样本下出现规范博弈,通过直接奖励优化扩大观察与隐藏奖励差距,且标准缓解措施无效。

Comments 28 pages, 16 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.11732 2026-06-05 cs.IR cs.CL cs.MA cs.MM 77%

AgentDisCo: Towards Disentanglement and Collaboration in Open-ended Deep Research Agents

AgentDisCo: 向开放深度研究代理中的解耦与协作迈进

Jiarui Jin, Zexuan Yan, Shijian Wang, Wenxiang Jiao, Yuan Lu

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 软件智能体 :agent(abstract);workflow(abstract);agentic(abstract);分类 cs.CL

AI总结 本文提出AgentDisCo,一种解耦且协作的代理架构,将深度研究视为信息探索与利用之间的对抗优化问题。通过批评代理评估生成的草稿并优化搜索查询,生成代理检索更新结果并修订草稿,最终生成综合报告。该框架通过元优化 harness 支持手工和自动发现的设计策略,并利用强大的代码生成代理自完善。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06310 2026-06-04 cs.SE 77%

Trustworthy AI Software Engineers

可信赖的AI软件工程师

Aldeida Aleti, Baishakhi Ray, Rashina Hoda, Simin Chen

专题命中 软件智能体 :agent(abstract);AI agent(abstract);agentic(abstract);分类 cs.SE

AI总结 本文探讨AI代理作为软件工程师的信任问题,提出以证据为中心的检查方法,从技术质量、透明度、认知谦逊和社会伦理等维度定义可信赖性。

Comments The first three authors contributed equally to this work

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.03394 2026-06-03 cs.SE 77%

Human-AI Collaboration and the Transformation of Software Engineering Work

人机协作与软件工程工作的转型

Mamdouh Alenezi

专题命中 软件智能体 :agentic(abstract,abstract_cn);agent(abstract);分类 cs.SE

AI总结 本文通过结构化综合分析方法,研究了生成式AI和智能体AI如何将软件工程从以人类编写代码为中心转变为以指导、验证和管理自主及半自主系统为中心,并提出了一个包含技术、认知、社会技术、治理和组织五个维度的未来工程师能力框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18552 2026-06-03 cs.SE cs.AI cs.CL cs.LG 77%

Toward Training Superintelligent Software Agents through Self-Play SWE-RL

通过自我对弈SWE-RL训练超级智能软件代理

Yuxiang Wei, Zhiqing Sun, Emily McMilin, Jonas Gehring, David Zhang, Gabriel Synnaeve, Daniel Fried, Lingming Zhang, Sida Wang

机构 * Meta FAIR University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Meta TBD Lab(Meta TBD 实验室) Carnegie Mellon University(卡内基梅隆大学)

专题命中 软件智能体 :agent(abstract);agentic(abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 提出自我对弈SWE-RL(SSR)方法,通过强化学习在自对弈环境中训练单一LLM代理,使其在无需人工标注问题或测试的情况下,在真实代码库中迭代注入和修复软件缺陷,在SWE-bench基准上实现显著自我改进并超越人类数据基线。

Comments Accepted to ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.28617 2026-05-28 cs.AI cs.PL 77%

LACUNA: Safe Agents as Recursive Program Holes

LACUNA: 作为递归程序空洞的安全智能体

Yaoyu Zhao, Yichen Xu, Oliver Bračevac, Cao Nguyen Pham, Frank Zhengqing Wu, Martin Odersky

机构 * EPFL(苏黎世联邦理工学院)

专题命中 软件智能体 :agent(abstract,abstract_cn);planning(abstract);分类 cs.AI

AI总结 提出LACUNA编程模型,通过类型化调用和编译时检查,让LLM智能体以递归程序空洞的方式安全地编写代码,实现表达性与安全性的统一。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20160 2026-05-26 cs.SE 77%

How do Agents Refactor: An Empirical Study

智能体如何重构:一项实证研究

Lukas Ottenhof, Daniel Penner, Abram Hindle, Thibaud Lutellier

专题命中 软件智能体 :agent(abstract,abstract_cn);agentic(abstract);分类 cs.SE

AI总结 通过对比86个Java项目中智能体与开发者的重构拉取请求,发现智能体重构以注释修改为主,而Cursor是唯一显著增加代码味道的模型。

Comments Accepted for publication in 23rd International Mining Software Repositories Conference (MSR 2026) : 5 pages, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24812 2026-05-26 cs.AI 77%

CoRe-Code: Collaborative Reinforcement Learning for Code Generation

CoRe-Code:面向代码生成的协作式强化学习

Zhihao Dou, Qinjian Zhao, Zhongwei Wan, Xiaoyu Xia, Sumon Biswas

机构 * The Ohio State University(俄亥俄州立大学) Royal Melbourne Institute of Technology(皇家墨尔本理工学院)

专题命中 软件智能体 :agent(abstract);planning(abstract);multi-agent(abstract);分类 cs.AI

AI总结 提出CoRe-Code框架,通过规划器-编码器范式和基于GRPO的协作感知强化学习,增强多智能体间的协调与专业化,提升代码生成的准确性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24659 2026-05-26 cs.LG 77%

IterInject: Indirect Prompt Injection Against LLM Agents via Feedback-Guided Iterative Optimization

IterInject: 通过反馈引导的迭代优化实现对LLM智能体的间接提示注入

Zixuan Chen, Jiaxiang Chen, Li Luo, Ke Xu, Xiaoxiang Huang, Tanfeng Sun, Xinghao Jiang

机构 * Shanghai Jiao Tong University(上海交通大学) The University of Hong Kong(香港大学)

专题命中 软件智能体 :agent(abstract);tool use(abstract);planning(abstract);分类 cs.LG

AI总结 提出IterInject框架,通过规则诊断器和LLM优化器迭代优化对抗载荷,实现对LLM智能体的间接提示注入攻击,在多个基准和实际系统中显著优于现有方法,并揭示了注意力介导的阈值机制。

Comments Submitted to EMNLP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18270 2026-05-12 cs.SE 77%

Can Old Tests Do New Tricks for Resolving SWE Issues?

旧测试能否为解决软件工程问题带来新用途?

Yang Chen, Toufique Ahmed, Reyhaneh Jabbarvand, Martin Hirzel

专题命中 软件智能体 :agent(abstract,abstract_cn);agentic(abstract);分类 cs.SE

AI总结 本文提出TestPrune,通过重用回归测试自动最小化测试套件,提升问题复现和修复效率,实验证明在多个基准测试中显著提高复现和修复率。

Comments Accepted to the main technical track of the Symposium on the Foundations of Software Engineering (FSE), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.08013 2026-05-11 cs.AI 77%

Learning CLI Agents with Structured Action Credit under Selective Observation

基于选择性观察的结构化行动信用学习CLI代理

Haoyang Su, Ying Wen

机构 * Fudan University(复旦大学) Shanghai Innovation Institute(上海创新研究院) Shanghai Jiao Tong University(上海交通大学)

专题命中 软件智能体 :agent(abstract,abstract_cn);agentic(abstract);分类 cs.AI

AI总结 本文研究了CLI代理在选择性观察和行动信用分配中的瓶颈,提出σ-Reveal和A³方法,构建ShellOps数据集以评估CLI任务学习。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23674 2026-04-28 cs.AI 77%

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work

Vibe Medicine:通过人机协作重新定义生物医学研究

Zihao Wu, Steven Xu, Bowen Chen, Shaowen Wan, Yiwei Li, Wei Ruan, Yanjun Lyu, Siyuan Li, Dajiang Zhu, Tianming Liu, Lin Zhao

机构 * School of Computing, University of Georgia(佐治亚大学计算学院) Department of Biomedical Engineering, New Jersey Institute of Technology(新泽西理工学院生物医学工程系) Department of Computer Science and Engineering, University of Texas at Arlington(德克萨斯大学阿灵顿分校计算机科学与工程系)

专题命中 软件智能体 :agent(abstract,abstract_cn);AI agent(abstract);分类 cs.AI

AI总结 本文提出Vibe Medicine,通过自然语言指导AI代理执行复杂生物医学流程,解决多领域数据整合与分析难题,提升研究效率与公平性。

详情

展开后加载摘要…

URL PDF HTML 收藏