arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2025-08-20 至 2025-08-20 共收录 51 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 多智能体 12 篇

2508.13247 2025-08-20 cs.MA cs.AI 70%

Goal-Directedness is in the Eye of the Beholder

Nina Rajcic, Anders Søgaard

机构 * University of Copenhagen, Department of Philosophy(哥本哈根大学哲学系) University of Copenhagen, Department of Computer Science(哥本哈根大学计算机科学系)

专题命中 多智能体 :agent(abstract);multi-agent(abstract);分类 cs.AI

Comments Submitted to Conference and Workshop on Neural Information Processing Systems 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 工作流自动化 4 篇

2505.18637 2025-08-20 cs.IT math.IT 71%

Neural Coding Is Not Always Semantic: Toward the Standardized Coding Workflow in Semantic Communications

Hai-Long Qin, Jincheng Dai, Sixian Wang, Xiaoqi Qin, Shuo Shao, Kai Niu, Wenjun Xu, Ping Zhang

专题命中 工作流自动化 :workflow(title)

Comments Accepted by IEEE COMSTD, project page: https://qin-jingyun.github.io/SemCod/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13828 2025-08-20 cs.AI 57%

Revisiting RAG Ensemble: A Theoretical and Mechanistic Analysis of Multi-RAG System Collaboration

Yifei Chen, Guanting Dong, Yutao Zhu, Zhicheng Dou

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学中关村人工智能学院)

专题命中 工作流自动化 :agentic(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13429 2025-08-20 q-fin.CP cs.AI 57%

AlphaX: An AI-Based Value Investing Strategy for the Brazilian Stock Market

Paulo André Lima de Castro

机构 * Artificial Intelligence Applied to Finance Research Group- AIAF(人工智能应用于金融研究组) Aeronautics Institute of Technology -ITA(航空技术研究所)

专题命中 工作流自动化 :autonomous agent(abstract);分类 cs.AI

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13738 2025-08-20 cs.GR 50%

Eliminating Rasterization: Direct Vector Floor Plan Generation with DiffPlanner

Shidong Wang, Renato Pajarola

专题命中 工作流自动化 :workflow(abstract)

Comments accepted to IEEE Transactions on Visualization and Computer Graphics

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 软件智能体 1 篇

2507.01457 2025-08-20 cs.LG cs.AI cs.SE 67%

Tensor Program Optimization for the RISC-V Vector Extension Using Probabilistic Programs

Federico Nicolas Peccia, Frederik Haxel, Oliver Bringmann

机构 * FZI Research Center for Information Technology, University of Tübingen Germany(弗赖堡研究所信息科技研究中心、图宾根大学德国)

专题命中 软件智能体 :workflow(abstract);分类 cs.AI、cs.LG、cs.SE

Comments 9 pages, 10 figures, 2 algorithms

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 记忆与上下文管理 6 篇

2508.11120 2025-08-20 cs.CL 92%

Towards Reliable Multi-Agent Systems for Marketing Applications via Reflection, Memory, and Planning

Lorenzo Jaime Yu Flores, Junyi Shen, Goodman Gu

专题命中 记忆与上下文管理 :agent(title,abstract);planning(title,abstract);multi-agent(title,abstract);AI agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13171 2025-08-20 cs.AI cs.CL 62%

Cognitive Workspace: Active Memory Management for LLMs -- An Empirical Study of Functional Infinite Context

Tao An

机构 * Hawaii Pacific University(夏威夷太平洋大学)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.CL

Comments 13 pages, 1 figure, code available at https://github.com/tao-hpu/cognitive-workspace

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19263 2025-08-20 cs.AI 57%

Modeling Uncertainty: Constraint-Based Belief States in Imperfect-Information Games

Achille Morenville, Éric Piette

机构 * ICTEAM, UCLouvain(ICTEAM,乌尔特拉-努夫大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08815 2025-08-20 physics.flu-dyn cs.AI 57%

Deep reinforcement learning for tracking a moving target in jellyfish-like swimming

Yihao Chen, Yue Yang

机构 * State Key Laboratory for Turbulence and Complex Systems, College of Engineering, Peking University(湍流与复杂系统国家重点实验室,工程学院,北京大学) HEDPS-CAPT, Peking University(HEDPS-CAPT,北京大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 22pages,14 figures

Journal ref J. Fluid Mech. 1017 (2025) A18

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13388 2025-08-20 cs.HC 50%

Data Work in Memory Institutions: Why and How Information Professionals Use Wikidata

Riya Sinha, Amelia Acker, Hanlin Li

专题命中 记忆与上下文管理 :workflow(abstract)

Comments 27 pages, 1 figure, 2 tables. Accepted to PACM HCI (CSCW 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19803 2025-08-20 cs.RO 50%

Integrating emotional intelligence, memory architecture, and gestures to achieve empathetic humanoid robot interaction in an educational setting

Fuze Sun, Lingyu Li, Shixiangyue Meng, Xiaoming Teng, Terry R. Payne, Paul Craig

机构 * University of Liverpool(利物浦大学) Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学)

专题命中 记忆与上下文管理 :agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

5. Agent评测 8 篇

2508.05668 2025-08-20 cs.IR cs.AI cs.CL 73%

A Survey of LLM-based Deep Search Agents: Paradigm, Optimization, Evaluation, and Challenges

Yunjia Xi, Jianghao Lin, Yongzhao Xiao, Zheli Zhou, Rong Shan, Te Gao, Jiachen Zhu, Weiwen Liu, Yong Yu, Weinan Zhang

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 Agent评测 :agent(abstract);planning(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13186 2025-08-20 cs.CL cs.AI cs.CV 73%

MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents

Shilong Li, Xingyuan Bu, Wenjie Wang, Jiaheng Liu, Jun Dong, Haoyang He, Hao Lu, Haozhe Zhang, Chenchen Jing, Zhen Li, Chuanhao Li, Jiayi Tian, Chenchen Zhang, Tianhao Peng, Yancheng He, Jihao Gu, Yuanxing Zhang, Jian Yang, Ge Zhang, Wenhao Huang, Wangchunshu Zhou, Zhaoxiang Zhang, Ruizhe Ding, Shilei Wen

机构 * Nanjing University(南京大学) Zhejiang University(浙江大学)

专题命中 Agent评测 :AI agent(abstract);tool use(abstract);分类 cs.AI、cs.CL

Comments The first two authors contribute equally, 26 pages, repo at https://github.com/MMBrowseComp/MM-BrowseComp

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13530 2025-08-20 cs.AI 57%

CrafterDojo: A Suite of Foundation Models for Building Open-Ended Embodied Agents in Crafter

Junyeong Park, Hyeonseo Cho, Sungjin Ahn

机构 * CrafterDojo: A Suite of Foundation Models for Building Open-Ended Embodied Agents in Crafter(CrafterDojo:构建开放性具身智能体的基础模型集合)

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13413 2025-08-20 cs.HC cs.SE 57%

Large Language Models as Visualization Agents for Immersive Binary Reverse Engineering

Dennis Brown, Samuel Mulder

专题命中 Agent评测 :agent(abstract);分类 cs.SE

Comments Accepted to IEEE VISSOFT 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13992 2025-08-20 eess.AS cs.SD 50%

MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Sonal Kumar, Šimon Sedláček, Vaibhavi Lokegaonkar, Fernando López, Wenyi Yu, Nishit Anand, Hyeonggon Ryu, Lichang Chen, Maxim Plička, Miroslav Hlaváček, William Fineas Ellingwood, Sathvik Udupa, Siyuan Hou, Allison Ferner, Sara Barahona, Cecilia Bolaños, Satish Rahi, Laura Herrera-Alarcón, Satvik Dixit, Siddhi Patil, Soham Deshmukh, Lasha Koroshinadze, Yao Liu, Leibny Paola Garcia Perera, Eleni Zanou, Themos Stafylakis, Joon Son Chung, David Harwath, Chao Zhang, Dinesh Manocha, Alicia Lozano-Diez, Santosh Kesiraju, Sreyan Ghosh, Ramani Duraiswami

专题命中 Agent评测 :AI agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13432 2025-08-20 cs.GT cs.DM 50%

Fair Division Among Couples and Small Groups

Paul Gölz, Hannane Yaghoubizade

专题命中 Agent评测 :agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17741 2025-08-20 cs.CY 50%

Experimental Evidence for the Propagation and Preservation of Machine Discoveries in Human Populations

Levin Brinkmann, Thomas F. Eisenmann, Anne-Marie Nussberger, Maxime Derex, Sara Bonati, Valerii Chirkov, Iyad Rahwan

专题命中 Agent评测 :agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.20396 2025-08-20 cs.GT 50%

Mechanism Design for Congested Facility Location

Cheng Peng, Houyu Zhou

专题命中 Agent评测 :agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 其他Agent 1 篇

2508.13960 2025-08-20 cs.GT cs.AI 57%

A Mechanism for Mutual Fairness in Cooperative Games with Replicable Resources -- Extended Version

Björn Filter, Ralf Möller, Özgür Lütfü Özçep

机构 * Institute for Humanities-Centered AI (CHAI), University of Hamburg, Germany(人文中心人工智能研究所(CHAI),汉堡大学)

专题命中 其他Agent :agentic(abstract);分类 cs.AI

Comments This paper is the extended version of a paper accepted at the European Conference on Artificial Intelligence 2025 (ECAI 2025), providing the proof of the main theorem in the appendix

详情

展开后加载摘要…

URL PDF HTML 收藏