arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2025-08-15 至 2025-08-15 共收录 11 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 11 篇

2507.23701 2025-08-15 cs.AI cs.CL 79%

TextQuests: How Good are LLMs at Text-Based Video Games?

Long Phan, Mantas Mazeika, Andy Zou, Dan Hendrycks

机构 * Center for AI Safety(人工智能安全中心) Carnegie Mellon University(卡内基梅隆大学) Gray Swan AI(灰天鹅AI)

专题命中 Agent评测 :agent(abstract);AI agent(abstract);tool use(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10769 2025-08-15 cs.AI cs.MM 57%

Modeling Human Responses to Multimodal AI Content

Zhiqi Shen, Shaojing Fan, Danni Xu, Terence Sim, Mohan Kankanhalli

专题命中 Agent评测 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10581 2025-08-15 cs.LG 57%

Technical Report: Facilitating the Adoption of Causal Inference Methods Through LLM-Empowered Co-Pilot

Jeroen Berrevoets, Julianna Piskorz, Robert Davis, Harry Amad, Jim Weatherall, Mihaela van der Schaar

机构 * University of Cambridge(剑桥大学) AstraZeneca(阿斯利康)

专题命中 Agent评测 :agentic(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.15729 2025-08-15 stat.ML cs.HC cs.LG 57%

A Two-Stage Learning-to-Defer Approach for Multi-Task Learning

Yannis Montreuil, Shu Heng Yeo, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi

机构 * School of Computing, National University of Singapore, Singapore(新加坡国立大学计算机学院) IRIT, Université de Toulouse, CNRS, Toulouse INP, Toulouse, France(图卢兹大学IRIT研究所,法国国家科学研究中心,图卢兹INP) Institute for Infocomm Research, Agency for Science, Technology and Research, Singapore(新加坡资讯与通信研究院,新加坡科技研究局)

专题命中 Agent评测 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15881 2025-08-15 cs.GT cs.LG 57%

Collaborative Mean Estimation Among Heterogeneous Strategic Agents: Individual Rationality, Fairness, and Truthful Contribution

Alex Clinton, Yiding Chen, Xiaojin Zhu, Kirthevasan Kandasamy

机构 * University of Wisconsin-Madison(威斯康星大学麦迪逊分校) Cornell University(康奈尔大学)

专题命中 Agent评测 :agent(abstract);分类 cs.LG

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10792 2025-08-15 cond-mat.mes-hall 50%

Reinforcement-Learning-Designed Field-Free Sub-Nanosecond Spin-Orbit-Torque Switching

Yuta Igarashi, Junji Fujimoto

专题命中 Agent评测 :agent(abstract)

Comments 5 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10466 2025-08-15 cs.SI cs.CY 50%

Online Homogeneity Can Emerge Without Filtering Algorithms or Homophily Preferences

Petter Törnberg

专题命中 Agent评测 :agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10401 2025-08-15 cs.IR 50%

Proxy Model-Guided Reinforcement Learning for Client Selection in Federated Recommendation

Liang Qu, Jianxin Li, Wei Yuan, Penghui Ruan, Yuhui Shi, Hongzhi Yin

专题命中 Agent评测 :agent(abstract)

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10185 2025-08-15 cs.CR cs.CY cs.NI 50%

An Architecture for Distributed Digital Identities in the Physical World

René Mayrhofer, Michael Roland, Tobias Höller, Philipp Hofer, Mario Lins

专题命中 Agent评测 :agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04485 2025-08-15 cs.GT 50%

Constant-Approximate and Constant-Strategyproof Two-Facility Location

Elijah Journey Fullerton, Zeyuan Hu, C. Gregory Plaxton

专题命中 Agent评测 :agent(abstract)

Comments Accepted at SAGT 2025. The latest version fixes minor typos and formatting issues

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06776 2025-08-15 cs.RO cs.SY eess.SY 50%

Chance-constrained Linear Quadratic Gaussian Games for Multi-robot Interaction under Uncertainty

Kai Ren, Giulio Salizzoni, Mustafa Emre Gürsoy, Maryam Kamgarpour

机构 * SYCAMORE Lab, École Polytechnique Fédérale de Lausanne (EPFL)(SYCAMORE实验室,瑞士联邦理工学院(EPFL))

专题命中 Agent评测 :agent(abstract)

Comments Published in IEEE Control Systems Letters

Journal ref IEEE Control Systems Letters, vol. 9, pp. 2061-2066, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏