arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2025-10-17 至 2025-10-17 共收录 7 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 软件智能体 7 篇

2509.06917 2025-10-17 cs.AI cs.CL cs.LG 85%

Paper2Agent: Reimagining Research Papers As Interactive and Reliable AI Agents

Jiacheng Miao, Joe R. Davis, Yaohui Zhang, Jonathan K. Pritchard, James Zou

机构 * Department of Genetics, Stanford University(遗传学系,斯坦福大学) Department of Biomedical Data Science, Stanford University(生物医学数据科学系,斯坦福大学) Department of Electrical Engineering, Stanford University(电气工程系,斯坦福大学) Department of Biology, Stanford University(生物学系,斯坦福大学) Department of Computer Science, Stanford University(计算机科学系,斯坦福大学)

专题命中 软件智能体 :AI agent(title,abstract);agent(abstract);分类 cs.AI、cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05179 2025-10-17 cs.CR cs.AI cs.LG 82%

Agentic Misalignment: How LLMs Could Be Insider Threats

Aengus Lynch, Benjamin Wright, Caleb Larson, Stuart J. Ritchie, Soren Mindermann, Evan Hubinger, Ethan Perez, Kevin Troy

机构 * University College London(伦敦大学学院) Anthropic MATS Mila

专题命中 软件智能体 :agentic(title,abstract);分类 cs.AI、cs.LG

Comments 20 pages, 12 figures. Code available at https://github.com/anthropic-experimental/agentic-misalignment

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08397 2025-10-17 eess.IV cs.AI cs.CV 80%

VoxelPrompt: A Vision Agent for End-to-End Medical Image Analysis

Andrew Hoopes, Neel Dey, Victor Ion Butoi, John V. Guttag, Adrian V. Dalca

机构 * Massachusetts Institute of Technology(麻省理工学院) Massachusetts General Hospital(麻省总医院) Harvard Medical School(哈佛医学院)

专题命中 软件智能体 :agent(title,abstract);分类 cs.AI

Comments 22 pages, vision-language agent, medical image analysis, neuroimage foundation model

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09714 2025-10-17 cs.CL cs.AI cs.LG 67%

All Code, No Thought: Current Language Models Struggle to Reason in Ciphered Language

Shiyuan Guo, Henry Sleight, Fabien Roger

机构 * Anthropic Fellows Program(Anthropic Fellow项目) Constellation Anthropic

专题命中 软件智能体 :AI agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments Version 2: updated related works section on LLM steganography

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13859 2025-10-17 cs.SE cs.AI 62%

Benchmarking Correctness and Security in Multi-Turn Code Generation

Ruchit Rawal, Jeffrey Yang Fan Chiang, Chihao Shen, Jeffery Siyuan Tian, Aastha Mahajan, Tom Goldstein, Yizheng Chen

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13709 2025-10-17 cs.AI cs.LG 62%

Training LLM Agents to Empower Humans

Evan Ellis, Vivek Myers, Jens Tuyls, Sergey Levine, Anca Dragan, Benjamin Eysenbach

专题命中 软件智能体 :AI agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01545 2025-10-17 cs.LG cs.AI cs.RO 62%

Predictive Preference Learning from Human Interventions

Haoyuan Cai, Zhenghao Peng, Bolei Zhou

机构 * Department of Computer Science, University of California, Los Angeles(计算机科学系,加州大学洛杉矶分校)

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2025 Spotlight. Project page: https://metadriverse.github.io/ppl

详情

展开后加载摘要…

URL PDF HTML 收藏