arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-01-22 至 2026-01-22 共收录 91 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 多智能体 20 篇

2601.10958 2026-01-22 cs.IT cs.NI math.IT 67%

Fundamental Limits of Quantum Semantic Communication via Sheaf Cohomology

量子语义通信的fundamental limits via sheaf cohomology

Christo Kurisummoottil Thomas, Mingzhe Chen

专题命中 多智能体 :agent(abstract);multi-agent(abstract)

AI总结 本文提出基于sheaf cohomology的量子语义通信框架,揭示语义模糊性的信息论限制,并通过量子纠缠和情境性降低上同调障碍,为自主系统提供新的通信理论基础。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 工作流自动化 9 篇

2601.14544 2026-01-22 cs.CR 82%

AI Agents vs. Human Investigators: Balancing Automation, Security, and Expertise in Cyber Forensic Analysis

AI代理与人类调查员:在自动化、安全与专业知识之间平衡网络取证分析

Sneha Sudhakaran, Naresh Kshetri

专题命中 工作流自动化 :AI agent(title,abstract);agent(abstract)

AI总结 本研究比较了AI代理与人类调查员在网络取证分析中的有效性,揭示AI的局限性及人类监督的重要性。

Comments 10 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15059 2026-01-22 cs.AI cs.SY eess.SY 79%

The Responsibility Vacuum: Organizational Failure in Scaled Agent Systems

责任真空:规模化智能体系统中的组织失效

Oleg Romanchuk, Roman Bondar

专题命中 工作流自动化 :agent(title,abstract);分类 cs.AI

AI总结 研究揭示规模化智能体系统中因决策生成吞吐量超出人类验证能力导致的责任归属问题,指出自动化加剧而非缓解责任真空现象,呼吁重新设计决策边界以缓解系统失效。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15247 2026-01-22 cs.CL 70%

Taxonomy-Aligned Risk Extraction from 10-K Filings with Autonomous Improvement Using LLMs

基于LLM的10-K文件风险提取与分类体系的自适应改进

Rian Dolphin, Joe Dursun, Jarrett Blankenship, Katie Adams, Quinton Pike

专题命中 工作流自动化 :agent(abstract);AI agent(abstract);分类 cs.CL

AI总结 本文提出了一种基于LLM的10-K文件风险提取方法,结合自动分类维护机制,实现对分类体系的持续改进和风险特征的准确捕捉。

Comments 4 figures, 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14790 2026-01-22 cs.AI 70%

CI4A: Semantic Component Interfaces for Agents Empowering Web Automation

CI4A: 为网络自动化赋能的语义组件接口

Zhi Qiu, Jiazheng Sun, Chenxiao Xia, Jun Zheng, Xin Peng

机构 * School of Cyberspace Science and Technology, Beijing Institute of Technology(电子信息学院,北京理工大学) College of Computer Science and Artificial Intelligence, Fudan University(计算机科学与人工智能学院,复旦大学)

专题命中 工作流自动化 :agent(abstract);planning(abstract);分类 cs.AI

AI总结 CI4A 通过构建语义组件接口,提升代理在网页自动化中的表现,实现 86.3% 的任务成功率。

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01951 2026-01-22 cs.LG 57%

NeuroClean: A Generalized Machine-Learning Approach to Neural Time-Series Conditioning

NeuroClean:一种通用的机器学习神经时间序列条件化方法

Manuel A. Hernandez Alonso, Michael Depass, Stephan Quessy, Ali Falaki, Soraya Rahimi, Numa Dancause, Ignasi Cos

专题命中 工作流自动化 :workflow(abstract);分类 cs.LG

AI总结 NeuroClean是一种通用的机器学习方法,用于神经时间序列的条件化处理,通过无监督算法有效去除伪影和噪声,提升信号质量和分类准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15253 2026-01-22 quant-ph physics.chem-ph physics.comp-ph 50%

QDK/Chemistry: A Modular Toolkit for Quantum Chemistry Applications

QDK/Chemistry:量子化学应用的模块化工具包

Nathan A. Baker, Brian Bilodeau, Chi Chen, Yingrong Chen, Marco Eckhoff, Alexandra Efimovskaya, Piero Gasparotto, Puck van Gerwen, Rushi Gong, Kevin Hoang, Zahra Hooshmand, Andrew J. Jenkins, Conrad S. N. Johnston, Run R. Li, Jiashu Liang, Hongbin Liu, Alexis Mills, Maximilian Mörchen, George Nishibuchi, Chong Sun, Bill Ticehurst, Matthias Troyer, Jan P. Unsleber, Stefan Wernli, David B. Williams-Young, Boqin Zhang

专题命中 工作流自动化 :workflow(abstract)

AI总结 QDK/Chemistry提供模块化架构,整合量子与经典计算,支持跨来源方法组合,助力可重复量子化学实验

Comments 32 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14574 2026-01-22 q-bio.BM 50%

De novo design of protein binders targeting the human sweet taste receptor as potential sweet proteins

从头设计靶向人类甜味受体的蛋白质结合物作为潜在甜蛋白

Saisai Ding, Yi Zhang

专题命中 工作流自动化 :workflow(abstract)

AI总结 本研究通过从头设计蛋白质结合物,模拟天然甜蛋白的功能特性,为开发新一代基于蛋白质的甜味剂提供了计算框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14529 2026-01-22 physics.flu-dyn physics.geo-ph 50%

From Columns to Heaps: Dimensionless Similarity with PSD-Distributed Damköhler Numbers and Dual-Porosity Flow

从柱体到堆体:基于PSD分布的无量纲相似性与双孔隙流体

Juan J. Segura

专题命中 工作流自动化 :workflow(abstract)

AI总结 本研究提出一个无量纲框架,用于比较不同尺度的反应多孔流系统,通过PSD分布和双孔隙结构分析,揭示了扩散控制浸出对PSD尾部和双孔隙结构的敏感性,并确定了确保相似性的关键无量纲组。

Comments 26 pages, 2 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.13967 2026-01-22 quant-ph 50%

A no free lunch theorem for untrained quantum circuits in machine learning

机器学习中未训练量子电路的无免费午餐定理

Steven Herbert

专题命中 工作流自动化 :workflow(abstract)

AI总结 本文提出机器学习中未训练量子电路的无免费午餐定理,指出未训练量子电路在平均意义上无优势,质疑其在提升性能上的理论依据。

Comments 12 pages; clearer presentation of some key results, and minor re-ordering

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 软件智能体 3 篇

2601.15195 2026-01-22 cs.SE cs.AI 84%

Where Do AI Coding Agents Fail? An Empirical Study of Failed Agentic Pull Requests in GitHub

AI 编程代理在哪里失败?对 GitHub 中失败的代理拉取请求的实证研究

Ramtin Ehsani, Sakshi Pathak, Shriya Rawal, Abdullah Al Mujahid, Mia Mohammad Imran, Preetha Chatterjee

机构 * Drexel University(德雷塞尔大学) Missouri University of Science and Technology(密苏里科技大学)

专题命中 软件智能体 :agentic(title,abstract);agent(abstract);分类 cs.AI、cs.SE

AI总结 研究分析了 GitHub 上 33k 个代理生成 PRs 的失败原因,发现文档、CI 等任务合并成功率高,而性能和修复任务失败率高,揭示了代理协作中的关键社会技术因素。

Comments Accepted at International Mining Software Repositories Conference (MSR 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14914 2026-01-22 cs.CL 77%

CodeDelegator: Mitigating Context Pollution via Role Separation in Code-as-Action Agents

CodeDelegator: 通过角色分离缓解上下文污染在代码作为行动代理中

Tianxiang Fei, Cheng Chen, Yue Pan, Mao Zheng, Mingyang Song

机构 * Large Language Model Department, Tencent(腾讯大语言模型部门)

专题命中 软件智能体 :agent(abstract);planning(abstract);multi-agent(abstract);分类 cs.CL

AI总结 CodeDelegator通过角色分离缓解上下文污染,利用多代理框架实现战略规划与实现的分离,提升代码作为行动代理的长期性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14523 2026-01-22 cs.AI cs.LG 73%

Large Language Model-Powered Evolutionary Code Optimization on a Phylogenetic Tree

基于进化树的大型语言模型驱动的代码优化

Leyi Zhao, Weijie Huang, Yitong Guo, Jiang Bian, Chenghong Wang, Xuhong Zhang

机构 * Department of Computer Science(计算机科学系) Luddy School of Informatics(信息学院) Indiana University(印第安纳大学)

专题命中 软件智能体 :agent(abstract);workflow(abstract);分类 cs.AI、cs.LG

AI总结 PhyloEvolve利用大型语言模型和进化算法优化GPU科学计算算法,通过进化树结构提升优化效率和可重复性。

详情

展开后加载摘要…

URL PDF HTML 收藏

4. GUI与网页智能体 2 篇

2601.14649 2026-01-22 cs.RO 50%

Spatially Generalizable Mobile Manipulation via Adaptive Experience Selection and Dynamic Imagination

基于自适应经验选择和动态想象的通用移动操作

Ping Zhong, Liangbai Liu, Bolei Chen, Tao Wu, Jiazhi Xia, Chaoxu Mu, Jianxin Wang

机构 * School of Computer Science and Engineering, Central South University(中南大学计算机科学与工程学院) Xiangjiang Laboratory(湘江实验室) School of Artificial Intelligence, Anhui University(安徽大学人工智能学院)

专题命中 GUI与网页智能体 :planning(abstract)

AI总结 本文提出基于自适应经验选择和动态想象的移动操作方法,提升技能学习效率和空间泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14307 2026-01-22 physics.soc-ph cs.CY 50%

Assessing the livability within the 15-minute city concept based on mobile phone data

基于移动电话数据评估15分钟城市概念内的宜居性

Tianqi Wang, Teemu Jama, Henrikki Tenkanen

专题命中 GUI与网页智能体 :planning(abstract)

AI总结 本研究利用移动电话数据评估15分钟城市概念中的宜居性,发现步行性和综合宜居指数与人类活动模式相关,但关系随时间波动,需更全面的城市发展方法。

Comments 29 pages

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 记忆与上下文管理 4 篇

2601.14602 2026-01-22 cs.CV 67%

3D Space as a Scratchpad for Editable Text-to-Image Generation

3D空间作为可编辑的文本到图像生成的草稿纸

Oindrila Saha, Vojtech Krs, Radomir Mech, Subhransu Maji, Matheus Gadelha, Kevin Blackburn-Matzen

机构 * University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校) Adobe Research(Adobe研究)

专题命中 记忆与上下文管理 :planning(abstract);agentic(abstract)

AI总结 本文提出一种基于3D空间推理的文本到图像生成方法,通过可编辑的3D网格实现空间一致性,提升图像生成的精度和可控性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15086 2026-01-22 cs.LG cs.AI 62%

Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning

在强化学习中,仅依靠记忆保持不足以掌握记忆任务

Oleg Shchendrigin, Egor Cherepanov, Alexey K. Kovalev, Aleksandr I. Panov

机构 * Innopolis University(因诺普利斯大学) Cognitive AI Systems Lab(认知人工智能系统实验室)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出一个基准测试持续记忆更新,发现经典循环模型在记忆重写任务中表现更优,强调需平衡稳定保留与适应更新的记忆机制。

Comments 11 pages, 6 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14798 2026-01-22 cs.LG cs.CL cs.CY 62%

Reflecting in the Reflection: Integrating a Socratic Questioning Framework into Automated AI-Based Question Generation

反思中的反思:将苏格拉底质疑框架整合到基于自动的AI问题生成中

Ondřej Holub, Essi Ryymin, Rodrigo Alves

机构 * Czech Technical University in Prague(捷克技术大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

AI总结 本文提出一种基于苏格拉底质疑框架的双代理模型,通过多轮对话生成高质量反思问题,提升教学效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14589 2026-01-22 cs.HC cs.AI cs.CL cs.CY 62%

Designing KRIYA: An AI Companion for Wellbeing Self-Reflection

设计KRIYA:一种用于幸福感自我反思的AI伴侣

Shanshan Zhu, Wenxuan Song, Jiayue Melissa Shi, Dong Whi Yoo, Karthik S. Bhat, Koustuv Saha

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Indiana University Indianapolis(印第安纳大学印第安纳波利斯分校) Drexel University(德雷塞尔大学)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.CL

AI总结 KRIYA是一种通过自我反思功能帮助用户理解个人幸福感数据的AI伴侣,旨在减少表现焦虑,提升用户对健康数据的反思性理解。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. Agent评测 12 篇

2512.24565 2026-01-22 cs.AI 88%

MCPAgentBench: A Real-world Task Benchmark for Evaluating LLM Agent MCP Tool Use

MCPAgentBench: 一个用于评估LLM代理MCP工具使用的现实任务基准

Wenrui Liu, Zixiang Liu, Elsie Dai, Wenhan Yu, Lei Yu, Tong Yang, Jinjun Han, Hong Gao

机构 * Peking University(北京大学) ZTE(中兴通讯)

专题命中 Agent评测 :agent(title);tool use(title);autonomous agent(abstract);tool-use(abstract)

AI总结 MCPAgentBench通过现实任务和动态沙盒环境评估LLM代理在复杂工具调用中的能力差异,提供开源代码以促进工具使用能力研究。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15153 2026-01-22 cs.AI 83%

How to Build AI Agents by Augmenting LLMs with Codified Human Expert Domain Knowledge? A Software Engineering Framework

如何通过将编码化的人类专家领域知识与大语言模型结合来构建AI代理?一种软件工程框架

Choro Ulan uulu, Mikhail Kulyabin, Iris Fuhrmann, Jan Joosten, Nuno Miguel Martins Pacheco, Filippos Petridis, Rebecca Johnson, Jan Bosch, Helena Holmström Olsson

机构 * Department of Computer Science and Engineering, Chalmers University of Technology(计算机科学与工程系,查尔姆斯理工大学) Department of Mathematics and Computer Science, Eindhoven University of Technology(数学与计算机科学系,埃因霍温理工大学) Department of Computer Science and Media Technology, Malmö University(计算机科学与媒体技术系,马尔默大学)

专题命中 Agent评测 :AI agent(title,abstract);agent(abstract);分类 cs.AI

AI总结 本文提出一种软件工程框架,通过增强大语言模型与编码化专家知识,构建能自主生成可视化内容的AI代理,实现非专家在专业领域内达到专家水平的成果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15034 2026-01-22 cs.HC cs.AI 79%

Visual and Cognitive Demands of a Large Language Model-Powered In-vehicle Conversational Agent

基于大型语言模型的车载对话代理的视觉与认知需求

Chris Monk, Allegra Ayala, Christine S. P. Yu, Gregory M. Fitch, Dara Gruber

机构 * Exponent, Inc.(Exponent公司) Google, Inc.(Google公司)

专题命中 Agent评测 :agent(title,abstract);分类 cs.AI

AI总结 本研究评估了基于大型语言模型的车载对话代理在驾驶中的视觉与认知需求,发现其与免提通话在认知负荷上相当,且视觉需求较低,支持其在驾驶环境中的安全应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15273 2026-01-22 q-bio.QM 78%

How high-resolution agent-based models can improve fundamental insights in tissue development and cell culturing methods

高分辨率基于代理的模型如何改进组织发育和细胞培养方法的基础见解

Paul Van Liedekerke, Jiří Pešek, Kevin Alessandri, Dirk Drasdo

专题命中 Agent评测 :agent(title,abstract)

AI总结 本文探讨了可变形细胞模型在组织发育和细胞培养方法中的应用,通过高分辨率模拟提升生物和生物技术问题的定量分析能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15258 2026-01-22 cs.GT 78%

Distributed Agent-Constrained Truthful Facility Location

分布式代理约束下的诚实设施定位

Argyrios Deligkas, Panagiotis Kanellopoulos, Alexandros A. Voudouris

专题命中 Agent评测 :agent(title,abstract)

AI总结 该研究提出了一种分布式设施定位机制,通过两阶段选择代表位置,确保代理无法通过策略性报告获益,并分析了两种成本变体下的近似比界。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15144 2026-01-22 q-bio.PE 78%

Modification speed and radius of higher-order interactions alter the oscillatory dynamics in an agent-based model

高阶相互作用的修改速度和半径改变代理模型中的振荡动力学

Thomas Van Giel, Hanna Jaspaert, Aisling J. Daly, Bernard De Baets, Jan M. Baetens

专题命中 Agent评测 :agent(title,abstract)

AI总结 本研究探讨了高阶相互作用在代理模型中对物种振荡动力学的影响,发现其修改速度和半径显著影响系统稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.01090 2026-01-22 cs.MA cs.AI cs.CY 77%

Harm in AI-Driven Societies: An Audit of Toxicity Adoption on Chirper.ai

AI驱动社会中的危害:对Chirper.ai上毒性采用的审计

Erica Coppolillo, Luca Luceri, Emilio Ferrara

机构 * University of Southern California, Los Angeles, California(美国南加州大学) University of Calabria, Rende, Italy(意大利卡拉布里亚大学)

专题命中 Agent评测 :agent(abstract);AI agent(abstract);autonomous agent(abstract);分类 cs.AI

AI总结 研究通过分析Chirper.ai上AI代理的毒性行为,揭示了暴露于有害内容如何影响代理行为,并提出通过监控毒性暴露来减轻有害行为的风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15016 2026-01-22 cs.CV 75%

LiViBench: An Omnimodal Benchmark for Interactive Livestream Video Understanding

LiViBench:面向交互式直播视频理解的多模态基准测试

Xiaodong Wang, Langling Huang, Zhirong Wu, Xu Zhao, Teng Xu, Xuhong Xia, Peixi Peng

专题命中 Agent评测 :agent(abstract);workflow(abstract);multi-agent(abstract)

AI总结 LiViBench是首个面向交互式直播视频的多模态基准测试,通过定制化两阶段指令微调和视频到评论检索模块,提升模型对直播视频的理解能力,并在多个基准测试中取得优异成绩。

Comments AAAI 2026 Main Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14606 2026-01-22 cs.CR 71%

An LLM Agent-based Framework for Whaling Countermeasures

基于LLM代理的鲸鱼攻击防御框架

Daisuke Miyamoto, Takuji Iimura, Narushige Michishita

专题命中 Agent评测 :agent(title)

AI总结 本研究提出基于LLM代理的鲸鱼攻击防御框架,通过构建个性化防御资料和分析电子邮件,提升对高权威目标的防御能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14235 2026-01-22 astro-ph.IM astro-ph.CO cs.AI cs.LG stat.ML 62%

Opportunities in AI/ML for the Rubin LSST Dark Energy Science Collaboration

人工智能/机器学习在Rubin LSST暗能量科学合作中的机遇

LSST Dark Energy Science Collaboration, Eric Aubourg, Camille Avestruz, Matthew R. Becker, Biswajit Biswas, Rahul Biswas, Boris Bolliet, Adam S. Bolton, Clecio R. Bom, Raphaël Bonnet-Guerrini, Alexandre Boucaud, Jean-Eric Campagne, Chihway Chang, Aleksandra Ćiprijanović, Johann Cohen-Tanugi, Michael W. Coughlin, John Franklin Crenshaw, Juan C. Cuevas-Tello, Juan de Vicente, Seth W. Digel, Steven Dillmann, Mariano Javier de León Dominguez Romero, Alex Drlica-Wagner, Sydney Erickson, Alexander T. Gagliano, Christos Georgiou, Aritra Ghosh, Matthew Grayling, Kirill A. Grishin, Alan Heavens, Lindsay R. House, Mustapha Ishak, Wassim Kabalan, Arun Kannawadi, François Lanusse, C. Danielle Leonard, Pierre-François Léget, Michelle Lochner, Yao-Yuan Mao, Peter Melchior, Grant Merz, Martin Millon, Anais Möller, Gautham Narayan, Yuuki Omori, Hiranya Peiris, Laurence Perreault-Levasseur, Andrés A. Plazas Malagón, Nesar Ramachandra, Benjamin Remy, Cécile Roucelle, Jaime Ruiz-Zapatero, Stefan Schuldt, Ignacio Sevilla-Noarbe, Ved G. Shah, Tjitske Starkenburg, Stephen Thorp, Laura Toribio San Cipriano, Tilman Tröster, Roberto Trotta, Padma Venkatraman, Amanda Wasserman, Tim White, Justine Zeghal, Tianqing Zhang, Yuanyuan Zhang

机构 * Université Paris Cité, CNRS, CEA, Astroparticule et Cosmologie, F-75013 Paris, France Department of Physics, University of Michigan, Ann Arbor, MI 48109, USA Leinweber Institute of Theoretical Physics, University of Michigan, Ann Arbor, MI 48109, USA Argonne National Laboratory, 9700 South Cass Avenue, Lemont, IL 60439, USA Cavendish Astrophysics, University of Cambridge, Madingley Road, Cambridge CB3 0HA, UK Kavli Institute for Cosmology, University of Cambridge, Madingley Road, Cambridge CB3 0HA, UK SLAC National Accelerator Laboratory, Menlo Park, CA 94025, USA Department of Computer Science, University of Milan, Milan, Italy Université Paris Cité, CNRS, Astroparticule et Cosmologie, F-75013 Paris, France Université Paris-Saclay, CNRS/IN2P3, IJCLab, 91405 Orsay, France Department of Astronomy Astrophysics, University of Chicago, Chicago, IL 60637, USA Kavli Institute for Cosmological Physics, University of Chicago, Chicago, IL 60637, USA NSF-Simons AI Institute for the Sky (SkAI), 172 E. Chestnut St., Chicago, IL 60611, USA Fermi National Accelerator Laboratory, P.O. Box 500, Batavia, IL 60510, USA Universit\'e Clermont-Auvergne, CNRS, LPCA, 63000 Clermont-Ferrand, France Kavli Institute for Particle Astrophysics Cosmology, Stanford University, Stanford, CA 94305, USA Department of Physics, Stanford University, 382 Via Pueblo Mall, Stanford, CA 94305, USA Engineering Faculty, Universidad Autonoma de San Luis Potosi, Zona Universitaria, San Luis Potosi, 78290, Mexico Stanford Artificial Intelligence Laboratory, Stanford University, Stanford, CA 94305, USA Kavli Institute of Cosmological Physics, University of Chicago, Chicago, IL 60637, USA The NSF AI Institute for Artificial Intelligence Center for Astrophysics Harvard \& Smithsonian, 60 Garden Street, Cambridge, MA 02138, USA Department of Physics Kavli Institute for Astrophysics Space Research, Massachusetts Institute of Technology, Cambridge, MA 02139, USA Institut de Física d'Altes Energies (IFAE), The Barcelona Institute of Science Institute of Astronomy Kavli Institute for Cosmology, University of Cambridge, Madingley Road, Cambridge, CB3 0HA, UK Imperial Centre for Inference Cosmology (ICIC), Imperial College London, Blackett Laboratory, Prince Consort Road, London SW7 2AZ, UK Data Science Institute, The University of Chicago, Chicago, IL 60615, USA Department of Physics, The University of Texas at Dallas, Richardson, TX 75080, USA Department of Physics, Duke University, Durham, NC 27708, USA Université Paris-Saclay, Université Paris Cité, CEA, CNRS, AIM, F-91191 Gif-sur-Yvette, France School of Mathematics, Statistics Physics, Newcastle University, Newcastle upon Tyne, NE1 7RU, United Kingdom Department of Astrophysical Sciences, Princeton University, Princeton, NJ 08544, USA Astronomy, University of the Western Cape, Bellville, Cape Town, 7535, South Africa Astronomy, University of Utah, Salt Lake City, UT 84112, USA Department of Astrophysical Sciences, Princeton University, Peyton Hall, Princeton, NJ 08544, USA Department of Astronomy, University of Illinois Urbana Champaign, 1002 W. Green St., Urbana, IL, 61801, USA Institute for Particle Physics Astrophysics, ETH Zürich, Wolfgang-Pauli-Strasse 27, CH-8093 Zurich, Switzerland Swinburne University of Technology, Hawthorn, Victoria 3122, Australia Ciela - Montr\'eal Institute for Astrophysical Data Analysis Mila - Quebec Artificial Intelligence Institute, Montréal, QC H2S 3H1, Canada Advanced Research Computing Centre, University College London, 90 High Holborn, London WC1V 6LJ, UK Finnish Centre for Astronomy with ESO (FINCA), University of Turku, FI-20014 Turku, Finland Department of Physics, P.O. Box 64, University of Helsinki, FI-00014 Helsinki, Finland Astronomy, Northwestern University, Evanston, IL, USA Center for Interdisciplinary Exploration Research in Astrophysics, Northwestern University, Evanston, IL, USA Scientific Data Science, International School for Advanced Study, Via Bonomea 265, I-34136 Trieste, Italy Department of Statistics, University of Michigan, Ann Arbor, MI 48109, USA PITT PACC, University of Pittsburgh, Pittsburgh, PA 15260, USA NSF NOIRLab, 950 N. Cherry Ave., Tucson, AZ 85719, USA

专题命中 Agent评测 :agentic(abstract);分类 cs.AI、cs.LG

AI总结 本文探讨了AI/ML在LSST暗能量科学合作中的应用机遇,强调了大规模贝叶斯推断、物理指导方法和主动学习等关键方法学优先事项,并讨论了新兴技术在重塑工作流程中的潜力。

Comments 84 pages. This is v1.0 of the DESC's white paper on AI/ML, a collaboration document that is being made public but which is not planned for submission to a journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.12787 2026-01-22 cs.LG cs.AI 62%

Impartial Games: A Challenge for Reinforcement Learning

impartial games: 一种对强化学习的挑战

Bei Zhou, Søren Riis

机构 * Imperial College London(帝国理工学院伦敦分校) Queen Mary University of London(女王玛丽大学)

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

AI总结 本文研究了AlphaZero风格强化学习在impartial games中的局限性,指出其在学习抽象数学原理如奇偶性时存在表示瓶颈,需发展新型算法以实现专家级AI。

详情

展开后加载摘要…

URL PDF HTML 收藏