arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15801 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15801 篇

2301.01219 2023-01-04 cs.LG cs.AI cs.FL math.OC 62%

Task-Guided IRL in POMDPs that Scales

Franck Djeumou, Christian Ellis, Murat Cubuktepe, Craig Lennon, Ufuk Topcu

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Final submission to the Artificial Intelligence journal (Elsevier). arXiv admin note: substantial text overlap with arXiv:2105.14073

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.07434 2022-12-26 cs.LG cs.AI 62%

Parallel Automatic History Matching Algorithm Using Reinforcement Learning

Omar S. Alolayan, Abdullah O. Alomar, John R. Williams

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.08302 2022-12-19 cs.LG cs.AI 62%

Safe Evaluation For Offline Learning: Are We Ready To Deploy?

Hager Radi, Josiah P. Hanna, Peter Stone, Matthew E. Taylor

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2021 Workshop on Deployable Decision Making in Embodied Systems [Spotlight]

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.07594 2022-12-16 eess.SY cs.AI cs.LG cs.SY 62%

Driver Assistance Eco-driving and Transmission Control with Deep Reinforcement Learning

Lindsey Kerbel, Beshah Ayalew, Andrej Ivanco, Keith Loiselle

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Journal ref 2022 American Control Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.01737 2022-12-14 cs.CV cs.AI cs.LG 62%

RLogist: Fast Observation Strategy on Whole-slide Images with Deep Reinforcement Learning

Boxuan Zhao, Jun Zhang, Deheng Ye, Jian Cao, Xiao Han, Qiang Fu, Wei Yang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments accepted by AAAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.04359 2022-12-09 cs.RO cs.AI cs.LG cs.NE 62%

HERD: Continuous Human-to-Robot Evolution for Learning from Human Demonstration

Xingyu Liu, Deepak Pathak, Kris M. Kitani

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments CoRL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01955 2022-12-09 cs.LG cs.AI 62%

Learning Dynamic Abstract Representations for Sample-Efficient Reinforcement Learning

Mehdi Dadvar, Rashmeet Kaur Nayyar, Siddharth Srivastava

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.02715 2022-12-07 eess.SY cs.AI cs.LG cs.SY math.OC 62%

Efficient Learning of Voltage Control Strategies via Model-based Deep Reinforcement Learning

Ramij R. Hossain, Tianzhixi Yin, Yan Du, Renke Huang, Jie Tan, Wenhao Yu, Yuan Liu, Qiuhua Huang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.02960 2022-12-02 cs.CV cs.AI cs.LG 62%

Simple and Effective Synthesis of Indoor 3D Scenes

Jing Yu Koh, Harsh Agrawal, Dhruv Batra, Richard Tucker, Austin Waters, Honglak Lee, Yinfei Yang, Jason Baldridge, Peter Anderson

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments AAAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.17188 2022-12-01 cs.LG cs.AI 62%

Automated Play-Testing Through RL Based Human-Like Play-Styles Generation

Pierre Le Pelletier de Woillemont, Rémi Labory, Vincent Corruble

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Proceedings of the AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment, 18(1)

Journal ref Vol. 18 No. 1 (2022): Eighteenth AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.17132 2022-12-01 cs.LG cs.AI cs.GT cs.MA stat.ML 62%

Targets in Reinforcement Learning to solve Stackelberg Security Games

Saptarashmi Bandyopadhyay, Chenqi Zhu, Philip Daniel, Joshua Morrison, Ethan Shay, John Dickerson

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Appears in Proceedings of AAAI FSS-22 Symposium "Lessons Learned for Autonomous Assessment of Machine Abilities (LLAAMA)"

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.15837 2022-11-30 cs.LG cs.AI cs.CV cs.GT 62%

Survey on Self-Supervised Multimodal Representation Learning and Foundation Models

Sushil Thapa

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.15233 2022-11-29 cs.LG cs.AI cs.CV 62%

Tackling Visual Control via Multi-View Exploration Maximization

Mingqi Yuan, Xin Jin, Bo Li, Wenjun Zeng

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 21 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.08217 2022-11-28 cs.RO cs.AI cs.IT cs.LG math.IT 62%

PI-QT-Opt: Predictive Information Improves Multi-Task Robotic Reinforcement Learning at Scale

Kuang-Huei Lee, Ted Xiao, Adrian Li, Paul Wohlhart, Ian Fischer, Yao Lu

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments CoRL 2022. 21 pages, 9 figures. The supplementary video is available at https://kuanghuei.github.io/piqtopt

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.10861 2022-11-22 cs.LG cs.AI 62%

Efficient Meta Reinforcement Learning for Preference-based Fast Adaptation

Zhizhou Ren, Anji Liu, Yitao Liang, Jian Peng, Jianzhu Ma

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Thirty-sixth Conference on Neural Information Processing Systems (NeurIPS 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.02052 2022-11-18 cs.LG cs.AI 62%

Theta-Resonance: A Single-Step Reinforcement Learning Method for Design Space Exploration

Masood S. Mortazavi, Tiancheng Qin, Ning Yan

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.13855 2022-11-17 cs.CL cs.AI cs.IR 62%

Actionable Entities Recognition Benchmark for Interactive Fiction

Alexey Tikhonov, Ivan P. Yamshchikov

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.08272 2022-11-16 cs.LG cs.AI 62%

Low-Thrust Orbital Transfer using Dynamics-Agnostic Reinforcement Learning

Carlos M. Casas, Belen Carro, Antonio Sanchez-Esguevillas

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.08016 2022-11-16 cs.LG cs.AI 62%

Contextual Transformer for Offline Meta Reinforcement Learning

Runji Lin, Ye Li, Xidong Feng, Zhaowei Zhang, Xian Hong Wu Fung, Haifeng Zhang, Jun Wang, Yali Du, Yaodong Yang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted by Foundation Models for Decision Making Workshop at Neural Information Processing Systems, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.07941 2022-11-16 cs.RO cs.AI cs.LG 62%

Automatic Evaluation of Excavator Operators using Learned Reward Functions

Pranav Agarwal, Marek Teichmann, Sheldon Andrews, Samira Ebrahimi Kahou

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 11 pages, 5 figures, Accepted at Reinforcement Learning for Real Life (RL4RealLife) Workshop at NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.13490 2022-11-15 cs.LG cs.AI 62%

Towards Continual Reinforcement Learning: A Review and Perspectives

Khimya Khetarpal, Matthew Riemer, Irina Rish, Doina Precup

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Journal of Artificial Intelligence Research (JAIR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.06376 2022-11-14 cs.AI cs.LG 62%

Global and Local Analysis of Interestingness for Competency-Aware Deep Reinforcement Learning

Pedro Sequeira, Jesse Hostetler, Melinda Gervasio

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Appears in Proceedings of AAAI FSS-22 Symposium "Lessons Learned for Autonomous Assessment of Machine Abilities (LLAAMA)"

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.05612 2022-11-11 cs.AI cs.LG 62%

Power Grid Congestion Management via Topology Optimization with AlphaZero

Matthias Dorfer, Anton R. Fuxjäger, Kristian Kozak, Patrick M. Blies, Marcel Wasserer

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.08524 2022-11-11 cs.LG cs.AI cs.CY stat.ML 62%

So2Sat POP -- A Curated Benchmark Data Set for Population Estimation from Space on a Continental Scale

Sugandha Doda, Yuanyuan Wang, Matthias Kahl, Eike Jens Hoffmann, Kim Ouan, Hannes Taubenböck, Xiao Xiang Zhu

专题命中 Agent评测 :planning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.03958 2022-11-11 cs.LG cs.AI cs.NE 62%

Evolving Reinforcement Learning Algorithms

John D. Co-Reyes, Yingjie Miao, Daiyi Peng, Esteban Real, Sergey Levine, Quoc V. Le, Honglak Lee, Aleksandra Faust

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments ICLR 2021 Oral. See project website at https://sites.google.com/view/evolvingrl

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.03525 2022-11-08 cs.LG cs.AI cs.NA math.NA 62%

Dynamic weights enabled Physics-Informed Neural Network for simulating the mobility of Engineered Nano-particles in a contaminated aquifer

Shikhar Nilabh, Fidel Grandia

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 5 pages, 3 Figures, Conference paper accepted in NeurIPS 2022 Workshop: Tackling Climate Change with Machine Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.02193 2022-11-07 cs.NE cs.AI cs.LG cs.RO 62%

Benchmarking Quality-Diversity Algorithms on Neuroevolution for Reinforcement Learning

Manon Flageat, Bryan Lim, Luca Grillotti, Maxime Allard, Simón C. Smith, Antoine Cully

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted at GECCO Workshop on Quality Diversity Algorithm Benchmarks

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.00789 2022-11-03 cs.LG cs.AI 62%

Beyond Not-Forgetting: Continual Learning with Backward Knowledge Transfer

Sen Lin, Li Yang, Deliang Fan, Junshan Zhang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Published as a conference paper at NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.00458 2022-11-02 cs.RO cs.AI cs.LG cs.SY eess.SY 62%

CPG-RL: Learning Central Pattern Generators for Quadruped Locomotion

Guillaume Bellegarda, Auke Ijspeert

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted for IEEE Robotics and Automation Letters, September 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.14876 2022-10-27 quant-ph cs.AI cs.ET cs.LG cs.NE 62%

Quantum deep recurrent reinforcement learning

Samuel Yen-Chi Chen

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏