arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15819 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15819 篇

2007.10442 2020-07-22 cs.AI cs.LG 62%

Unlocking the Potential of Deep Counterfactual Value Networks

Ryan Zarick, Bryan Pellegrino, Noam Brown, Caleb Banister

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 11 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.08229 2020-07-17 cs.LG cs.AI stat.ML 62%

Mixture of Step Returns in Bootstrapped DQN

Po-Han Chiang, Hsuan-Kung Yang, Zhang-Wei Hong, Chun-Yi Lee

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.12442 2020-07-14 cs.CL cs.AI 62%

Open-Domain Conversational Agents: Current Progress, Open Problems, and Future Directions

Stephen Roller, Y-Lan Boureau, Jason Weston, Antoine Bordes, Emily Dinan, Angela Fan, David Gunning, Da Ju, Margaret Li, Spencer Poff, Pratik Ringshia, Kurt Shuster, Eric Michael Smith, Arthur Szlam, Jack Urbanek, Mary Williamson

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.00089 2020-07-14 cs.LG cs.AI cs.RO stat.ML 62%

Safe, Efficient, and Comfortable Velocity Control based on Reinforcement Learning for Autonomous Driving

Meixin Zhu, Yinhai Wang, Ziyuan Pu, Jingyun Hu, Xuesong Wang, Ruimin Ke

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Under the first-round revision for transportation research part c

Journal ref Transportation Research Part C: Emerging Technologies 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.04017 2020-07-10 cs.LG cs.AI stat.ML 62%

Provable Self-Play Algorithms for Competitive Reinforcement Learning

Yu Bai, Chi Jin

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Appearing at ICML 2020. Fixed typos from v1

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.01820 2020-07-09 q-fin.TR cs.AI cs.LG stat.ML 62%

Robust Market Making via Adversarial Reinforcement Learning

Thomas Spooner, Rahul Savani

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 7 pages, 3 figures; IJCAI-PRICAI '20 Conference Proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.03964 2020-07-09 math.OC cs.AI cs.LG 62%

Responsive Safety in Reinforcement Learning by PID Lagrangian Methods

Adam Stooke, Joshua Achiam, Pieter Abbeel

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments ICML 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.12399 2020-07-06 cs.LG cs.AI 62%

Reinforcement Learning Generalization with Surprise Minimization

Jerry Zikun Chen

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Inductive biases, invariances and generalization in RL Workshop, ICML 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.15762 2020-06-30 cs.AI cs.LG stat.ML 62%

Empirically Verifying Hypotheses Using Reinforcement Learning

Kenneth Marino, Rob Fergus, Arthur Szlam, Abhinav Gupta

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.09026 2020-06-23 cs.LG cs.AI stat.ML 62%

Bandits with Temporal Stochastic Constraints

Priyank Agrawal, Theja Tulabandhula

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments An extended abstract appeared in the 4th Multi-disciplinary Conference on Reinforcement Learning and Decision Making (RLDM 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.09497 2020-06-18 cs.LG cs.AI stat.ML 62%

Task-agnostic Exploration in Reinforcement Learning

Xuezhou Zhang, Yuzhe ma, Adish Singla

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.07113 2020-06-15 cs.HC cs.AI cs.LG stat.ML 62%

Large-scale Hybrid Approach for Predicting User Satisfaction with Conversational Agents

Dookun Park, Hao Yuan, Dongmin Kim, Yinglei Zhang, Matsoukas Spyros, Young-Bum Kim, Ruhi Sarikaya, Edward Guo, Yuan Ling, Kevin Quinn, Pham Hung, Benjamin Yao, Sungjin Lee

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.06026 2020-06-12 cs.CL cs.AI 62%

Report from the NSF Future Directions Workshop, Toward User-Oriented Agents: Research Directions and Challenges

Maxine Eskenazi, Tiancheng Zhao

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments Final report of the NSF Future Directions Workshop, Toward User-Oriented Agents: Research Directions and Challenges

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.13186 2020-05-28 cs.SE cs.AI 62%

Beware the evolving 'intelligent' web service! An integration architecture tactic to guard AI-first components

Alex Cummaudo, Scott Barnett, Rajesh Vasa, John Grundy, Mohamed Abdelrazek

专题命中 Agent评测 :planning(abstract);分类 cs.AI、cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.07594 2020-05-20 quant-ph cond-mat.mes-hall cs.AI cs.ET cs.LG 62%

Measurement-based adaptation protocol with quantum reinforcement learning in a Rigetti quantum computer

J. Olivares-Sánchez, J. Casanova, E. Solano, L. Lamata

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Journal ref Quantum Reports 2, 293 (2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.02627 2020-05-07 cs.LG cs.AI cs.NE stat.ML 62%

Unity: A General Platform for Intelligent Agents

Arthur Juliani, Vincent-Pierre Berges, Ervin Teng, Andrew Cohen, Jonathan Harper, Chris Elion, Chris Goy, Yuan Gao, Hunter Henry, Marwan Mattar, Danny Lange

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.00496 2020-04-17 cs.LG cs.AI stat.ML 62%

Uncertainty-Based Out-of-Distribution Classification in Deep Reinforcement Learning

Andreas Sedlmeier, Thomas Gabor, Thomy Phan, Lenz Belzner, Claudia Linnhoff-Popien

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments arXiv admin note: text overlap with arXiv:1901.02219

Journal ref Proceedings of the 12th International Conference on Agents and Artificial Intelligence - Volume 2: ICAART, 2020, ISBN 978-989-758-395-7, pages 522-529

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.07093 2020-04-16 cs.LG cs.CL stat.ML 62%

lamBERT: Language and Action Learning Using Multimodal BERT

Kazuki Miyazawa, Tatsuya Aoki, Takato Horii, Takayuki Nagai

专题命中 Agent评测 :agent(abstract);分类 cs.CL、cs.LG

Comments 8 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.05940 2020-04-14 q-fin.TR cs.AI cs.LG 62%

A Deep Reinforcement Learning Framework for Continuous Intraday Market Bidding

Ioannis Boukas, Damien Ernst, Thibaut Théate, Adrien Bolland, Alexandre Huynen, Martin Buchwald, Christelle Wynants, Bertrand Cornélusse

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.13056 2020-04-13 cs.AI cs.LG 62%

Distributed Soft Actor-Critic with Multivariate Reward Representation and Knowledge Distillation

Dmitry Akimov

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.04647 2020-04-10 cs.CR cs.AI cs.LG 62%

Adversarial Genetic Programming for Cyber Security: A Rising Application Domain Where GP Matters

Una-May O'Reilly, Jamal Toutouh, Marcos Pertierra, Daniel Prado Sanchez, Dennis Garcia, Anthony Erb Luogo, Jonathan Kelly, Erik Hemberg

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.00801 2020-04-03 cs.LG cs.AI cs.RO 62%

Exploration of Reinforcement Learning for Event Camera using Car-like Robots

Riku Arakawa, Shintaro Shiba

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.11939 2020-04-01 cs.LG cs.AI stat.ML 62%

MERL: Multi-Head Reinforcement Learning

Yannis Flet-Berliac, Philippe Preux

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Deep Reinforcement Learning Workshop, NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.09969 2020-03-25 cs.CL cs.AI cs.IR 62%

The JDDC Corpus: A Large-Scale Multi-Turn Chinese Dialogue Dataset for E-commerce Customer Service

Meng Chen, Ruixue Liu, Lei Shen, Shaozu Yuan, Jingyan Zhou, Youzheng Wu, Xiaodong He, Bowen Zhou

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments This paper is accepted by LREC 2020 (International Conference on Language Resources and Evaluation )

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.10019 2020-03-18 cs.LG cs.AI stat.ML 62%

Adversarial Active Exploration for Inverse Dynamics Model Learning

Zhang-Wei Hong, Tsu-Jui Fu, Tzu-Yun Shann, Yi-Hsiang Chang, Chun-Yi Lee

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Published as a conference paper at CoRL 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.04310 2020-03-11 eess.SP cs.AI cs.LG cs.MA cs.SY eess.SY stat.ML 62%

Advancing Renewable Electricity Consumption With Reinforcement Learning

Filip Tolovski

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments To be presented at the Workshop on Tackling Climate Change with Machine Learning at ICLR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.06349 2020-03-03 cs.LG cs.AI stat.ML 62%

Slice-based Learning: A Programming Model for Residual Learning in Critical Data Slices

Vincent S. Chen, Sen Wu, Zhenzhen Weng, Alexander Ratner, Christopher Ré

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.10695 2020-02-26 cs.CL cs.CV cs.LG 62%

Multimodal Transformer with Pointer Network for the DSTC8 AVSD Challenge

Hung Le, Nancy F. Chen

专题命中 Agent评测 :agent(abstract);分类 cs.CL、cs.LG

Comments Accepted at DSTC Workshop at AAAI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.03568 2020-02-17 cs.LG cs.AI stat.ML 62%

Behaviour Suite for Reinforcement Learning

Ian Osband, Yotam Doron, Matteo Hessel, John Aslanides, Eren Sezener, Andre Saraiva, Katrina McKinney, Tor Lattimore, Csaba Szepesvari, Satinder Singh, Benjamin Van Roy, Richard Sutton, David Silver, Hado Van Hasselt

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.08277 2020-02-11 cs.LG cs.AI 62%

Federated Deep Reinforcement Learning

Hankz Hankui Zhuo, Wenfeng Feng, Yufeng Lin, Qian Xu, Qiang Yang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏