arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 5118 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 工具调用 5118 篇

1905.12197 2019-05-30 cs.RO cs.AI cs.LG 62%

LeTS-Drive: Driving in a Crowd by Learning from Tree Search

Panpan Cai, Yuanfu Luo, Aseem Saxena, David Hsu, Wee Sun Lee

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Journal ref Proc. Robotics: Science & Systems (RSS), 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.12255 2019-04-30 cs.RO cs.AI cs.LG 62%

Non-myopic Planetary Exploration Combining In Situ and Remote Measurements

Suhit Kodgule, Alberto Candela, David Wettergreen

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments Preprint. Under review for IROS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.05538 2019-04-12 cs.RO cs.AI cs.LG 62%

Improvisation through Physical Understanding: Using Novel Objects as Tools with Visual Foresight

Annie Xie, Frederik Ebert, Sergey Levine, Chelsea Finn

专题命中 工具调用 :tool use(abstract);分类 cs.AI、cs.LG

Comments Videos available at https://sites.google.com/view/gvf-tool

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.02477 2019-04-11 cs.LG cs.AI cs.PL stat.ML 62%

Programmatically Interpretable Reinforcement Learning

Abhinav Verma, Vijayaraghavan Murali, Rishabh Singh, Pushmeet Kohli, Swarat Chaudhuri

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments Published at The 35th International Conference on Machine Learning (ICML 2018)

Journal ref PMLR 80:5045-5054

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.01843 2019-02-19 cs.LG cs.AI stat.ML 62%

How to Combine Tree-Search Methods in Reinforcement Learning

Yonathan Efroni, Gal Dalal, Bruno Scherrer, Shie Mannor

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments AAAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.06176 2018-12-18 cs.AI cs.CL 62%

Bootstrapping Conversational Agents With Weak Supervision

Neil Mallinar, Abhishek Shah, Rajendra Ugrani, Ayush Gupta, Manikandan Gurusankar, Tin Kam Ho, Q. Vera Liao, Yunfeng Zhang, Rachel K. E. Bellamy, Robert Yates, Chris Desmarais, Blake McGregor

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL

Comments 6 pages, 3 figures, 1 table, Accepted for publication in IAAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.01672 2018-12-10 cs.LG cs.AI stat.ML 62%

Ranked Reward: Enabling Self-Play Reinforcement Learning for Combinatorial Optimization

Alexandre Laterre, Yunguan Fu, Mohamed Khalil Jabri, Alain-Sam Cohen, David Kas, Karl Hajjar, Torbjorn S. Dahl, Amine Kerkeni, Karim Beguir

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Journal ref Presented at the Thirty-second Conference on Neural Information Processing Systems (NeurIPS 2018), Deep Reinforcement Learning Workshop, Montreal, Canada, December 3-8, 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.00045 2018-12-04 cs.LG cs.AI cs.NE 62%

Using Monte Carlo Tree Search as a Demonstrator within Asynchronous Deep RL

Bilal Kartal, Pablo Hernandez-Leal, Matthew E. Taylor

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 6 figures, To appear at AAAI-19 Workshop on Reinforcement Learning in Games

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.03875 2018-10-30 cs.LG cs.AI cs.LO 62%

Learning Task Specifications from Demonstrations

Marcell Vazquez-Chanlatte, Susmit Jha, Ashish Tiwari, Mark K. Ho, Sanjit A. Seshia

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments NIPS 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.00781 2018-10-11 cs.LG cs.AI 62%

Unsupervised Learning of Goal Spaces for Intrinsically Motivated Goal Exploration

Alexandre Péré, Sébastien Forestier, Olivier Sigaud, Pierre-Yves Oudeyer

专题命中 工具调用 :tool use(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.04697 2018-07-18 cs.AI cs.LG stat.ML 62%

Learning to Search with MCTSnets

Arthur Guez, Théophane Weber, Ioannis Antonoglou, Karen Simonyan, Oriol Vinyals, Daan Wierstra, Rémi Munos, David Silver

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments ICML 2018 (camera-ready version)

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.02448 2018-06-08 cs.LG cs.AI cs.NE stat.ML 62%

Deep Reinforcement Learning for General Video Game AI

Ruben Rodriguez Torrado, Philip Bontrager, Julian Togelius, Jialin Liu, Diego Perez-Liebana

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments 8 pages, 4 figures, Accepted at the conference on Computational Intelligence and Games 2018 IEEE

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.00503 2018-05-24 stat.ML cs.AI cs.LG 62%

Mean Actor Critic

Cameron Allen, Kavosh Asadi, Melrose Roderick, Abdel-rahman Mohamed, George Konidaris, Michael Littman

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.10937 2018-03-30 cs.LG cs.AI stat.ML 62%

Best arm identification in multi-armed bandits with delayed feedback

Aditya Grover, Todor Markov, Peter Attia, Norman Jin, Nicholas Perkins, Bryan Cheong, Michael Chen, Zi Yang, Stephen Harris, William Chueh, Stefano Ermon

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments AISTATS 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.08456 2018-03-23 cs.AI cs.LG stat.ML 62%

Deep Reinforcement Learning with Model Learning and Monte Carlo Tree Search in Minecraft

Stephan Alaniz

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments The 3rd Multidisciplinary Conference on Reinforcement Learning and Decision Making (RLDM) 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.04873 2017-11-22 cs.LG cs.AI 62%

Efficient Architecture Search by Network Transformation

Han Cai, Tianyao Chen, Weinan Zhang, Yong Yu, Jun Wang

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments The Thirty-Second AAAI Conference on Artificial Intelligence (AAAI-18). We change the title from "Reinforcement Learning for Architecture Search by Network Transformation" to "Efficient Architecture Search by Network Transformation"

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.01391 2017-11-08 cs.AI cs.LG cs.RO 62%

Guiding the search in continuous state-action spaces by learning an action sampling distribution from off-target samples

Beomjoon Kim, Leslie Pack Kaelbling, Tomas Lozano-Perez

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.05958 2017-10-18 cs.LG cs.AI cs.CV 62%

Gradient-free Policy Architecture Search and Adaptation

Sayna Ebrahimi, Anna Rohrbach, Trevor Darrell

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted in Conference on Robot Learning, 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.03034 2017-07-12 cs.RO cs.AI cs.LG 62%

Learning Heuristic Search via Imitation

Mohak Bhardwaj, Sanjiban Choudhury, Sebastian Scherer

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.08520 2017-05-25 cs.AI cs.LG cs.NE 62%

An effective algorithm for hyperparameter optimization of neural networks

Gonzalo Diaz, Achille Fokoue, Giacomo Nannicini, Horst Samulowitz

专题命中 工具调用 :tool use(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.00777 2017-04-21 cs.CL cs.LG 62%

Towards End-to-End Reinforcement Learning of Dialogue Agents for Information Access

Bhuwan Dhingra, Lihong Li, Xiujun Li, Jianfeng Gao, Yun-Nung Chen, Faisal Ahmed, Li Deng

专题命中 工具调用 :agent(abstract);分类 cs.CL、cs.LG

Comments Accepted at ACL 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.00388 2017-01-12 cs.CL cs.LG 62%

Learning to Translate in Real-time with Neural Machine Translation

Jiatao Gu, Graham Neubig, Kyunghyun Cho, Victor O. K. Li

专题命中 工具调用 :agent(abstract);分类 cs.CL、cs.LG

Comments 10 pages, camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.01644 2015-11-30 cs.AI cs.LG 62%

Reinforcement Learning with Parameterized Actions

Warwick Masson, Pravesh Ranchod, George Konidaris

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted for AAAI 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1205.3109 2015-03-19 cs.LG cs.AI stat.ML 62%

Efficient Bayes-Adaptive Reinforcement Learning using Sample-Based Search

Arthur Guez, David Silver, Peter Dayan

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments 14 pages, 7 figures, includes supplementary material. Advances in Neural Information Processing Systems (NIPS) 2012

Journal ref (2012) Advances in Neural Information Processing Systems 25, pages 1034-1042

详情

展开后加载摘要…

URL PDF HTML 收藏
1306.4753 2013-06-21 cs.LG cs.AI math.OC 62%

Galerkin Methods for Complementarity Problems and Variational Inequalities

Geoffrey J. Gordon

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1209.3818 2013-04-03 cs.AI cs.LG 62%

Evolution and the structure of learning agents

Alok Raj

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments total 4 pages. Submitted to IEEE Congress on Evolutionary Computation 2013

详情

展开后加载摘要…

URL PDF HTML 收藏
1301.5488 2013-01-24 cs.LG cs.AI stat.ML 62%

Multi-class Generalized Binary Search for Active Inverse Reinforcement Learning

Francisco Melo, Manuel Lopes

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1202.2112 2012-02-10 cs.AI cs.LG cs.RO 62%

Predicting Contextual Sequences via Submodular Function Maximization

Debadeepta Dey, Tian Yu Liu, Martial Hebert, J. Andrew Bagnell

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
0909.0801 2010-12-30 cs.AI cs.IT cs.LG math.IT 62%

A Monte Carlo AIXI Approximation

Joel Veness, Kee Siong Ng, Marcus Hutter, William Uther, David Silver

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG

Comments 51 LaTeX pages, 11 figures, 6 tables, 4 algorithms

详情

展开后加载摘要…

URL PDF HTML 收藏
0912.5029 2010-01-14 cs.LG cs.AI 62%

Complexity of stochastic branch and bound methods for belief tree search in Bayesian reinforcement learning

Christos Dimitrakakis

专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG

Comments 13 pages, 1 figure, ICAART 2010

详情

展开后加载摘要…

URL PDF HTML 收藏