arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 15819 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. Agent评测 15819 篇

1712.03890 2017-12-13 cs.NI cs.AI cs.LG 62%

DeepConfig: Automating Data Center Network Topologies Management with Machine Learning

Christopher Streiffer, Huan Chen, Theophilus Benson, Asim Kadav

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.09883 2017-11-29 cs.LG cs.AI 62%

AI Safety Gridworlds

Jan Leike, Miljan Martic, Victoria Krakovna, Pedro A. Ortega, Tom Everitt, Andrew Lefrancq, Laurent Orseau, Shane Legg

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.03732 2017-11-27 cs.AI cs.LG 62%

Deep Q-learning from Demonstrations

Todd Hester, Matej Vecerik, Olivier Pietquin, Marc Lanctot, Tom Schaul, Bilal Piot, Dan Horgan, John Quan, Andrew Sendonaris, Gabriel Dulac-Arnold, Ian Osband, John Agapiou, Joel Z. Leibo, Audrunas Gruslys

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Published at AAAI 2018. Previously on arxiv as "Learning from Demonstrations for Real World Reinforcement Learning"

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.07613 2017-11-22 cs.CV cs.AI cs.CL 62%

Are You Talking to Me? Reasoned Visual Dialog Generation through Adversarial Learning

Qi Wu, Peng Wang, Chunhua Shen, Ian Reid, Anton van den Hengel

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.05427 2017-11-07 cs.AI cs.LG 62%

Repeated Inverse Reinforcement Learning

Kareem Amin, Nan Jiang, Satinder Singh

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments The first two authors contributed equally to this work. The paper appears in NIPS 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1504.02247 2017-11-02 cs.AI cs.LG stat.ML 62%

Projective simulation with generalization

Alexey A. Melnikov, Adi Makmal, Vedran Dunjko, Hans J. Briegel

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 14 pages, 9 figures

Journal ref Sci. Rep. 7, 14430 (2017)

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.05298 2017-10-25 cs.LG cs.CL cs.RO 62%

Text2Action: Generative Adversarial Synthesis from Language to Action

Hyemin Ahn, Timothy Ha, Yunho Choi, Hwiyeon Yoo, Songhwai Oh

专题命中 Agent评测 :agent(abstract);分类 cs.CL、cs.LG

Comments 8 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.09816 2017-09-29 cs.CL cs.AI 62%

Edina: Building an Open Domain Socialbot with Self-dialogues

Ben Krause, Marco Damonte, Mihai Dobre, Daniel Duma, Joachim Fainberg, Federico Fancellu, Emmanuel Kahembwe, Jianpeng Cheng, Bonnie Webber

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments 10 pages; submitted to the 1st Proceedings of the Alexa Prize

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.08292 2017-09-26 cs.RO cs.AI cs.LG 62%

Underwater Multi-Robot Convoying using Visual Tracking by Detection

Florian Shkurti, Wei-Di Chang, Peter Henderson, Md Jahidul Islam, Juan Camilo Gamboa Higuera, Jimmy Li, Travis Manderson, Anqi Xu, Gregory Dudek, Junaed Sattar

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.08417 2017-08-22 cs.AI cs.LG stat.ML 62%

Reinforcement Learning with a Corrupted Reward Channel

Tom Everitt, Victoria Krakovna, Laurent Orseau, Marcus Hutter, Shane Legg

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments A shorter version of this report was accepted to IJCAI 2017 AI and Autonomy track

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.05122 2017-08-18 cs.HC cs.AI cs.CL cs.CV 62%

Evaluating Visual Conversational Agents via Cooperative Human-AI Games

Prithvijit Chattopadhyay, Deshraj Yadav, Viraj Prabhu, Arjun Chandrasekaran, Abhishek Das, Stefan Lee, Dhruv Batra, Devi Parikh

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments HCOMP 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.06887 2017-07-24 cs.LG cs.AI stat.ML 62%

A Distributional Perspective on Reinforcement Learning

Marc G. Bellemare, Will Dabney, Rémi Munos

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments ICML 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.05173 2017-07-18 cs.AI cs.LG cs.NE 62%

Trial without Error: Towards Safe Reinforcement Learning via Human Intervention

William Saunders, Girish Sastry, Andreas Stuhlmueller, Owain Evans

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.06131 2017-07-12 cs.AI cs.LG stat.ML 62%

Learning to Acquire Information

Yewen Pu, Leslie P Kaelbling, Armando Solar-Lezama

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.07147 2017-06-23 cs.LG cs.AI q-bio.NC stat.ML 62%

A Useful Motif for Flexible Task Learning in an Embodied Two-Dimensional Visual Environment

Kevin T. Feigelis, Daniel L. K. Yamins

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1410.0949 2017-06-08 cs.LG cs.AI math.OC stat.ML 62%

Tight Regret Bounds for Stochastic Combinatorial Semi-Bandits

Branislav Kveton, Zheng Wen, Azin Ashkan, Csaba Szepesvari

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Proceedings of the 18th International Conference on Artificial Intelligence and Statistics

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.10229 2017-05-30 cs.CL cs.LG cs.NE stat.ML 62%

Latent Intention Dialogue Models

Tsung-Hsien Wen, Yishu Miao, Phil Blunsom, Steve Young

专题命中 Agent评测 :agent(abstract);分类 cs.CL、cs.LG

Comments Accepted at ICML 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.06230 2017-05-09 cs.LG cs.AI 62%

Beating the World's Best at Super Smash Bros. with Deep Reinforcement Learning

Vlad Firoiu, William F. Whitney, Joshua B. Tenenbaum

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.00673 2017-05-04 cs.AI cs.SE 62%

MACA: A Modular Architecture for Conversational Agents

Hoai Phuoc Truong, Prasanna Parthasarathi, Joelle Pineau

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.SE

Comments The architecture needs to be tested further. Sorry for the inconvenience. We should be putting up the paper up soon

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.01086 2017-03-28 cs.AI cs.LG cs.RO 62%

Deep Learning of Robotic Tasks without a Simulator using Strong and Weak Human Supervision

Bar Hilleli, Ran El-Yaniv

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.04936 2017-02-14 cs.CL cs.AI 62%

Learning through Dialogue Interactions by Asking Questions

Jiwei Li, Alexander H. Miller, Sumit Chopra, Marc'Aurelio Ranzato, Jason Weston

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1406.7443 2017-02-01 cs.LG cs.AI stat.ML 62%

Efficient Learning in Large-Scale Combinatorial Semi-Bandits

Zheng Wen, Branislav Kveton, Azin Ashkan

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.09894 2016-12-04 cs.AI cs.LG stat.ML 62%

Exploration for Multi-task Reinforcement Learning with Deep Generative Models

Sai Praveen Bangaru, JS Suhas, Balaraman Ravindran

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 5 figures; NIPS Deep Reinforcement Learning Workshop 2016, Barcelona

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.10328 2016-12-01 cs.LG cs.AI 62%

The observer-assisted method for adjusting hyper-parameters in deep learning algorithms

Maciej Wielgosz

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.08666 2016-11-29 cs.LG cs.AI cs.RO 62%

Training an Interactive Humanoid Robot Using Multimodal Deep Reinforcement Learning

Heriberto Cuayáhuitl, Guillaume Couly, Clément Olalainty

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments NIPS Workshop on Future of Interactive Learning Machines, 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.04994 2016-11-21 cs.LG cs.AI 62%

Exploration Potential

Jan Leike

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments 10 pages, including proofs

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.05379 2016-11-17 cs.AI cs.CL cs.HC cs.RO 62%

PCT and Beyond: Towards a Computational Framework for `Intelligent' Communicative Systems

Prof. Roger K. Moore

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

Comments To appear in A. McElhone & W. Mansell (Eds.), Living Control Systems IV: Perceptual Control Theory and the Future of the Life and Social Sciences, Benchmark Publications Inc

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.06620 2016-10-25 cs.CL cs.AI cs.CV 62%

Proposing Plausible Answers for Open-ended Visual Question Answering

Omid Bakhshandeh, Trung Bui, Zhe Lin, Walter Chang

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1605.07157 2016-10-19 cs.LG cs.AI cs.CV cs.RO 62%

Unsupervised Learning for Physical Interaction through Video Prediction

Chelsea Finn, Ian Goodfellow, Sergey Levine

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments To appear in NIPS '16; Video results, code, and data available at: http://www.sites.google.com/site/robotprediction

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.02164 2016-10-10 cs.LG cs.AI 62%

Deep Reinforcement Learning From Raw Pixels in Doom

Danijar Hafner

专题命中 Agent评测 :agent(abstract);分类 cs.AI、cs.LG

Comments Bachelor's thesis

详情

展开后加载摘要…

URL PDF HTML 收藏