arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

作者

David Silver

Reinforcement Learning

至 收录 68
1501.03959 2015-01-19 cs.AI cs.LG stat.ML

Value Iteration with Options and State Aggregation

Kamil Ciosek, David Silver

URL PDF HTML 收藏
1312.6055 2014-02-26 cs.LG

Unit Tests for Stochastic Optimization

Tom Schaul, Ioannis Antonoglou, David Silver

Comments Final submission to ICLR 2014 (revised according to reviews, additional results added)

URL PDF HTML 收藏
1402.1958 2014-02-11 cs.AI cs.LG stat.ML

Better Optimism By Bayes: Adaptive Planning with Rich Models

Arthur Guez, David Silver, Peter Dayan

Comments 11 pages, 11 figures

URL PDF HTML 收藏
1401.5390 2014-01-22 cs.CL cs.AI cs.LG

Learning to Win by Reading Manuals in a Monte-Carlo Framework

S. R. K. Branavan, David Silver, Regina Barzilay

Journal ref Journal Of Artificial Intelligence Research, Volume 43, pages 661-704, 2012

详情
URL PDF HTML 收藏
1312.5602 2013-12-20 cs.LG

Playing Atari with Deep Reinforcement Learning

Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, Martin Riedmiller

Comments NIPS Deep Learning Workshop 2013

URL PDF HTML 收藏
1206.6473 2012-07-02 cs.AI cs.LG

Compositional Planning Using Optimal Option Models

David Silver, Kamil Ciosek

Comments Appears in Proceedings of the 29th International Conference on Machine Learning (ICML 2012)

URL PDF HTML 收藏
0909.0801 2010-12-30 cs.AI cs.IT cs.LG math.IT

A Monte Carlo AIXI Approximation

Joel Veness, Kee Siong Ng, Marcus Hutter, William Uther, David Silver

Comments 51 LaTeX pages, 11 figures, 6 tables, 4 algorithms

URL PDF HTML 收藏
1007.2049 2010-10-04 cs.LG

Reinforcement Learning via AIXI Approximation

Joel Veness, Kee Siong Ng, Marcus Hutter, David Silver

Comments 8 LaTeX pages, 1 figure

Journal ref Proc. 24th AAAI Conference on Artificial Intelligence (AAAI 2010) pages 605-611

URL PDF HTML 收藏