arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

作者

Joelle Pineau

Reinforcement Learning

至 收录 126
1811.02549 2020-02-21 cs.CL cs.LG

Language GANs Falling Short

Massimo Caccia, Lucas Caccia, William Fedus, Hugo Larochelle, Joelle Pineau, Laurent Charlin

Journal ref ICLR 2020 - Proceedings of the Seventh International Conference on Learning Representation

URL PDF HTML 收藏
1812.01180 2019-12-04 cs.CV

Deep Generative Modeling of LiDAR Data

Lucas Caccia, Herke van Hoof, Aaron Courville, Joelle Pineau

Comments Presented at IROS 2019

URL PDF HTML 收藏
1911.06970 2019-12-03 cs.LG cs.AI stat.ML

Off-Policy Policy Gradient Algorithms by Constraining the State Distribution Shift

Riashat Islam, Komal K. Teru, Deepak Sharma, Joelle Pineau

Comments Accepted at NeurIPS 2019 workshop on Deep Reinforcement Learning

URL PDF HTML 收藏
1911.09033 2019-11-21 cs.LG cs.CV stat.ML

Exploiting Spatial Invariance for Scalable Unsupervised Object Tracking

Eric Crawford, Joelle Pineau

Comments Accepted at AAAI 2020. Code: https://github.com/e2crawfo/silot. Visualizations: https://sites.google.com/view/silot

URL PDF HTML 收藏
1909.02128 2019-11-21 cs.AI cs.LG cs.MA

No Press Diplomacy: Modeling Multi-Agent Gameplay

Philip Paquette, Yuchen Lu, Steven Bocco, Max O. Smith, Satya Ortiz-Gagne, Jonathan K. Kummerfeld, Satinder Singh, Joelle Pineau, Aaron Courville

Comments Accepted at NeurIPS 2019

URL PDF HTML 收藏
1910.01708 2019-10-07 cs.LG cs.AI stat.ML

Benchmarking Batch Deep Reinforcement Learning Algorithms

Scott Fujimoto, Edoardo Conti, Mohammad Ghavamzadeh, Joelle Pineau

Comments Deep RL Workshop NeurIPS 2019

URL PDF HTML 收藏
1905.06893 2019-09-25 cs.LG stat.ML

Leveraging exploration in off-policy algorithms via normalizing flows

Bogdan Mazoure, Thang Doan, Audrey Durand, R Devon Hjelm, Joelle Pineau

Comments Accepted to 3rd Conference on Robot Learning (CoRL 2019); Keywords: Exploration, soft actor-critic, normalizing flow, off-policy; maximum entropy, reinforcement learning; deceptive reward; sparse reward; inverse autoregressive flow

URL PDF HTML 收藏
1908.06177 2019-09-05 cs.LG cs.CL cs.LO stat.ML

CLUTRR: A Diagnostic Benchmark for Inductive Reasoning from Text

Koustuv Sinha, Shagun Sodhani, Jin Dong, Joelle Pineau, William L. Hamilton

Comments Accepted at EMNLP 2019, 9 page content + Appendix

URL PDF HTML 收藏
1806.02315 2019-07-02 cs.LG stat.ML

Randomized Value Functions via Multiplicative Normalizing Flows

Ahmed Touati, Harsh Satija, Joshua Romoff, Joelle Pineau, Pascal Vincent

Journal ref UAI 2019: Conference on Uncertainty in Artificial Intelligence 2019

URL PDF HTML 收藏
1902.01883 2019-05-28 cs.LG cs.AI stat.ML

Separating value functions across time-scales

Joshua Romoff, Peter Henderson, Ahmed Touati, Emma Brunskill, Joelle Pineau, Yann Ollivier

Comments Full version accepted to ICML 2019. Extended abstract also to be presented at RLDM 2019

URL PDF HTML 收藏
1905.09562 2019-05-24 cs.LG stat.ML

Recurrent Value Functions

Pierre Thodoroff, Nishanth Anand, Lucas Caccia, Doina Precup, Joelle Pineau

URL PDF HTML 收藏
1811.00429 2019-04-12 cs.LG stat.ML

Temporal Regularization in Markov Decision Process

Pierre Thodoroff, Audrey Durand, Joelle Pineau, Doina Precup

Comments Published as a conference paper at NIPS 2018

URL PDF HTML 收藏
1903.05168 2019-03-14 cs.LG cs.AI cs.CL stat.ML

On the Pitfalls of Measuring Emergent Communication

Ryan Lowe, Jakob Foerster, Y-Lan Boureau, Joelle Pineau, Yann Dauphin

Comments AAMAS 2019. 13 pages

URL PDF HTML 收藏
1808.00020 2019-03-13 cs.LG stat.ML

On-line Adaptative Curriculum Learning for GANs

Thang Doan, Joao Monteiro, Isabela Albuquerque, Bogdan Mazoure, Audrey Durand, Joelle Pineau, R Devon Hjelm

Comments Accepted to the Thirty-Third AAAI Conference On Artificial Intelligence, 2019 (Added 128x128 CelebA samples to the end of the appendix)

Journal ref Proceedings of 33rd AAAI Conference on Artificial Intelligence (AAAI 2019)

URL PDF HTML 收藏
1709.07796 2019-02-07 stat.ML cs.AI cs.LG

On overfitting and asymptotic bias in batch reinforcement learning with partial observability

Vincent Francois-Lavet, Guillaume Rabusseau, Joelle Pineau, Damien Ernst, Raphael Fonteneau

Comments Accepted at the Journal of Artificial Intelligence Research (JAIR) - 31 pages

URL PDF HTML 收藏
1902.00098 2019-02-04 cs.AI cs.CL cs.HC

The Second Conversational Intelligence Challenge (ConvAI2)

Emily Dinan, Varvara Logacheva, Valentin Malykh, Alexander Miller, Kurt Shuster, Jack Urbanek, Douwe Kiela, Arthur Szlam, Iulian Serban, Ryan Lowe, Shrimai Prabhumoye, Alan W Black, Alexander Rudnicky, Jason Williams, Joelle Pineau, Mikhail Burtsev, Jason Weston

URL PDF HTML 收藏
1709.06560 2019-01-31 cs.LG stat.ML

Deep Reinforcement Learning that Matters

Peter Henderson, Riashat Islam, Philip Bachman, Joelle Pineau, Doina Precup, David Meger

Comments Accepted to the Thirthy-Second AAAI Conference On Artificial Intelligence (AAAI), 2018

URL PDF HTML 收藏
1811.12560 2018-12-04 cs.LG cs.AI stat.ML

An Introduction to Deep Reinforcement Learning

Vincent Francois-Lavet, Peter Henderson, Riashat Islam, Marc G. Bellemare, Joelle Pineau

Journal ref Foundations and Trends in Machine Learning: Vol. 11, No. 3-4, 2018

详情
URL PDF HTML 收藏
1809.04506 2018-11-20 cs.LG cs.AI stat.ML

Combined Reinforcement Learning via Abstract Representations

Vincent François-Lavet, Yoshua Bengio, Doina Precup, Joelle Pineau

Comments Accepted to the Thirty-Third AAAI Conference On Artificial Intelligence, 2019

URL PDF HTML 收藏
1811.06032 2018-11-16 cs.LG cs.AI stat.ML

Natural Environment Benchmarks for Reinforcement Learning

Amy Zhang, Yuxin Wu, Joelle Pineau

Comments 12 figures

URL PDF HTML 收藏
1811.02959 2018-11-09 cs.CL cs.AI

Compositional Language Understanding with Text-based Relational Reasoning

Koustuv Sinha, Shagun Sodhani, William L. Hamilton, Joelle Pineau

Comments 4 pages of main content, to be presented at Relational Representation Learning Workshop, NIPS 2018, Montreal

URL PDF HTML 收藏
1811.02714 2018-11-09 cs.CL

The RLLChatbot: a solution to the ConvAI challenge

Nicolas Gontier, Koustuv Sinha, Peter Henderson, Iulian Serban, Michael Noseworthy, Prasanna Parthasarathi, Joelle Pineau

Comments 46 pages including references and appendix, 14 figures, 12 tables; Under review for the Dialogue & Discourse journal

URL PDF HTML 收藏
1805.03359 2018-11-09 cs.LG cs.AI stat.ML

Reward Estimation for Variance Reduction in Deep Reinforcement Learning

Joshua Romoff, Peter Henderson, Alexandre Piché, Vincent Francois-Lavet, Joelle Pineau

Comments Version 1 as appears in the International Conference on Learning Representations (ICLR) 2018 Workshop Track; Version 2 as appears in the Proceedings of The 2nd Conference on Robot Learning

URL PDF HTML 收藏
1811.01302 2018-11-06 cs.LG cs.CL stat.ML

Adversarial Gain

Peter Henderson, Koustuv Sinha, Rosemary Nan Ke, Joelle Pineau

URL PDF HTML 收藏
1810.02525 2018-10-08 cs.LG cs.AI stat.ML

Where Did My Optimum Go?: An Empirical Analysis of Gradient Descent Optimization in Policy Gradient Methods

Peter Henderson, Joshua Romoff, Joelle Pineau

Comments Accepted at the European Workshop on Reinforcement Learning 2018 (EWRL14)

URL PDF HTML 收藏
1809.05524 2018-09-17 cs.CL cs.AI

Extending Neural Generative Conversational Model using External Knowledge Sources

Prasanna Parthasarathi, Joelle Pineau

Comments Accepted in EMNLP 2018

URL PDF HTML 收藏
1809.04988 2018-09-14 cs.LG cs.AI stat.ML

Sequential Coordination of Deep Models for Learning Visual Arithmetic

Eric Crawford, Guillaume Rabusseau, Joelle Pineau

URL PDF HTML 收藏
1807.04723 2018-07-13 cs.LG cs.AI cs.CL cs.NE stat.ML

The Bottleneck Simulator: A Model-based Deep Reinforcement Learning Approach

Iulian Vlad Serban, Chinnadhurai Sankar, Michael Pieper, Joelle Pineau, Yoshua Bengio

Comments 26 pages, 2 figures, 4 tables

URL PDF HTML 收藏
1806.07937 2018-06-26 cs.LG cs.AI stat.ML

A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning

Amy Zhang, Nicolas Ballas, Joelle Pineau

Comments 20 pages, 16 figures

URL PDF HTML 收藏
1806.04342 2018-06-13 stat.ML cs.LG

Focused Hierarchical RNNs for Conditional Sequence Processing

Nan Rosemary Ke, Konrad Zolna, Alessandro Sordoni, Zhouhan Lin, Adam Trischler, Yoshua Bengio, Joelle Pineau, Laurent Charlin, Chris Pal

Comments To appear at ICML 2018

URL PDF HTML 收藏