arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

作者

Joelle Pineau

Reinforcement Learning

至 收录 126
2102.07097 2021-02-16 cs.LG cs.AI

Domain Adversarial Reinforcement Learning

Bonnie Li, Vincent François-Lavet, Thang Doan, Joelle Pineau

URL PDF HTML 收藏
2007.07206 2021-02-15 cs.LG cs.AI stat.ML

Learning Robust State Abstractions for Hidden-Parameter Block MDPs

Amy Zhang, Shagun Sodhani, Khimya Khetarpal, Joelle Pineau

Comments Accepted at the 9th International Conference on Learning Representations. 22 pages, 14 figures

URL PDF HTML 收藏
1906.10437 2021-02-09 cs.LG stat.ML

Learning Causal State Representations of Partially Observable Environments

Amy Zhang, Zachary C. Lipton, Luis Pineda, Kamyar Azizzadenesheli, Anima Anandkumar, Laurent Itti, Joelle Pineau, Tommaso Furlanello

Comments 35 pages, 8 figures

URL PDF HTML 收藏
2102.03419 2021-02-09 cs.AI cs.CL cs.IR cs.LG cs.SI

Exploring the Limits of Few-Shot Link Prediction in Knowledge Graphs

Dora Jambor, Komal Teru, Joelle Pineau, William L. Hamilton

Comments code available at https://github.com/dorajam/few-shot-link-prediction-paper

Journal ref European Chapter of the ACL (EACL), 2021

URL PDF HTML 收藏
2101.04909 2021-01-26 cs.CV cs.LG

COVID-19 Prognosis via Self-Supervised Representation Learning and Multi-Image Prediction

Anuroop Sriram, Matthew Muckley, Koustuv Sinha, Farah Shamout, Joelle Pineau, Krzysztof J. Geras, Lea Azour, Yindalon Aphinyanaphongs, Nafissa Yakubova, William Moore

URL PDF HTML 收藏
2003.12206 2021-01-01 cs.LG stat.ML

Improving Reproducibility in Machine Learning Research (A Report from the NeurIPS 2019 Reproducibility Program)

Joelle Pineau, Philippe Vincent-Lamarre, Koustuv Sinha, Vincent Larivière, Alina Beygelzimer, Florence d'Alché-Buc, Emily Fox, Hugo Larochelle

Comments To appear at JMLR, 16 pages + Appendix

URL PDF HTML 收藏
2012.02055 2020-12-04 cs.RO cs.LG

Intervention Design for Effective Sim2Real Transfer

Melissa Mozifian, Amy Zhang, Joelle Pineau, David Meger

URL PDF HTML 收藏
2010.15896 2020-12-04 cs.MA cs.AI

Exploring Zero-Shot Emergent Communication in Embodied Multi-Agent Populations

Kalesha Bullard, Franziska Meier, Douwe Kiela, Joelle Pineau, Jakob Foerster

URL PDF HTML 收藏
2010.03691 2020-12-04 cs.LG

Regularized Inverse Reinforcement Learning

Wonseok Jeon, Chen-Yang Su, Paul Barde, Thang Doan, Derek Nowrouzezahrai, Joelle Pineau

Comments 26 pages, 7 figures

URL PDF HTML 收藏
2003.00898 2020-10-28 stat.AP

The importance of transparency and reproducibility in artificial intelligence research

Benjamin Haibe-Kains, George Alexandru Adam, Ahmed Hosny, Farnoosh Khodakarami, MAQC Society Board, Levi Waldron, Bo Wang, Chris McIntosh, Anshul Kundaje, Casey S. Greene, Michael M. Hoffman, Jeffrey T. Leek, Wolfgang Huber, Alvis Brazma, Joelle Pineau, Robert Tibshirani, Trevor Hastie, John P. A. Ioannidis, John Quackenbush, Hugo J. W. L. Aerts

Journal ref Nature 586 (2020) E14-E16

URL PDF HTML 收藏
2002.02863 2020-10-16 cs.LG stat.ML

Representation of Reinforcement Learning Policies in Reproducing Kernel Hilbert Spaces

Bogdan Mazoure, Thang Doan, Tianyu Li, Vladimir Makarenkov, Joelle Pineau, Doina Precup, Guillaume Rabusseau

URL PDF HTML 收藏
2008.11811 2020-08-28 cs.LG math.OC stat.ML

Constrained Markov Decision Processes via Backward Value Functions

Harsh Satija, Philip Amortila, Joelle Pineau

URL PDF HTML 收藏
2008.10427 2020-08-25 cs.CL cs.AI

How To Evaluate Your Dialogue System: Probe Tasks as an Alternative for Token-level Evaluation Metrics

Prasanna Parthasarathi, Joelle Pineau, Sarath Chandar

URL PDF HTML 收藏
1911.08019 2020-08-24 cs.LG cs.CV stat.ML

Online Learned Continual Compression with Adaptive Quantization Modules

Lucas Caccia, Eugene Belilovsky, Massimo Caccia, Joelle Pineau

URL PDF HTML 收藏
1910.01741 2020-07-10 cs.LG cs.AI cs.RO stat.ML

Improving Sample Efficiency in Model-Free Reinforcement Learning from Images

Denis Yarats, Amy Zhang, Ilya Kostrikov, Brandon Amos, Joelle Pineau, Rob Fergus

URL PDF HTML 收藏
1909.07543 2020-07-10 cs.LG cs.AI cs.MA stat.ML

Attraction-Repulsion Actor-Critic for Continuous Control Reinforcement Learning

Thang Doan, Bogdan Mazoure, Moloud Abdar, Audrey Durand, Joelle Pineau, R Devon Hjelm

URL PDF HTML 收藏
2007.02786 2020-07-07 cs.LG stat.ML

TDprop: Does Jacobi Preconditioning Help Temporal Difference Learning?

Joshua Romoff, Peter Henderson, David Kanaa, Emmanuel Bengio, Ahmed Touati, Pierre-Luc Bacon, Joelle Pineau

Comments Presented at the Theoretical Foundations of Reinforcement Learning workshop at ICML 2020

URL PDF HTML 收藏
2007.01516 2020-07-06 cs.LG q-bio.GN stat.AP stat.ML

Deep interpretability for GWAS

Deepak Sharma, Audrey Durand, Marc-André Legault, Louis-Philippe Lemieux Perreault, Audrey Lemaçon, Marie-Pierre Dubé, Joelle Pineau

Comments Accepted at ICML 2020 workshop on ML Interpretability for Scientific Discovery

URL PDF HTML 收藏
2002.01093 2020-06-24 cs.CL cs.AI cs.LG cs.MA stat.ML

On the interaction between supervision and self-play in emergent communication

Ryan Lowe, Abhinav Gupta, Jakob Foerster, Douwe Kiela, Joelle Pineau

Comments The first two authors contributed equally. Accepted at ICLR 2020

URL PDF HTML 收藏
2003.04108 2020-06-22 cs.LG stat.ML

Stable Policy Optimization via Off-Policy Divergence Regularization

Ahmed Touati, Amy Zhang, Joelle Pineau, Pascal Vincent

Journal ref Proceedings of the 36th Conference on Uncertainty in Artificial Intelligence (UAI), PMLR volume 124, 2020

URL PDF HTML 收藏
2003.06016 2020-06-15 cs.LG cs.AI stat.ML

Invariant Causal Prediction for Block MDPs

Amy Zhang, Clare Lyle, Shagun Sodhani, Angelos Filos, Marta Kwiatkowska, Joelle Pineau, Yarin Gal, Doina Precup

Comments Accepted to ICML 2020. 16 pages, 8 figures

URL PDF HTML 收藏
2005.06616 2020-05-15 cs.CY cs.AI cs.CL cs.HC cs.LG

A Large-Scale, Open-Domain, Mixed-Interface Dialogue-Based ITS for STEM

Iulian Vlad Serban, Varun Gupta, Ekaterina Kochmar, Dung D. Vu, Robert Belfer, Joelle Pineau, Aaron Courville, Laurent Charlin, Yoshua Bengio

Comments 6 pages, 1 figure, 1 table, accepted for publication in the 21st International Conference on Artificial Intelligence in Education (AIED 2020)

URL PDF HTML 收藏
2005.02431 2020-05-11 cs.CL cs.AI

Automated Personalized Feedback Improves Learning Gains in an Intelligent Tutoring System

Ekaterina Kochmar, Dung Do Vu, Robert Belfer, Varun Gupta, Iulian Vlad Serban, Joelle Pineau

Comments To be published in Proceedings of the the 21st International Conference on Artificial Intelligence in Education (AIED 2020)

URL PDF HTML 收藏
2005.03648 2020-05-08 cs.LG cs.AI stat.ML

Plan2Vec: Unsupervised Representation Learning by Latent Plans

Ge Yang, Amy Zhang, Ari S. Morcos, Joelle Pineau, Pieter Abbeel, Roberto Calandra

Comments code available at https://geyang.github.io/plan2vec

Journal ref Proceedings of Machine Learning Research, the 2nd Annual Conference on Learning for Dynamics and Control (2020) Volume 120, 1-12

URL PDF HTML 收藏
2005.00583 2020-05-05 cs.CL cs.LG

Learning an Unreferenced Metric for Online Dialogue Evaluation

Koustuv Sinha, Prasanna Parthasarathi, Jasmine Wang, Ryan Lowe, William L. Hamilton, Joelle Pineau

Comments Accepted at ACL 2020, 5 pages

URL PDF HTML 收藏
1906.04585 2020-04-23 cs.LG cs.AI cs.MA math.OC stat.ML

Gossip-based Actor-Learner Architectures for Deep Reinforcement Learning

Mahmoud Assran, Joshua Romoff, Nicolas Ballas, Joelle Pineau, Michael Rabbat

Journal ref Advances in Neural Information Processing Systems (2019) 13299-13309

URL PDF HTML 收藏
2003.06560 2020-03-17 cs.LG stat.ML

Evaluating Logical Generalization in Graph Neural Networks

Koustuv Sinha, Shagun Sodhani, Joelle Pineau, William L. Hamilton

URL PDF HTML 收藏
2003.06350 2020-03-16 cs.LG stat.ML

Interference and Generalization in Temporal Difference Learning

Emmanuel Bengio, Joelle Pineau, Doina Precup

Comments Submitted to ICML 2020. 20 pages, 14 figures

URL PDF HTML 收藏
2002.10525 2020-02-26 cs.MA cs.LG

Scalable Multi-Agent Inverse Reinforcement Learning via Actor-Attention-Critic

Wonseok Jeon, Paul Barde, Derek Nowrouzezahrai, Joelle Pineau

URL PDF HTML 收藏
1810.11187 2020-02-25 cs.LG cs.AI cs.MA stat.ML

TarMAC: Targeted Multi-Agent Communication

Abhishek Das, Théophile Gervet, Joshua Romoff, Dhruv Batra, Devi Parikh, Michael Rabbat, Joelle Pineau

Comments ICML 2019

URL PDF HTML 收藏