Takip et
Rémi Munos
Rémi Munos
DeepMind
inria.fr üzerinde doğrulanmış e-posta adresine sahip - Ana Sayfa
Başlık
Alıntı yapanlar
Alıntı yapanlar
Yıl
Bootstrap your own latent-a new approach to self-supervised learning
JB Grill, F Strub, F Altché, C Tallec, P Richemond, E Buchatskaya, ...
Advances in neural information processing systems 33, 21271-21284, 2020
57072020
A distributional perspective on reinforcement learning
MG Bellemare, W Dabney, R Munos
International conference on machine learning, 449-458, 2017
16362017
Unifying count-based exploration and intrinsic motivation
M Bellemare, S Srinivasan, G Ostrovski, T Schaul, D Saxton, R Munos
Advances in neural information processing systems 29, 2016
15972016
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
L Espeholt, H Soyer, R Munos, K Simonyan, V Mnih, T Ward, Y Doron, ...
International conference on machine learning, 1407-1416, 2018
15072018
Learning to reinforcement learn
JX Wang, Z Kurth-Nelson, D Tirumala, H Soyer, JZ Leibo, R Munos, ...
arXiv preprint arXiv:1611.05763, 2016
9882016
Sample efficient actor-critic with experience replay
Z Wang, V Bapst, N Heess, V Mnih, R Munos, K Kavukcuoglu, ...
arXiv preprint arXiv:1611.01224, 2016
9452016
Best arm identification in multi-armed bandits
JY Audibert, S Bubeck
COLT-23th Conference on learning theory-2010, 13 p., 2010
8972010
Minimax regret bounds for reinforcement learning
MG Azar, I Osband, R Munos
International conference on machine learning, 263-272, 2017
7742017
Exploration–exploitation tradeoff using variance estimates in multi-armed bandits
JY Audibert, R Munos, C Szepesvári
Theoretical Computer Science 410 (19), 1876-1902, 2009
7582009
Thompson sampling: An asymptotically optimal finite-time analysis
E Kaufmann, N Korda, R Munos
International conference on algorithmic learning theory, 199-213, 2012
7532012
Distributional reinforcement learning with quantile regression
W Dabney, M Rowland, M Bellemare, R Munos
Proceedings of the AAAI conference on artificial intelligence 32 (1), 2018
7352018
Count-based exploration with neural density models
G Ostrovski, MG Bellemare, A Oord, R Munos
International conference on machine learning, 2721-2730, 2017
6832017
Safe and efficient off-policy reinforcement learning
R Munos, T Stepleton, A Harutyunyan, M Bellemare
Advances in neural information processing systems 29, 2016
6782016
Finite-Time Bounds for Fitted Value Iteration.
R Munos, C Szepesvári
Journal of Machine Learning Research 9 (5), 2008
6052008
Pure exploration in multi-armed bandits problems
S Bubeck, R Munos, G Stoltz
Algorithmic Learning Theory: 20th International Conference, ALT 2009, Porto …, 2009
5822009
Automated curriculum learning for neural networks
A Graves, MG Bellemare, J Menick, R Munos, K Kavukcuoglu
international conference on machine learning, 1311-1320, 2017
5802017
Successor features for transfer in reinforcement learning
A Barreto, W Dabney, R Munos, JJ Hunt, T Schaul, HP van Hasselt, ...
Advances in neural information processing systems 30, 2017
5762017
Implicit quantile networks for distributional reinforcement learning
W Dabney, G Ostrovski, D Silver, R Munos
International conference on machine learning, 1096-1105, 2018
5382018
Modification of UCT with patterns in Monte-Carlo Go
S Gelly, Y Wang, R Munos, O Teytaud
INRIA, 2006
5362006
Recurrent experience replay in distributed reinforcement learning
S Kapturowski, G Ostrovski, J Quan, R Munos, W Dabney
International conference on learning representations, 2018
5042018
Sistem, işlemi şu anda gerçekleştiremiyor. Daha sonra yeniden deneyin.
Makaleler 1–20