Mastering the game of Go with deep neural networks and tree search D Silver, A Huang, CJ Maddison, A Guez, L Sifre, G Van Den Driessche, ... nature 529 (7587), 484-489, 2016 | 17310 | 2016 |
Mastering the game of go without human knowledge D Silver, J Schrittwieser, K Simonyan, I Antonoglou, A Huang, A Guez, ... nature 550 (7676), 354-359, 2017 | 9884 | 2017 |
A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play D Silver, T Hubert, J Schrittwieser, I Antonoglou, M Lai, A Guez, M Lanctot, ... Science 362 (6419), 1140-1144, 2018 | 3802 | 2018 |
Mastering chess and shogi by self-play with a general reinforcement learning algorithm D Silver, T Hubert, J Schrittwieser, I Antonoglou, M Lai, A Guez, M Lanctot, ... arXiv preprint arXiv:1712.01815, 2017 | 1984 | 2017 |
Mastering atari, go, chess and shogi by planning with a learned model J Schrittwieser, I Antonoglou, T Hubert, K Simonyan, L Sifre, S Schmitt, ... Nature 588 (7839), 604-609, 2020 | 1883 | 2020 |
Starcraft ii: A new challenge for reinforcement learning O Vinyals, T Ewalds, S Bartunov, P Georgiev, AS Vezhnevets, M Yeo, ... arXiv preprint arXiv:1708.04782, 2017 | 951 | 2017 |
Deepmind lab C Beattie, JZ Leibo, D Teplyashin, T Ward, M Wainwright, H Küttler, ... arXiv preprint arXiv:1612.03801, 2016 | 548 | 2016 |
Competition-level code generation with alphacode Y Li, D Choi, J Chung, N Kushman, J Schrittwieser, R Leblond, T Eccles, ... Science 378 (6624), 1092-1097, 2022 | 425 | 2022 |
Discovering faster matrix multiplication algorithms with reinforcement learning A Fawzi, M Balog, A Huang, T Hubert, B Romera-Paredes, M Barekatain, ... Nature 610 (7930), 47-53, 2022 | 307 | 2022 |
OpenSpiel: A framework for reinforcement learning in games M Lanctot, E Lockhart, JB Lespiau, V Zambaldi, S Upadhyay, J Pérolat, ... arXiv preprint arXiv:1908.09453, 2019 | 206 | 2019 |
Cyprien de Masson d’Autume, Igor Babuschkin, Xinyun Chen, Po-Sen Huang, Johannes Welbl, Sven Gowal, Alexey Cherepanov, James Molloy, Daniel J Y Li, D Choi, J Chung, N Kushman, J Schrittwieser, R Leblond, T Eccles, ... Mankowitz, Esme Sutherland Robson, Pushmeet Kohli, Nando de Freitas, Koray …, 2022 | 139 | 2022 |
Bayesian optimization in alphago Y Chen, A Huang, Z Wang, I Antonoglou, J Schrittwieser, D Silver, ... arXiv preprint arXiv:1812.06855, 2018 | 134 | 2018 |
and Hassabis, D. 2016. Mastering the game of go with deep neural networks and tree search D Silver, A Huang, CJ Maddison, A Guez, L Sifre, G Van Den Driessche, ... Nature 529 (7587), 484-489, 0 | 90 | |
Online and offline reinforcement learning by planning with a learned model J Schrittwieser, T Hubert, A Mandhane, M Barekatain, I Antonoglou, ... Advances in Neural Information Processing Systems 34, 27580-27591, 2021 | 87 | 2021 |
Adrian Bolton και others D Silver, J Schrittwieser, K Simonyan, I Antonoglou, A Huang, A Guez, ... Mastering the game of go without human knowledge. nature 550 (7676), 354-359, 2017 | 81 | 2017 |
& Hassabis, D.(2017) D Silver, J Schrittwieser, K Simonyan, I Antonoglou, A Huang, A Guez Mastering the game of go without human knowledge. nature 550 (7676), 354-359, 0 | 57 | |
et almbox. 2016. Mastering the game of Go with deep neural networks and tree search D Silver, A Huang, CJ Maddison, A Guez, L Sifre, G Van Den Driessche, ... Nature 529 (7587), 484-489, 2016 | 55 | 2016 |
Learning and planning in complex action spaces T Hubert, J Schrittwieser, I Antonoglou, M Barekatain, S Schmitt, D Silver International Conference on Machine Learning, 4476-4486, 2021 | 51 | 2021 |
Panneershelvam Veda S David, H Aja, J Maddison Chris, G Arthur, S Laurent, ... Lanctot Marc, Dieleman Sander, Grewe Dominik, Nham John, Kalchbrenner Nal …, 2016 | 49 | 2016 |
Planning in stochastic environments with a learned model I Antonoglou, J Schrittwieser, S Ozair, TK Hubert, D Silver International Conference on Learning Representations, 2021 | 36 | 2021 |