| Genie: Generative interactive environments J Bruce, MD Dennis, A Edwards, J Parker-Holder, Y Shi, E Hughes, M Lai, ... Forty-first International Conference on Machine Learning, 2024 | 1134 | 2024 |
| Adversarial policies: Attacking deep reinforcement learning A Gleave, M Dennis, C Wild, N Kant, S Levine, S Russell (ICLR 2020) - Eighth International Conference on Learning Representations, 2020 | 645 | 2020 |
| Emergent Complexity and Zero-shot Transfer via Unsupervised Environment Design M Dennis, N Jaques, E Vinitsky, A Bayen, S Russell, A Critch, S Levine (NeurIPS 2020) - Advances in Neural Information Processing Systems 33, 2020 | 453 | 2020 |
| Multi-agent risks from advanced ai L Hammond, A Chan, J Clifton, J Hoelscher-Obermaier, A Khan, ... arXiv preprint arXiv:2502.14143, 2025 | 273 | 2025 |
| Evolving curricula with regret-based environment design J Parker-Holder, M Jiang, M Dennis, M Samvelyan, J Foerster, ... International Conference on Machine Learning, 17473-17498, 2022 | 252 | 2022 |
| Replay-guided adversarial environment design M Jiang, M Dennis, J Parker-Holder, J Foerster, E Grefenstette, ... Advances in Neural Information Processing Systems 34, 1884-1897, 2021 | 185 | 2021 |
| Genie 2: A large-scale foundation world model J Parker-Holder, P Ball, J Bruce, V Dasagi, K Holsheimer, C Kaplanis, ... URL: https://deepmind. google/discover/blog/genie-2-a-large-scale-foundation …, 2024 | 162* | 2024 |
| Open-Endedness is Essential for Artificial Superhuman Intelligence E Hughes, M Dennis, J Parker-Holder, F Behbahani, A Mavalankar, Y Shi, ... arXiv preprint arXiv:2406.04268, 2024 | 153 | 2024 |
| Adversarial policies beat superhuman go AIs TT Wang, A Gleave, T Tseng, K Pelrine, N Belrose, J Miller, MD Dennis, ... International Conference on Machine Learning, 35655-35739, 2023 | 122* | 2023 |
| Quantifying Differences in Reward Functions A Gleave, M Dennis, S Legg, S Russell, J Leike (ICLR 2021) - Ninth International Conference on Learning Representations, 2021 | 108 | 2021 |
| MAESTRO: Open-ended environment design for multi-agent reinforcement learning M Samvelyan, A Khan, M Dennis, M Jiang, J Parker-Holder, J Foerster, ... arXiv preprint arXiv:2303.03376, 2023 | 68 | 2023 |
| A new formalism, method and open issues for zero-shot coordination J Treutlein, M Dennis, C Oesterheld, J Foerster International Conference on Machine Learning, 10413-10423, 2021 | 68 | 2021 |
| Benefits of Assistance over Reward Learning R Shah, P Freire, N Alex, R Freedman, D Krasheninnikov, L Chan, ... | 48 | |
| Refining Minimax Regret for Unsupervised Environment Design M Beukman, S Coward, M Matthews, M Fellows, M Jiang, M Dennis, ... arXiv preprint arXiv:2402.12284, 2024 | 28 | 2024 |
| Grounding aleatoric uncertainty for unsupervised environment design M Jiang, M Dennis, J Parker-Holder, A Lupu, H Küttler, E Grefenstette, ... Advances in Neural Information Processing Systems 35, 32868-32881, 2022 | 26 | 2022 |
| Stabilizing unsupervised environment design with a learned adversary I Mediratta, M Jiang, J Parker-Holder, M Dennis, E Vinitsky, T Rocktäschel Conference on Lifelong Learning Agents, 270-291, 2023 | 22 | 2023 |
| BAMDP Shaping: a Unified Framework for Intrinsic Motivation and Reward Shaping A Lidayan, MD Dennis, S Russell The Thirteenth International Conference on Learning Representations, 0 | 18* | |
| Cooperative and uncooperative institution designs: Surprises and problems in open-source game theory A Critch, M Dennis, S Russell arXiv preprint arXiv:2208.07006, 2022 | 17 | 2022 |
| minimax: Efficient Baselines for Autocurricula in JAX M Jiang, M Dennis, E Grefenstette, T Rocktäschel arXiv preprint arXiv:2311.12716, 2023 | 16 | 2023 |
| The Stretch Factor of Hexagon-Delaunay Triangulations L Perkovic, M Dennis, DT Türkoğlu Journal of Computational Geometry 12 (2), 86–125-86–125, 2021 | 12* | 2021 |