Fetching the paper…
Reading the bibliography…
Recent progress in Quality Diversity Reinforcement Learning (QD-RL) has enabled learning a collection of behaviorally diverse, high performing policies.
Robots that can adapt like animals
Antoine Cully, Jeff Clune, Danesh Tarapore, and Jean-Baptiste Mouret · 2015
Earlier work this paper cites.
Illuminating search spaces by mapping elites
Jean-Baptiste Mouret and Jeff Clune · 2015
Earlier work this paper cites.
David Ha, Andrew M. Dai, and Quoc V. Le · 2016
Earlier work this paper cites.
Generative adversarial policy networks for behavioural repertoire
Marija Jegorova, Stéphane Doncieux, and Timothy M. Hospedales · 2018
Earlier work this paper cites.
Using centroidal voronoi tessellations to scale up the multidimensional archive of phenotypic elites algorithm
Vassilis Vassiliades, Konstantinos I. Chatzilygeroudis, and Jean-Baptiste Mouret · 2018
Earlier work this paper cites.
Discovering the elite hypervolume by leveraging interspecies correlation
Vassilis Vassiliades and Jean-Baptiste Mouret · 2018
Earlier work this paper cites.
Go-explore: a new approach for hard-exploration problems
Adrien Ecoffet, Joost Huizinga, Joel Lehman, Kenneth O. Stanley, and Jeff Clune · 2019
Earlier work this paper cites.
Continual learning with hypernetworks
Johannes von Oswald, Christian Henning, João Sacramento, and Benjamin F. Grewe · 2019
Earlier work this paper cites.
Graph hypernetworks for neural architecture search
Chris Zhang, Mengye Ren, and Raquel Urtasun · 2019
Earlier work this paper cites.
Discovering representations for black-box optimization
Adam Gaier, Alexander Asteroth, and Jean-Baptiste Mouret · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Narain Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2020
Cited alongside, same era.
Decision transformer: Reinforcement learning via sequence modeling
Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Michael Laskin, P. Abbeel, A. Srinivas, and Igor Mordatch · 2021
Cited alongside, same era.
Diffusion models beat gans on image synthesis
Prafulla Dhariwal and Alexander Quinn Nichol · 2021
Cited alongside, same era.
Differentiable quality diversity
Matthew Fontaine and Stefanos Nikolaidis · 2021
Cited alongside, same era.
Brax - a differentiable physics engine for large scale rigid body simulation, 2021
C. Daniel Freeman, Erik Frey, Anton Raichuk, Sertan Girgin, Igor Mordatch, and Olivier Bachem · 2021
Cited alongside, same era.
Efficiently learning small policies for locomotion and manipulation
Shashank Hegde and Gaurav S Sukhatme · 2022
Later among the works it cites.
Elucidating the design space of diffusion-based generative models
Tero Karras, Miika Aittala, Timo Aila, and Samuli Laine · 2022
Later among the works it cites.
Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps
Cheng Lu, Yuhao Zhou, Fan Bao, Jianfei Chen, Chongxuan Li, and Jun Zhu · 2022
Later among the works it cites.
Diversity policy gradient for sample efficient quality-diversity optimization
Thomas Pierrot, Valentin Macé, Félix Chalumeau, Arthur Flajolet, Geoffrey Cideron, Karim Beguir, Antoine Cully, Olivier Sigaud, and Nicolas Perrin-Gilbert · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Boris Knyazev, Michal Drozdzal, Graham W Taylor, and Adriana Romero · 2021
Cited alongside, same era.
Policy gradient assisted map-elites
Olle Nilsson and Antoine Cully · 2021
Cited alongside, same era.
Policy manifold search: exploring the manifold hypothesis for diversity-based neuroevolution
Nemanja Rakicevic, Antoine Cully, and Petar Kormushev · 2021
Cited alongside, same era.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2021
Cited alongside, same era.
Deep surrogate assisted generation of environments
Varun Bhatt, Bryon Tjanaka, Matthew C. Fontaine, and Stefanos Nikolaidis · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models, 2022
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Eric Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Zhao, Yanping Huang, Andrew Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei · 2022
Cited alongside, same era.
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
Approximating gradients for differentiable quality diversity in reinforcement learning
Bryon Tjanaka, Matthew C. Fontaine, Julian Togelius, and Stefanos Nikolaidis · 2022
Later among the works it cites.
Approximating gradients for differentiable quality diversity in reinforcement learning
Bryon Tjanaka, Matthew C Fontaine, Julian Togelius, and Stefanos Nikolaidis · 2022
Later among the works it cites.
Scaling covariance matrix adaptation map-annealing to high-dimensional controllers
Bryon Tjanaka, Matthew Christopher Fontaine, Aniruddha Kalkar, and Stefanos Nikolaidis · 2022
Later among the works it cites.
Proximal policy gradient arborescence for quality diversity reinforcement learning
Sumeet Batra, Bryon Tjanaka, Matthew C Fontaine, Aleksei Petrenko, Stefanos Nikolaidis, and Gaurav Sukhatme · 2023
Closest in time.
Map-elites with descriptor-conditioned gradients and archive distillation into a single policy
Maxence Faldor, Félix Chalumeau, Manon Flageat, and Antoine Cully · 2023
Closest in time.
Valentin Mac’e, Raphael Boige, Félix Chalumeau, Thomas Pierrot, Guillaume Richard, and Nicolas Perrin-Gilbert · 2023
Closest in time.