Off-policy deep reinforcement learning without exploration
Original
Scott Fujimoto, David Meger, and Doina Precup · 2018
Later among the works it cites.
Curiosity driven exploration of learned disentangled goal spaces
Adrien Laversanne-Finot, Alexandre Pere, and Pierre-Yves Oudeyer · 2018
Later among the works it cites.
Unicorn: Continual learning with a universal, off-policy agent
Original
Daniel J. Mankowitz, Augustin Zídek, André Barreto, Dan Horgan, Matteo Hessel, John Quan, Junhyuk Oh, Hado van Hasselt, David Silver, and Tom Schaul · 2018
Later among the works it cites.
Visual reinforcement learning with imagined goals
Ashvin V Nair, Vitchyr Pong, Murtaza Dalal, Shikhar Bahl, Steven Lin, and Sergey Levine · 2018
Later among the works it cites.
Multi-goal reinforcement learning: Challenging robotics environments and request for research
Original
Matthias Plappert, Marcin Andrychowicz, Alex Ray, Bob McGrew, Bowen Baker, Glenn Powell, Jonas Schneider, Josh Tobin, Maciek Chociej, Peter Welinder, et al · 2018
Later among the works it cites.
Monet: Unsupervised scene decomposition and representation
Original
Christopher P Burgess, Loic Matthey, Nicholas Watters, Rishabh Kabra, Irina Higgins, Matt Botvinick, and Alexander Lerchner · 2019
Later among the works it cites.
Actrce: Augmenting experience via teacher’s advice for multi-goal reinforcement learning, 2019
Harris Chan, Yuhuai Wu, Jamie Kiros, Sanja Fidler, and Jimmy Ba · 2019
Later among the works it cites.
Baby{AI}: First Steps Towards Grounded Language Learning With a Human In the Loop
Maxime Chevalier-Boisvert, Dzmitry Bahdanau, Salem Lahlou, Lucas Willems, Chitwan Saharia, Thien Huu Nguyen, and Yoshua Bengio · 2019
Later among the works it cites.
Self-educated language agent with hindsight experience replay for instruction following
Original
Geoffrey Cideron, Mathieu Seurin, Florian Strub, and Olivier Pietquin · 2019
Later among the works it cites.
CURIOUS: intrinsically motivated modular multi-goal reinforcement learning
Cédric Colas, Pierre-Yves Oudeyer, Olivier Sigaud, Pierre Fournier, and Mohamed Chetouani · 2019
Later among the works it cites.
Go-explore: a new approach for hard-exploration problems
Original
Adrien Ecoffet, Joost Huizinga, Joel Lehman, Kenneth O Stanley, and Jeff Clune · 2019
Later among the works it cites.
From Language to Goals: Inverse Reinforcement Learning for Vision-Based Instruction Following
Justin Fu, Anoop Korattikara, Sergey Levine, and Sergio Guadarrama · 2019
Later among the works it cites.
Multi-object representation learning with iterative variational inference
Original
Klaus Greff, Raphaël Lopez Kaufmann, Rishab Kabra, Nick Watters, Chris Burgess, Daniel Zoran, Loic Matthey, Matthew Botvinick, and Alexander Lerchner · 2019
Later among the works it cites.
Emergent systematic generalization in a situated agent, 2019
Felix Hill, Andrew Lampinen, Rosalia Schneider, Stephen Clark, Matthew Botvinick, James L. McClelland, and Adam Santoro · 2019
Later among the works it cites.
Measuring compositional generalization: A comprehensive method on realistic data, 2019
Daniel Keysers, Nathanael Schärli, Nathan Scales, Hylke Buisman, Daniel Furrer, Sergii Kashubin, Nikola Momchev, Danila Sinopalnikov, Lukasz Stafiniak, Tibor Tihon, Dmitry Tsarkov, Xiao Wang, Marc van Zee, and Olivier Bousquet · 2019
Later among the works it cites.
Extending machine language models toward human-level language understanding
Original
James L McClelland, Felix Hill, Maja Rudolph, Jason Baldridge, and Hinrich Schütze · 2019
Later among the works it cites.
Contextual imagined goals for self-supervised robotic learning
Original
Ashvin Nair, Shikhar Bahl, Alexander Khazatsky, Vitchyr Pong, Glen Berseth, and Sergey Levine · 2019
Later among the works it cites.
Vision-based navigation with language-based assistance via imitation learning with indirect intervention
Khanh Nguyen, Debadeepta Dey, Chris Brockett, and Bill Dolan · 2019
Later among the works it cites.
Skew-fit: State-covering self-supervised reinforcement learning
Original
Vitchyr H Pong, Murtaza Dalal, Steven Lin, Ashvin Nair, Shikhar Bahl, and Sergey Levine · 2019
Later among the works it cites.
Automated curricula through setter-solver interactions
Original
Sebastien Racaniere, Andrew K Lampinen, Adam Santoro, David P Reichert, Vlad Firoiu, and Timothy P Lillicrap · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Later among the works it cites.
Self-supervised learning of distance functions for goal-conditioned reinforcement learning
Original
Srinivas Venkattaramanujam, Eric Crawford, Thang Doan, and Doina Precup · 2019
Later among the works it cites.
Good-enough compositional data augmentation
Jacob Andreas · 2020
Closest in time.
Exploratory play, rational action, and efficient search
Junyi Chu and Laura Schulz · 2020
Closest in time.
Deep sets for generalization in rl, 2020
Tristan Karch, Cédric Colas, Laetitia Teodorescu, Clément Moulin-Frier, and Pierre-Yves Oudeyer · 2020
Closest in time.
Alfred: A benchmark for interpreting grounded instructions for everyday tasks
Mohit Shridhar, Jesse Thomason, Daniel Gordon, Yonatan Bisk, Winson Han, Roozbeh Mottaghi, Luke Zettlemoyer, and Dieter Fox · 2020
Closest in time.