Fetching the paper…
Reading the bibliography…
We introduce RLDS (Reinforcement Learning Datasets), an ecosystem for recording, replaying, manipulating, annotating and sharing data in the context of Sequential Decision Making (SDM) including Reinforcement Learning (RL), Learning from Demonstrations, Offline RL or Imitation Learning.
Efficient training of artificial neural networks for autonomous navigation
D. A. Pomerleau · 1991
Earlier work this paper cites.
Running experiments on amazon mechanical turk
G. Paolacci, J. Chandler, and P. G. Ipeirotis · 2010
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling · 2013
Earlier work this paper cites.
Openml: networked science in machine learning
J. Vanschoren, J. N. Van Rijn, B. Bischl, and L. Torgo · 2014
Earlier work this paper cites.
C. Beattie, J. Z. Leibo, D. Teplyashin, T. Ward, M. Wainwright, H. Küttler, A. Lefrancq, S. Green, V. Valdés, A. Sadik, et al · 2016
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Earlier work this paper cites.
Generative adversarial imitation learning
J. Ho and S. Ermon · 2016
Earlier work this paper cites.
SQuAD: 100,000+ questions for machine comprehension of text
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang · 2016
Earlier work this paper cites.
Deep reinforcement learning from human preferences
P. Christiano, J. Leike, T. B. Brown, M. Martic, S. Legg, and D. Amodei · 2017
Earlier work this paper cites.
UCI machine learning repository, 2017
D. Dua and C. Graff · 2017
Earlier work this paper cites.
Leveraging demonstrations for deep reinforcement learning on robotics problems with sparse rewards
M. Vecerik, T. Hester, J. Scholz, F. Wang, O. Pietquin, B. Piot, N. Heess, T. Rothörl, T. Lampe, and M. Riedmiller · 2017
Earlier work this paper cites.
Addressing function approximation error in actor-critic methods
S. Fujimoto, H. Hoof, and D. Meger · 2018
Earlier work this paper cites.
TF-Agents: A library for reinforcement learning in tensorflow
S. Guadarrama, A. Korattikara, O. Ramirez, P. Castro, E. Holly, S. Fishman, K. Wang, E. Gonina, N. Wu, E. Kokiopoulou, L. Sbaiz, J. Smith, G. Bartók, J. Berent, C. Harris, V. Vanhoucke, and E. Brevdo · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Earlier work this paper cites.
Deep q-learning from demonstrations
T. Hester, M. Vecerik, O. Pietquin, M. Lanctot, T. Schaul, B. Piot, D. Horgan, J. Quan, A. Sendonaris, I. Osband, et al · 2018
Earlier work this paper cites.
I. Kostrikov, K. K. Agrawal, D. Dwibedi, S. Levine, and J. Tompson · 2018
Cited alongside, same era.
Roboturk: A crowdsourcing platform for robotic skill learning through imitation
A. Mandlekar, Y. Zhu, A. Garg, J. Booher, M. Spero, A. Tung, J. Gao, J. Emmons, A. Gupta, E. Orbay, et al · 2018
Cited alongside, same era.
Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine · 2018
Cited alongside, same era.
Multiple interactions made easy (mime): Large scale demonstrations data for imitation
P. Sharma, L. Mohan, L. Pinto, and A. Gupta · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Cited alongside, same era.
D4rl: Datasets for deep data-driven reinforcement learning
J. Fu, A. Kumar, O. Nachum, G. Tucker, and S. Levine · 2020
Later among the works it cites.
Rl unplugged: Benchmarks for offline reinforcement learning
C. Gulcehre, Z. Wang, A. Novikov, T. L. Paine, S. G. Colmenarejo, K. Zolna, R. Agarwal, J. Merel, D. Mankowitz, C. Paduraru, et al · 2020
Later among the works it cites.
Acme: A research framework for distributed reinforcement learning
M. Hoffman, B. Shahriari, J. Aslanides, G. Barth-Maron, F. Behbahani, T. Norman, A. Abdolmaleki, A. Cassirer, F. Yang, K. Baumli, et al · 2020
Later among the works it cites.
The nethack learning environment
H. Küttler, N. Nardelli, A. H. Miller, R. Raileanu, M. Selvatici, E. Grefenstette, and T. Rocktäschel · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Torabi, G. Warnell, and P. Stone · 2018
Cited alongside, same era.
Self-attentional credit assignment for transfer in reinforcement learning
J. Ferret, R. Marinier, M. Geist, and O. Pietquin · 2019
Cited alongside, same era.
Learning from a learner
A. Jacq, M. Geist, A. Paiva, and O. Pietquin · 2019
Cited alongside, same era.
Imitation learning via off-policy distribution matching
I. Kostrikov, O. Nachum, and J. Tompson · 2019
Cited alongside, same era.
Simitate: A hybrid imitation learning benchmark
R. Memmesheimer, I. Kramer, V. Seib, and D. Paulus · 2019
Cited alongside, same era.
dm_env: A python interface for reinforcement learning environments, 2019
A. Muldal, Y. Doron, J. Aslanides, T. Harley, T. Ward, and S. Liu · 2019
Cited alongside, same era.
Making efficient use of demonstrations to solve hard exploration problems
T. L. Paine, C. Gulcehre, B. Shahriari, M. Denil, M. Hoffman, H. Soyer, R. Tanburn, S. Kapturowski, N. Rabinowitz, D. Williams, et al · 2019
Cited alongside, same era.
S. Levine, A. Kumar, G. Tucker, and J. Fu · 2020
Later among the works it cites.
Reinforcement learning with videos: Combining offline observations with interaction
K. Schmeckpeper, O. Rybkin, K. Daniilidis, S. Levine, and C. Finn · 2020
Later among the works it cites.
Robust imitation learning from noisy demonstrations
V. Tangkaratt, N. Charoenphakdee, and M. Sugiyama · 2020
Later among the works it cites.
The magical benchmark for robust imitation
S. Toyer, R. Shah, A. Critch, and S. Russell · 2020
Later among the works it cites.
Transporter networks: Rearranging the visual world for robotic manipulation
A. Zeng, P. Florence, J. Tompson, S. Welker, J. Chien, M. Attarian, T. Armstrong, I. Krasin, D. Duong, V. Sindhwani, et al · 2020
Later among the works it cites.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, and R. Martín-Martín · 2020
Later among the works it cites.
Offline reinforcement learning with pseudometric learning
R. Dadashi, S. Rezaeifar, N. Vieillard, L. Hussenot, O. Pietquin, and M. Geist · 2021
Closest in time.
An empirical investigation of the challenges of real-world reinforcement learning, 2021
G. Dulac-Arnold, N. Levine, D. J. Mankowitz, J. Li, C. Paduraru, S. Gowal, and T. Hester · 2021
Closest in time.
Benchmarks for deep off-policy evaluation
J. Fu, M. Norouzi, O. Nachum, G. Tucker, Z. Wang, A. Novikov, M. Yang, M. R. Zhang, Y. Chen, A. Kumar, et al · 2021
Closest in time.
Show me the way: Intrinsic motivation from demonstrations
L. Hussenot, R. Dadashi, M. Geist, and O. Pietquin · 2021
Closest in time.
What matters in learning from offline human demonstrations for robot manipulation
A. Mandlekar, D. Xu, J. Wong, S. Nasiriany, C. Wang, R. Kulkarni, L. Fei-Fei, S. Savarese, Y. Zhu, and R. Martín-Martín · 2021
Closest in time.