IKEA furniture assembly environment for long-horizon complex manipulation tasks
Original
Y. Lee, E. S. Hu, Z. Yang, A. Yin, and J. J. Lim · 1911
Earlier work this paper cites.
Torcs, the open racing car simulator
B. Wymann, E. Espié, C. Guionneau, C. Dimitrakakis, R. Coulom, and A. Sumner · 2000
Earlier work this paper cites.
The columbia grasp database
C. Goldfeder, M. Ciocarlie, H. Dang, and P. K. Allen · 2008
Earlier work this paper cites.
A list of household objects for robotic retrieval prioritized by people with als
Y. S. Choi, T. Deyle, T. Chen, J. D. Glass, and C. C. Kemp · 2009
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
The kit object models database: An object model database for object recognition, localization and manipulation in service robotics
A. Kasper, Z. Xue, and R. Dillmann · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Learning to select and generalize striking movements in robot table tennis
K. Mülling, J. Kober, O. Kroemer, and J. Peters · 2013
Earlier work this paper cites.
Actor-mimic: Deep multitask and transfer reinforcement learning
Original
E. Parisotto, J. L. Ba, and R. Salakhutdinov · 2015
Earlier work this paper cites.
Policy distillation
Original
A. A. Rusu, S. G. Colmenarejo, C. Gulcehre, G. Desjardins, J. Kirkpatrick, R. Pascanu, V. Mnih, K. Kavukcuoglu, and R. Hadsell · 2015
Earlier work this paper cites.
Benchmarking in manipulation research: The ycb object and model set and benchmarking protocols
Original
B. Calli, A. Walsman, A. Singh, S. Srinivasa, P. Abbeel, and A. M. Dollar · 2015
Earlier work this paper cites.
Deep learning for detecting robotic grasps
I. Lenz, H. Lee, and A. Saxena · 2015
Earlier work this paper cites.
The ycb object and model set: Towards common benchmarks for manipulation research
B. Calli, A. Singh, A. Walsman, S. Srinivasa, P. Abbeel, and A. M. Dollar · 2015
Earlier work this paper cites.
Leveraging big data for grasp planning
D. Kappler, J. Bohg, and S. Schaal · 2015
Earlier work this paper cites.
Trust region policy optimization
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz · 2015
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Earlier work this paper cites.
Openai gym
Original
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Earlier work this paper cites.
Rl$ˆ2$: Fast reinforcement learning via slow reinforcement learning
Original
Y. Duan, J. Schulman, X. Chen, P. L. Bartlett, I. Sutskever, and P. Abbeel · 2016
Earlier work this paper cites.
Learning to reinforcement learn, 2016
Original
J. X. Wang, Z. Kurth-Nelson, D. Tirumala, H. Soyer, J. Z. Leibo, R. Munos, C. Blundell, D. Kumaran, and M. Botvinick · 2016
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
C. Finn, I. Goodfellow, and S. Levine · 2016
Earlier work this paper cites.
More than a million ways to be pushed. a high-fidelity experimental dataset of planar pushing
K.-T. Yu, M. Bauza, N. Fazeli, and A. Rodriguez · 2016
Earlier work this paper cites.