Fetching the paper…
Reading the bibliography…
Connector insertion and many other tasks commonly found in modern manufacturing settings involve complex contact dynamics and friction.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” in NIPS , 1989, pp. 305–313
1989
Earlier work this paper cites.
A. Y. Ng, D. Harada, and S. Russell, “Policy invariance under reward transformations: Theory and application to reward shaping,” in ICML , 1999
1999
Earlier work this paper cites.
J. Nakanishi, J. Morimoto, G. Endo, G. Cheng, S. Schaal, and M. Kawato, “Learning from demonstration and adaptation of biped locomotion,” in Robotics and Autonomous Systems , vol. 47, no. 2-3, 2004, pp. 79–91
2004
Earlier work this paper cites.
J. Peters and S. Schaal, “Reinforcement learning of motor skills with policy gradients,” Neural Networks , vol. 21, no. 4, pp. 682–697, 2008
2008
Earlier work this paper cites.
J. Kober and J. Peter, “Policy search for motor primitives in robotics,” in NIPS , vol. 97, 2008, pp. 83–117
2008
Earlier work this paper cites.
B. D. Ziebart, A. Maas, J. A. Bagnell, and A. K. Dey, “Maximum Entropy Inverse Reinforcement Learning.” in AAAI , 2008, pp. 1433–1438
2008
Earlier work this paper cites.
J. Peters, K. Mülling, and Y. Altün, “Relative Entropy Policy Search,” in AAAI , 2010, pp. 1607–1612
2010
Earlier work this paper cites.
M. P. Deisenroth and C. E. Rasmussen, “PILCO: A model-based and data-efficient approach to policy search,” in ICML , 2011, pp. 465–472
2011
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing Atari with Deep Reinforcement Learning,” in NIPS Workshop on Deep Learning , 2013, pp. 1–9
2013
Earlier work this paper cites.
C. Daniel, M. Viering, J. Metz, O. Kroemer, and J. Peters, “Active Reward Learning,” in RSS , 2014
2014
Earlier work this paper cites.
R. Li, R. Platt, W. Yuan, A. Ten Pas, N. Roscup, M. A. Srinivasan, and E. Adelson, “Localization and Manipulation of Small Parts Using GelSight Tactile Sensing,” in IROS , 2014
2014
Earlier work this paper cites.
A. Giusti, J. J. Guzzi, D. C. Cirean et al. , “A Machine Learning Approach to Visual Perception of Forest Trails for Mobile Robots,” in IEEE Robotics and Automation Letters. , vol. 1, no. 2, 2015, pp. 2377–3766
2015
Earlier work this paper cites.
L. Pinto and A. Gupta, “Supersizing Self-supervision: Learning to Grasp from 50K Tries and 700 Robot Hours,” ICRA , 2016
2016
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-End Training of Deep Visuomotor Policies,” Journal of Machine Learning Research (JMLR) , vol. 17, no. 1, pp. 1334–1373, 2016
2016
Earlier work this paper cites.
S. Gu, T. Lillicrap, I. Sutskever, and S. Levine, “Continuous Deep Q-Learning with Model-based Acceleration,” in ICML , 2016
2016
Cited alongside, same era.
C. Finn, S. Levine, and P. Abbeel, “Guided Cost Learning: Deep Inverse Optimal Control via Policy Optimization,” in ICML , 2016
2016
Cited alongside, same era.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” in ICLR , 2016
2016
Cited alongside, same era.
H. Van Hasselt, A. Guez, and D. Silver, “Deep Reinforcement Learning with Double Q-learning,” in AAAI , 2016
2016
Cited alongside, same era.
I. Popov, N. Heess, T. Lillicrap et al. , “Data-efficient Deep Reinforcement Learning for Dexterous Manipulation,” CoRR , vol. abs/1704.0, 2017
2017
Cited alongside, same era.
H. Zhu, A. Gupta, A. Rajeswaran, S. Levine, and V. Kumar, “Dexterous Manipulation with Deep Reinforcement Learning: Efficient, General, and Low-Cost,” in ICRA , oct 2018
2018
Later among the works it cites.
D. Kalashnikov, A. Irpan, P. Pastor et al. , “QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation,” in CoRL , 2018
2018
Later among the works it cites.
G. Thomas, M. Chien, A. Tamar, J. A. Ojea, and P. Abbeel, “Learning Robotic Assembly from CAD,” in ICRA , 2018
2018
Later among the works it cites.
T. Hester, M. Vecerik, O. Pietquin et al. , “Learning from Demonstrations for Real World Reinforcement Learning,” in AAAI , 2018
2018
Later among the works it cites.
A. Nair, B. Mcgrew, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Overcoming Exploration in Reinforcement Learning with Demonstrations,” in ICRA , 2018
2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Chebotar, K. Hausman, M. Zhang, G. Sukhatme, S. Schaal, and S. Levine, “Combining Model-Based and Model-Free Updates for Trajectory-Centric Reinforcement Learning,” in ICML , 2017
2017
Cited alongside, same era.
A. Tamar, G. Thomas, T. Zhang, S. Levine, and P. Abbeel, “Learning from the hindsight plan — episodic mpc improvement,” in ICRA , 2017, pp. 336–343
2017
Cited alongside, same era.
T. Inoue, G. De Magistris, A. Munawar, T. Yokoya, and R. Tachibana, “Deep reinforcement learning for high precision assembly tasks,” in IROS , 2017, pp. 819–825
2017
Cited alongside, same era.
M. Andrychowicz, F. Wolski, A. Ray, J. Schneider, R. Fong, P. Welinder, B. Mcgrew, J. Tobin, P. Abbeel, and W. Zaremba, “Hindsight Experience Replay,” in NIPS , 2017
2017
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor,” in ICML , 2018
2018
Cited alongside, same era.
M. Hessel, J. Modayil, H. Van Hasselt, T. Schaul, G. Ostrovski, W. Dabney, D. Horgan, B. Piot, M. Azar, and D. Silver, “Rainbow: Combining improvements in deep reinforcement learning,” in AAAI , 2018
2018
Cited alongside, same era.
A. Nair, V. Pong, M. Dalal, S. Bahl, S. Lin, and S. Levine, “Visual Reinforcement Learning with Imagined Goals,” in NeurIPS , 2018
2018
Cited alongside, same era.
Later among the works it cites.
A. Rajeswaran, V. Kumar, A. Gupta, J. Schulman, E. Todorov, and S. Levine, “Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations,” in RSS , 2018
2018
Later among the works it cites.
J. Falco, Y. Sun, and M. Roa, “Robotic grasping and manipulation competition: Competitor feedback and lessons learned,” in Robotic Grasping and Manipulation , Y. Sun and J. Falco, Eds. Cham: Springer International Publishing, 2018, pp. 180–189
2018
Later among the works it cites.
S. Fujimoto, H. van Hoof, and D. Meger, “Addressing Function Approximation Error in Actor-Critic Methods,” in ICML , 2018
2018
Later among the works it cites.
V. Pong, S. Gu, M. Dalal, and S. Levine, “Temporal Difference Models: Model-Free Deep RL For Model-Based Control,” in ICLR , 2018
2018
Later among the works it cites.
T. Johannink, S. Bahl, A. Nair, J. Luo, A. Kumar, M. Loskyll, J. Aparicio Ojea, E. Solowjow, and S. Levine, “Residual Reinforcement Learning for Robot Control,” in ICRA , 2019
2019
Closest in time.
M. Zhang, S. Vikram, L. Smith, P. Abbeel, M. J. Johnson, and S. Levine, “SOLAR: Deep Structured Representations for Model-Based Reinforcement Learning,” in ICML , aug 2019
2019
Closest in time.
A. Singh, L. Yang, K. Hartikainen, C. Finn, and S. Levine, “End-to-End Robotic Reinforcement Learning without Reward Engineering,” in RSS , 2019
2019
Closest in time.
J. Luo, E. Solowjow, C. Wen, J. Aparicio Ojea, A. Agogino, A. Tamar, and A. P, “Reinforcement learning on variable impedance controller for high-precision robotic assembly,” in ICRA , 2019
2019
Closest in time.