Fetching the paper…
Reading the bibliography…
Robotic insertion tasks are characterized by contact and friction mechanics, making them challenging for conventional feedback control methods due to unmodeled physical effects.
D. E. Whitney, “Force feedback control of manipulator fine motions,” ” 1977
1977
Earlier work this paper cites.
——, “Quasi-static assembly of compliantly supported rigid parts,”
1982
Earlier work this paper cites.
S. R. Chhatpar and M.S. Branicky, “Search strategies for peg-in-hole assemblies with position uncertainty,” in
2001
Earlier work this paper cites.
J. Peters, K. Mülling, and Y. Altün, “Relative Entropy Policy Search,” in
2010
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “MuJoCo: A physics engine for model-based control,” in
2012
Earlier work this paper cites.
H. Park, J.-H. Bae, J.-H. Park, M.-H. Baeg, and J. Park, “Intuitive peg-in-hole assembly strategy with a compliant manipulator,” in
2013
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing Atari with Deep Reinforcement Learning,” in
2013
Earlier work this paper cites.
R. Li, R. Platt, W. Yuan, A. Ten Pas, N. Roscup, M. A. Srinivasan, and E. Adelson, “Localization and Manipulation of Small Parts Using GelSight Tactile Sensing,” in
2014
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, and D. Hassabis, “Mastering the game of Go with deep neural networks and tree search,”
2016
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-End Training of Deep Visuomotor Policies,”
2016
Earlier work this paper cites.
Y. Duan, J. Schulman, X. Chen, P. L. Bartlett, I. Sutskever, and P. Abbeel, “RL$
2016
Earlier work this paper cites.
M. Andrychowicz, M. Denil, S. G. Colmenarejo, M. W. Hoffman, D. Pfau, T. Schaul, B. Shillingford, and N. De Freitas, “Learning to learn by gradient descent by gradient descent,” in
2016
Earlier work this paper cites.
A. Tamar, G. Thomas, T. Zhang, S. Levine, and P. Abbeel, “Learning from the hindsight plan — episodic mpc improvement,” in
2017
Earlier work this paper cites.
T. Inoue, G. De Magistris, A. Munawar, T. Yokoya, and R. Tachibana, “Deep reinforcement learning for high precision assembly tasks,” in
2017
Earlier work this paper cites.
M. Večerík, T. Hester, J. Scholz, F. Wang, O. Pietquin, B. Piot, N. Heess, T. Rothörl, T. Lampe, and M. Riedmiller, “Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards,”
2017
Cited alongside, same era.
F. Sadeghi and S. Levine, “CAD 2 RL: Real Single-Image Flight Without a Single Real Image,” in
2017
Cited alongside, same era.
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel, “Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World,”
2017
Cited alongside, same era.
C. Finn, P. Abbeel, and S. Levine, “Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks,” in
2017
Cited alongside, same era.
J. Falco, Y. Sun, and M. Roa, “Robotic grasping and manipulation competition: Competitor feedback and lessons learned,” in
2018
A. Srinivas, A. Jabri, P. Abbeel, S. Levine, and C. Finn, “Universal Planning Networks,” in
2018
Later among the works it cites.
J. A. Marvel, R. Bostelman, and J. Falco, “Multi-Robot Assembly Strategies and Metrics,”
2018
Later among the works it cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor,” in
2018
Later among the works it cites.
K. Rakelly, A. Zhou, C. Finn, S. Levine, and D. Quillen, “Efficient off-policy meta-reinforcement learning via probabilistic context variables,” in
2019
Later among the works it cites.
J. Luo, E. Solowjow, C. Wen, J. Aparicio Ojea, A. Agogino, A. Tamar, and P. Abbeel, “Reinforcement learning on variable impedance controller for high-precision robotic assembly,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
G. Kahn, A. Villaflor, B. Ding, P. Abbeel, and S. Levine, “Self-supervised Deep Reinforcement Learning with Generalized Computation Graphs for Robot Navigation,” in
2018
Cited alongside, same era.
C. Florensa, D. Held, M. Wulfmeier, and P. Abbeel, “Reverse Curriculum Generation for Reinforcement Learning,” in
2018
Cited alongside, same era.
T. Hester, M. Vecerik, O. Pietquin, M. Lanctot, T. Schaul, B. Piot, D. Horgan, J. Quan, A. Sendonaris, G. Dulac-Arnold, I. Osband, J. Agapiou, J. Z. Leibo, and A. Gruslys, “Learning from Demonstrations for Real World Reinforcement Learning,” in
2018
Cited alongside, same era.
A. Nair, B. Mcgrew, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Overcoming Exploration in Reinforcement Learning with Demonstrations,” in
2018
Cited alongside, same era.
A. Rajeswaran, V. Kumar, A. Gupta, J. Schulman, E. Todorov, and S. Levine, “Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations,” in
2018
Cited alongside, same era.
T. Silver, K. Allen, J. Tenenbaum, and L. Kaelbling, “Residual Policy Learning,”
2018
Cited alongside, same era.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-Real Transfer of Robotic Control with Dynamics Randomization,” in
2018
Cited alongside, same era.
G. Schoettler, A. Nair, J. Luo, S. Bahl, J. Aparicio Ojea, E. Solowjow, and S. Levine, “Deep Reinforcement Learning for Industrial Insertion Tasks with Visual Inputs and Natural Rewards,”
2019
Later among the works it cites.
T. Johannink, S. Bahl, A. Nair, J. Luo, A. Kumar, M. Loskyll, J. Aparicio Ojea, E. Solowjow, and S. Levine, “Residual Reinforcement Learning for Robot Control,” in
2019
Later among the works it cites.
OpenAI, I. Akkaya, M. Andrychowicz, M. Chociej, M. Litwin, B. McGrew, A. Petron, A. Paino, M. Plappert, G. Powell, R. Ribas, J. Schneider, N. Tezak, J. Tworek, P. Welinder, L. Weng, Q. Yuan, W. Zaremba, and L. Zhang, “Solving Rubik’s Cube with a Robot Hand,” ” oct 2019
2019
Later among the works it cites.
F. Ramos, R. Carvalhaes Possas, and D. Fox, “BayesSim: adaptive domain randomization via probabilistic inference for robotics simulators,” in
2019
Later among the works it cites.
B. Mehta, M. Diaz, F. Golemo, C. J. Pal, and L. Paull, “Active Domain Randomization,” in
2019
Later among the works it cites.
W. Zhou, L. Pinto, and A. Gupta, “Environment Probing Interaction Policies,” in
2019
Later among the works it cites.
Y. Chebotar, A. Handa, V. Makoviychuk, M. Macklin, J. Issac, N. Ratliff, and D. Fox, “Closing the Sim-to-Real Loop: Adapting Simulation Randomization with Real World Experience,” in
2019
Later among the works it cites.
H. Bharadhwaj, Z. Wang, Y. Bengio, and L. Paull, “A data-efficient framework for training and sim-to-real transfer of navigation policies,” in
2019
Later among the works it cites.
M. Wortsman, K. Ehsani, M. Rastegari, A. Farhadi, and R. Mottaghi, “Learning to Learn How to Learn: Self-Adaptive Visual Navigation Using Meta-Learning,” in
2019
Later among the works it cites.