Fetching the paper…
Reading the bibliography…
Manipulation skills involving contact and friction are inherent to many robotics tasks.
C. Pottle, “The digital adaptive control of a linear process modulated by random noise,” IEEE Transactions on Automatic Control , vol. 8, no. 3, pp. 228–234, 1963
1963
Earlier work this paper cites.
C. Grubin, “Derivation of the quaternion scheme via the euler axis and angle,” Journal of Spacecraft and Rockets , vol. 7, no. 10, pp. 1261–1263, 1970
1970
Earlier work this paper cites.
D. E. Whitney, “Quasi-static assembly of compliantly supported rigid parts,” Journal of Dynamic Systems, Measurement, and Control , vol. 104, no. 1, pp. 65–77, 1982
1982
Earlier work this paper cites.
S. Schaal, “Learning from demonstration,” in Advances in neural information processing systems , 1997, pp. 1040–1046
1997
Earlier work this paper cites.
J. Peters and S. Schaal, “Applying the episodic natural actor-critic architecture to motor primitive learning.” in ESANN , 2007, pp. 295–300
2007
Earlier work this paper cites.
J. Peters and S. Schaal, “Reinforcement learning of motor skills with policy gradients,” Neural networks , vol. 21, no. 4, pp. 682–697, 2008
2008
Earlier work this paper cites.
M. D. Shuster, “The nature of the quaternion,” The Journal of the Astronautical Sciences , vol. 56, no. 3, pp. 359–373, 2008
2008
Earlier work this paper cites.
J. Kober and J. R. Peters, “Policy search for motor primitives in robotics,” in Advances in neural information processing systems , 2009, pp. 849–856
2009
Earlier work this paper cites.
S. Schaal, “The sl simulation and real-time control software package,” Technical report, University of Southern California, Tech. Rep., 2009
2009
Earlier work this paper cites.
F. Sehnke, C. Osendorfer, T. Rückstieß, A. Graves, J. Peters, and J. Schmidhuber, “Parameter-exploring policy gradients.” Neural Networks , 2010
2010
Earlier work this paper cites.
F. Stulp, E. A. Theodorou, and S. Schaal, “Reinforcement learning with sequences of motion primitives for robust manipulation,” IEEE Transactions on robotics , vol. 28, no. 6, pp. 1360–1370, 2012
2012
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
A. J. Ijspeert, J. Nakanishi, H. Hoffmann, P. Pastor, and S. Schaal, “Dynamical movement primitives: learning attractor models for motor behaviors,” Neural computation , vol. 25, no. 2, pp. 328–373, 2013
2013
Cited alongside, same era.
I. Havoutis and S. Ramamoorthy, “Motion planning and reactive control on learnt skill manifolds,” The International Journal of Robotics Research , vol. 32, no. 9-10, pp. 1120–1150, 2013
2013
Cited alongside, same era.
A. Ude, B. Nemec, T. Petrić, and J. Morimoto, “Orientation in cartesian space dynamic movement primitives,” in IEEE International Conference on Robotics and Automation . IEEE, 2014, pp. 2997–3004
2014
Cited alongside, same era.
A. Gams, B. Nemec, A. J. Ijspeert, and A. Ude, “Coupling movement primitives: Interaction with the environment and bimanual tasks,” IEEE Transactions on Robotics , vol. 30, no. 4, pp. 816–830, 2014
2014
Cited alongside, same era.
Z. Zhu, H. Hu, and D. Gu, “Robot performing peg-in-hole operations by learning from human demonstration,” in 2018 10th Computer Science and Electronic Engineering (CEEC) . IEEE, 2018, pp. 30–35
2018
Later among the works it cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in Proceedings of the 35th International Conference on Machine Learning , J. Dy and A. Krause, Eds., vol. 80, 2018, pp. 1861–1870
2018
Later among the works it cites.
O. Kroemer, S. Niekum, and G. Konidaris, “A review of robot learning for manipulation: Challenges,” Representations, and Algorithms. , 2019
2019
Later among the works it cites.
J. Luo, E. Solowjow, C. Wen, J. A. Ojea, A. M. Agogino, A. Tamar, and P. Abbeel, “Reinforcement learning on variable impedance controller for high-precision robotic assembly,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 3080–3087
2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Gams, T. Petric, B. Nemec, and A. Ude, “Learning and adaptation of periodic motion primitives based on force feedback and human coaching interaction,” in 2014 IEEE-RAS International Conference on Humanoid Robots . IEEE, 2014, pp. 166–171
2014
Cited alongside, same era.
N. Likar, B. Nemec, L. Žlajpah, S. Ando, and A. Ude, “Adaptation of bimanual assembly tasks using iterative learning framework,” in 2015 IEEE-RAS 15th International Conference on Humanoid Robots (Humanoids) . IEEE, 2015, pp. 771–776
2015
Cited alongside, same era.
F. J. Abu-Dakka, B. Nemec, J. A. Jørgensen, T. R. Savarimuthu, N. Krüger, and A. Ude, “Adaptation of manipulation skills in physical contact with the environment to reference force profiles,” Autonomous Robots , vol. 39, no. 2, pp. 199–217, 2015
2015
Cited alongside, same era.
V. Kumar, E. Todorov, and S. Levine, “Optimal control with learned local models: Application to dexterous manipulation,” in 2016 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2016, pp. 378–383
2016
Cited alongside, same era.
2017
Cited alongside, same era.
T. Inoue, G. De Magistris, A. Munawar, T. Yokoya, and R. Tachibana, “Deep reinforcement learning for high precision assembly tasks,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2017, pp. 819–825
2017
Cited alongside, same era.
2017
Cited alongside, same era.
G. Sutanto, Z. Su, S. Schaal, and F. Meier, “Learning sensor feedback models from demonstrations via phase-modulated neural networks,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 1142–1149
2018
Cited alongside, same era.
Later among the works it cites.
T. Johannink, S. Bahl, A. Nair, J. Luo, A. Kumar, M. Loskyll, J. A. Ojea, E. Solowjow, and S. Levine, “Residual reinforcement learning for robot control,” in 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 6023–6029
2019
Later among the works it cites.
S. Guadarrama, A. Korattikara, O. Ramirez, P. Castro, E. Holly, S. Fishman, K. Wang, E. Gonina, N. Wu, E. Kokiopoulou, L. Sbaiz, J. Smith, G. Bartók, J. Berent, C. Harris, V. Vanhoucke, and E. Brevdo, “Tf-agents: A library for reinforcement learning in tensorflow,” https://github.com/tensorflow/agents
2019
Later among the works it cites.
Y. Huang, F. J. Abu-Dakka, J. Silvério, and D. G. Caldwell, “Toward orientation learning and adaptation in cartesian space,” IEEE Transactions on Robotics , 2020
2020
Closest in time.
G. Schoettler, A. Nair, J. Luo, S. Bahl, J. Aparicio Ojea, E. Solowjow, and S. Levine, “Deep reinforcement learning for industrial insertion tasks with visual inputs and natural rewards,” in International Conference on Intelligent Robots and Systems (IROS) , 2020, pp. 5548–5555
2020
Closest in time.
L. Koutras and Z. Doulgeri, “Dynamic movement primitives for moving goals with temporal scaling adaptation,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2020, pp. 144–150
2020
Closest in time.
A. Zeng, S. Song, J. Lee, A. Rodriguez, and T. Funkhouser, “Tossingbot: Learning to throw arbitrary objects with residual physics,” IEEE Transactions on Robotics , vol. 36, no. 4, pp. 1307–1319, 2020
2020
Closest in time.
2020
Closest in time.
M. A. Lee, C. Florensa, J. Tremblay, N. Ratliff, A. Garg, F. Ramos, and D. Fox, “Guided uncertainty-aware policy optimization: Combining learning and model-based strategies for sample-efficient policy learning,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2020, pp. 7505–7512
2020
Closest in time.