L. P. Kaelbling, “Learning to achieve goals,” in International Joint Conference on Artificial Intelligence (IJCAI) , vol. vol.2, 1993, pp. 1094 – 8
1993
Earlier work this paper cites.
S. Ekvall and D. Kragic, “Interactive grasp learning based on human demonstration,” in IEEE International Conference on Robotics and Automation, 2004. Proceedings. ICRA’04. 2004 , vol. 4. IEEE, 2004, pp. 3519–3524
2004
Earlier work this paper cites.
J. Peters and S. Schaal, “Reinforcement learning of motor skills with policy gradients,” Neural Networks , vol. 21, no. 4, pp. 682 – 697, 2008, robotics and Neuroscience
2008
Earlier work this paper cites.
J. Peters and S. Schaal, “Reinforcement learning of motor skills with policy gradients,” Neural Networks , vol. 21, no. 4, pp. 682–697, 2008
2008
Earlier work this paper cites.
O. Kroemer, R. Detry, J. Piater, and J. Peters, “Combining Active Learning and Reactive Control for Robot Grasping,” Robotics and Autonomous systems , vol. 58, no. 9, pp. 1105–1116, 2010
2010
Earlier work this paper cites.
J. Bohg and D. Kragic, “Learning grasping points with shape context,” Robotics and Autonomous Systems , vol. 58, no. 4, pp. 362–377, 2010
2010
Earlier work this paper cites.
J. Peters, K. Mülling, and Y. Altün, “Relative Entropy Policy Search,” in AAAI Conference on Artificial Intelligence , 2010, pp. 1607–1612
2010
Earlier work this paper cites.
M. P. Deisenroth and C. E. Rasmussen, “PILCO: A model-based and data-efficient approach to policy search,” in International Conference on Machine Learning (ICML) , 2011, pp. 465–472
2011
Earlier work this paper cites.
M. Krainin, B. Curless, and D. Fox, “Autonomous generation of complete 3d object models using next best view manipulation planning,” in 2011 IEEE International Conference on Robotics and Automation . IEEE, 2011, pp. 5031–5037
2011
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “MuJoCo: A physics engine for model-based control,” in IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2012, pp. 5026–5033
2012
Earlier work this paper cites.
J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement learning in robotics: A survey,” The International Journal of Robotics Research , vol. 32, no. 11, pp. 1238–1274, 2013
2013
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing Atari with Deep Reinforcement Learning,” in NIPS Workshop on Deep Learning , 2013, pp. 1–9
2013
Earlier work this paper cites.
M. P. Deisenroth, P. Englert, J. Peters, and D. Fox, “Multi-task policy search for robotics,” in 2014 IEEE International Conference on Robotics and Automation (ICRA) , May 2014, pp. 3876–3881
2014
Earlier work this paper cites.
D. Martinez, G. Alenya, P. Jimenez, C. Torras, J. Rossmann, N. Wantia, E. E. Aksoy, S. Haller, and J. Piater, “Active learning of manipulation sequences,” in 2014 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2014, pp. 5671–5678
2014
Earlier work this paper cites.