Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (RL) algorithms can learn complex robotic skills from raw sensory inputs, but have yet to achieve the kind of broad generalization and applicability demonstrated by deep learning methods in supervised domains.
D. Sherer, “Fetal grasping at 16 weeks’ gestation,” Journal of ultrasound in medicine , 1993
1993
Earlier work this paper cites.
G. Tesauro, “Temporal difference learning and td-gammon,” Communications of the ACM , vol. 38, no. 3, pp. 58–68, 1995
1995
Earlier work this paper cites.
J. T. Betts, “Survey of numerical methods for trajectory optimization,” Journal of guidance, control, and dynamics , 1998
1998
Earlier work this paper cites.
B. Babenko, M.-H. Yang, and S. Belongie, “Visual tracking with online multiple instance learning,” in Computer Vision and Pattern Recognition (CVPR) . IEEE, 2009
2009
Earlier work this paper cites.
A. Bubic, D. Y. Von Cramon, and R. Schubotz, “Prediction, cognition and the brain,” Frontiers in Human Neuroscience , 2010
2010
Earlier work this paper cites.
M. Deisenroth and C. E. Rasmussen, “Pilco: A model-based and data-efficient approach to policy search,” in International Conference on Machine Learning (ICML) , 2011
2011
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in International Conference on Intelligent Robots and Systems (IROS) , 2012
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
M. P. Deisenroth, G. Neumann, J. Peters et al. , “A survey on policy search for robotics,” Foundations and Trends® in Robotics , 2013
2013
Earlier work this paper cites.
B. Boots, A. Byravan, and D. Fox, “Learning predictive models of a depth camera & manipulator from raw execution traces,” in International Conference on Robotics and Automation (ICRA) , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
I. Lenz and A. Saxena, “Deepmpc: Learning deep latent features for model predictive control,” in In RSS . Citeseer, 2015
2015
Earlier work this paper cites.
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller, “Embed to control: A locally linear latent dynamics model for control from raw images,” in Neural Information Processing Systems , 2015
2015
Earlier work this paper cites.
J. Oh, X. Guo, H. Lee, R. L. Lewis, and S. Singh, “Action-conditional video prediction using deep networks in atari games,” in Neural Information Processing Systems , 2015
2015
Earlier work this paper cites.
S. Xingjian, Z. Chen, H. Wang, D.-Y. Yeung, W.-K. Wong, and W.-c. Woo, “Convolutional lstm network: A machine learning approach for precipitation nowcasting,” in Neural Information Processing Systems , 2015
2015
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end learning of deep visuomotor policies,” Journal of Machine Learning Research (JMLR) , 2016
2016
Earlier work this paper cites.
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel, “Deep spatial autoencoders for visuomotor learning,” in International Conference on Robotics and Automation (ICRA) , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
P. Agrawal, A. V. Nair, P. Abbeel, J. Malik, and S. Levine, “Learning to poke by poking: Experiential learning of intuitive physics,” in Neural Information Processing Systems , 2016
2016
Cited alongside, same era.
L. Pinto and A. Gupta, “Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours,” in International Conference on Robotics and Automation (ICRA) , 2016
2016
Cited alongside, same era.
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen, “Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,” International Journal of Robotics Research (IJRR) , 2016
2016
Cited alongside, same era.
L. Pinto, D. Gandhi, Y. Han, Y.-L. Park, and A. Gupta, “The curious robot: Learning visual representations via physical interactions,” in European Conference on Computer Vision , 2016
2016
Cited alongside, same era.
2017
Later among the works it cites.
D. Gandhi, L. Pinto, and A. Gupta, “Learning to fly by crashing,” arXiv:1704.05588 , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Finn, I. Goodfellow, and S. Levine, “Unsupervised learning for physical interaction through video prediction,” in Neural Information Processing Systems (NIPS) , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
M. Mathieu, C. Couprie, and Y. LeCun, “Deep multi-scale video prediction beyond mean square error,” International Conference on Learning Representations (ICLR) , 2016
2016
Cited alongside, same era.
C. Vondrick, H. Pirsiavash, and A. Torralba, “Generating videos with scene dynamics,” in Neural Information Processing Systems , 2016
2016
Cited alongside, same era.
B. De Brabandere, X. Jia, T. Tuytelaars, and L. Van Gool, “Dynamic filter networks,” in Neural Information Processing Systems (NIPS) , 2016
2016
Cited alongside, same era.
T. Zhou, S. Tulsiani, W. Sun, J. Malik, and A. A. Efros, “View synthesis by appearance flow,” in European conference on computer vision , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
A. Odena, V. Dumoulin, and C. Olah, “Deconvolution and checkerboard artifacts,” Distill , 2016
2016
Cited alongside, same era.
W. Lotter, G. Kreiman, and D. Cox, “Deep predictive coding networks for video prediction and unsupervised learning,” International Conference on Learning Representations (ICLR) , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” International Conference on Machine Learning (ICML) , 2017
2017
Later among the works it cites.
E. Grant, C. Finn, J. Peterson, J. Abbott, S. Levine, T. Griffiths, and T. Darrell, “Concept acquisition via meta-learning: Few-shot learning from positive examples,” in NIPS Workshop on Cognitively-Informed Artificial Intelligence , 2017
2017
Later among the works it cites.
S. Niklaus, L. Mai, and F. Liu, “Video frame interpolation via adaptive separable convolution,” in International Conference on Computer Vision (ICCV) , 2017
2017
Later among the works it cites.
F. Ebert, S. Dasari, A. X. Lee, S. Levine, and C. Finn, “Robustness via retrying: Closed-loop robotic manipulation with self-supervised learning,” Conference on Robot Learning (CoRL) , 2018
2018
Closest in time.
A. Xie, A. Singh, S. Levine, and C. Finn, “Few-shot goal inference for visuomotor learning and planning,” Conference on Robot Learning (CoRL) , 2018
2018
Closest in time.
D. Ha and J. Schmidhuber, “World models,” arXiv:1803.10122 , 2018
2018
Closest in time.
2018
Closest in time.
2018
Closest in time.
M. Babaeizadeh, C. Finn, D. Erhan, R. H. Campbell, and S. Levine, “Stochastic variational video prediction,” International Conference on Learning Representations (ICLR) , 2018
2018
Closest in time.
2018
Closest in time.