Fetching the paper…
Reading the bibliography…
In this paper, we explore deep reinforcement learning algorithms for vision-based robotic grasping.
Q-learning
C. J. Watkins and P. Dayan · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Neural fitted q iteration-first experiences with a data efficient neural reinforcement learning method
M. Riedmiller · 2005
Earlier work this paper cites.
The columbia grasp database
C. Goldfeder, M. Ciocarlie, H. Dang, and P. K. Allen · 2009
Earlier work this paper cites.
Autonomous reinforcement learning on raw visual input data in a real world application
S. Lange, M. Riedmiller, and A. Voigtlander · 2012
Earlier work this paper cites.
From caging to grasping
A. Rodriguez, M. T. Mason, and S. Ferry · 2012
Earlier work this paper cites.
Pose error robust grasping from contact wrench space metrics
J. Weisz and P. K. Allen · 2012
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
Data-driven grasp synthesis—a survey
J. Bohg, A. Morales, T. Asfour, and D. Kragic · 2014
Earlier work this paper cites.
Learning of grasp selection based on shape-templates
A. Herzog, P. Pastor, M. Kalakrishnan, L. Righetti, J. Bohg, T. Asfour, and S. Schaal · 2014
Earlier work this paper cites.
Leveraging big data for grasp planning
D. Kappler, J. Bohg, and S. Schaal · 2015
Earlier work this paper cites.
Deep learning for detecting robotic grasps
I. Lenz, H. Lee, and A. Saxena · 2015
Earlier work this paper cites.
Trust region policy optimization
J. Schulman, S. Levine, P. Abbeel, M. I. Jordan, and P. Moritz · 2015
Earlier work this paper cites.
Towards vision-based deep reinforcement learning for robotic motion control
F. Zhang, J. Leitner, M. Milford, B. Upcroft, and P. Corke · 2015
Earlier work this paper cites.
Openai gym
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Cited alongside, same era.
Benchmarking deep reinforcement learning for continuous control
Y. Duan, X. Chen, R. Houthooft, J. Schulman, and P. Abbeel · 2016
Cited alongside, same era.
Unsupervised learning for physical interaction through video prediction
C. Finn, I. Goodfellow, and S. Levine · 2016
Cited alongside, same era.
Deep spatial autoencoders for visuomotor learning
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel · 2016
Cited alongside, same era.
Continuous deep Q-learning with model-based acceleration
S. Gu, T. Lillicrap, I. Sutskever, and S. Levine · 2016
Cited alongside, same era.
3d simulation for robot arm control with deep q-learning
S. James and E. Johns · 2016
Cited alongside, same era.
Path integral guided policy search
Y. Chebotar, M. Kalakrishnan, A. Yahya, A. Li, S. Schaal, and S. Levine · 2017
Later among the works it cites.
pybullet, a python module for physics simulation, games, robotics and machine learning
E. Coumans and Y. Bai · 2017
Later among the works it cites.
Deep visual foresight for planning robot motion
C. Finn and S. Levine · 2017
Later among the works it cites.
Deep predictive policy training using reinforcement learning
A. Ghadirzadeh, A. Maki, D. Kragic, and M. Björkman · 2017
Later among the works it cites.
Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates
S. Gu, E. Holly, T. Lillicrap, and S. Levine · 2017
Later among the works it cites.
Q-prop: Sample-efficient policy gradient with an off-policy critic
S. Gu, T. Lillicrap, Z. Ghahramani, R. E. Turner, and S. Levine · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Cited alongside, same era.
Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen · 2016
Cited alongside, same era.
Continuous control with deep reinforcement learning
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. P. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Cited alongside, same era.
PGQ: Combining policy gradient and Q-learning
B. O’Donoghue, R. Munos, K. Kavukcuoglu, and V. Mnih · 2016
Cited alongside, same era.
Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours
L. Pinto and A. Gupta · 2016
Cited alongside, same era.
Later among the works it cites.
Reinforcement learning with deep energy-based policies
T. Haarnoja, H. Tang, P. Abbeel, and S. Levine · 2017
Later among the works it cites.
Reproducibility of benchmarked deep reinforcement learning tasks for continuous control
R. Islam, P. Henderson, M. Gomrokchi, and D. Precup · 2017
Later among the works it cites.
Dex-net 2.0: Deep learning to plan robust grasps with synthetic point clouds and analytic grasp metrics
J. Mahler, J. Liang, S. Niyaz, M. Laskey, R. Doan, X. Liu, J. A. Ojea, and K. Goldberg · 2017
Later among the works it cites.
Bridging the gap between value and policy based reinforcement learning
O. Nachum, M. Norouzi, K. Xu, and D. Schuurmans · 2017
Later among the works it cites.
Trust-pcl: An off-policy trust region method for continuous control
O. Nachum, M. Norouzi, K. Xu, and D. Schuurmans · 2017
Later among the works it cites.
Grasp pose detection in point clouds
A. t. Pas, M. Gualtieri, K. Saenko, and R. Platt · 2017
Later among the works it cites.
Data-efficient deep reinforcement learning for dexterous manipulation
I. Popov, N. Heess, T. Lillicrap, R. Hafner, G. Barth-Maron, M. Vecerik, T. Lampe, Y. Tassa, T. Erez, and M. Riedmiller · 2017
Later among the works it cites.
Domain randomization for transferring deep neural networks from simulation to the real world
J. Tobin, R. Fong, A. Ray, J. Schneider, W. Zaremba, and P. Abbeel · 2017
Later among the works it cites.
Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation
Y. Wu, E. Mansimov, S. Liao, R. Grosse, and J. Ba · 2017
Later among the works it cites.