Fetching the paper…
Reading the bibliography…
In this paper, we study the problem of learning vision-based dynamic manipulation skills using a scalable reinforcement learning approach.
Acceleration of stochastic approximation by averaging
B. T. Polyak and A. B. Juditsky · 1992
Earlier work this paper cites.
TD-Gammon, a Self-Teaching Backgammon Program, Achieves Master-Level Play
G. Tesauro · 1994
Earlier work this paper cites.
Introduction to Reinforcement Learning
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
The Cross-Entropy Method
R. Rubinstein and D. Kroese · 2004
Earlier work this paper cites.
Reinforcement learning of motor skills with policy gradients
J. Peters and S. Schaal · 2008
Earlier work this paper cites.
Double Q-learning
H. V. Hasselt · 2010
Earlier work this paper cites.
Learning Force Control Policies for Compliant Manipulation
M. Kalakrishnan, L. Righetti, P. Pastor, and S. Schaal · 2011
Earlier work this paper cites.
Reinforcement learning in feedback control
R. Hafner and M. Riedmiller · 2011
Earlier work this paper cites.
Pose Error Robust Grasping from Contact Wrench Space Metrics
J. Weisz and P. K. Allen · 2012
Earlier work this paper cites.
Large scale distributed deep networks
J. Dean, G. S. Corrado, R. Monga, K. Chen, M. Devin, Q. V. Le, M. Z. Mao, M. Ranzato, A. Senior, P. Tucker, K. Yang, and A. Y. Ng · 2012
Earlier work this paper cites.
Data-Driven Grasp Synthesis — A Survey
J. Bohg, A. Morales, T. Asfour, and D. Kragic · 2014
Earlier work this paper cites.
Continuous Control with Deep Reinforcement Learning
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2015
Earlier work this paper cites.
Deep Learning for Detecting Robotic Grasps
I. Lenz, H. Lee, and A. Saxena · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
V. Mnih et al · 2015
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
A. X. Chang, T. A. Funkhouser, L. J. Guibas, P. Hanrahan, Q. Huang, Z. Li, S. Savarese, M. Savva, S. Song, H. Su, J. Xiao, L. Yi, and F. Yu · 2015
Earlier work this paper cites.
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift
S. Ioffe and C. Szegedy · 2015
Earlier work this paper cites.
Massively parallel methods for deep reinforcement learning
A. Nair, P. Srinivasan, S. Blackwell, C. Alcicek, R. Fearon, A. D. Maria, V. Panneershelvam, M. Suleyman, C. Beattie, S. Petersen, S. Legg, V. Mnih, K. Kavukcuoglu, and D. Silver · 2015
Cited alongside, same era.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
M. Abadi et. al · 2015
Cited alongside, same era.
Openai gym, 2016
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Cited alongside, same era.
Benchmarking Deep Reinforcement Learning for Continuous Control
Y. Duan, X. Chen, R. Houthooft, J. Schulman, and P. Abbeel · 2016
Cited alongside, same era.
End-to-end Training of Deep Visuomotor Policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Cited alongside, same era.
Learning hand-eye coordination for robotic grasping with large-scale data collection
Learning a visuomotor controller for real world robotic grasping using simulated depth images
U. Viereck, A. ten Pas, K. Saenko, and R. Platt · 2017
Later among the works it cites.
Regrasping using Tactile Perception and Supervised Policy Learning
K. Hausman, Y. Chebotar, O. Kroemer, G. S. Sukhatme, and S. Schaal · 2017
Later among the works it cites.
Input convex neural networks
B. Amos, L. Xu, and J. Z. Kolter · 2017
Later among the works it cites.
Learning Synergies between Pushing and Grasping with Self-supervised Deep Reinforcement Learning
A. Zeng, S. Song, S. Welker, J. Lee, A. Rodriguez, and T. Funkhouser · 2018
Closest in time.
Cartman: The low-cost Cartesian Manipulator that won the Amazon Robotics Challenge
D. Morrison et. al · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Levine, P. Pastor, A. Krizhevsky, and D. Quillen · 2016
Cited alongside, same era.
Supersizing self-supervision: Learning to grasp from 50K tries and 700 robot hours
L. Pinto and A. Gupta · 2016
Cited alongside, same era.
Continuous Deep Q-learning with Model-based Acceleration
S. Gu, T. Lillicrap, I. Sutskever, and S. Levine · 2016
Cited alongside, same era.
Continuous control with deep reinforcement learning
T. Lillicrap, J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2016
Cited alongside, same era.
Deep Reinforcement Learning with Double Q-Learning
H. v. Hasselt, A. Guez, and D. Silver · 2016
Cited alongside, same era.
Collective robot reinforcement learning with distributed asynchronous guided policy search
A. Yahya, A. Li, M. Kalakrishnan, Y. Chebotar, and S. Levine · 2017
Cited alongside, same era.
Deep predictive policy training using reinforcement learning
A. Ghadirzadeh, A. Maki, D. Kragic, and M. Björkman · 2017
Cited alongside, same era.
N. Chavan-Dafle and A. Rodriguez · 2018
Closest in time.
Self-Supervised Deep Reinforcement Learning with Generalized Computation Graphs for Robot Navigation
G. Kahn, A. Villaflor, B. Ding, P. Abbeel, and S. Levine · 2018
Closest in time.
Deep Reinforcement Learning for Vision-Based Robotic Grasping: A Simulated Comparative Evaluation of Off-Policy Methods
D. Quillen, E. Jang, O. Nachum, C. Finn, J. Ibarz, and S. Levine · 2018
Closest in time.
Realtime State Estimation with Tactile and Visual sensing. Application to Planar Manipulation
K. Yu and A. Rodriguez · 2018
Closest in time.
Closing the Loop for Robotic Grasping: A Real-time, Generative Grasp Synthesis Approach
D. Morrison, P. Corke, and J. Leitner · 2018
Closest in time.
Addressing Function Approximation Error in Actor-Critic Methods
S. Fujimoto, H. van Hoof, and D. Meger · 2018
Closest in time.
Pybullet, a python module for physics simulation for games, robotics and machine learning
E. Coumans and Y. Bai · 2018
Closest in time.
Accelerated Methods for Deep Reinforcement Learning
A. Stooke and P. Abbeel · 2018
Closest in time.
IMPALA: scalable distributed deep-rl with importance weighted actor-learner architectures
L. Espeholt, H. Soyer, R. Munos, K. Simonyan, V. Mnih, T. Ward, Y. Doron, V. Firoiu, T. Harley, I. Dunning, S. Legg, and K. Kavukcuoglu · 2018
Closest in time.
Distributed Prioritized Experience Replay
D. Horgan, J. Quan, D. Budden, G. Barth-Maron, M. Hessel, H. van Hasselt, and D. Silver · 2018
Closest in time.