Fetching the paper…
Reading the bibliography…
Deep reinforcement learning provides a promising approach for vision-based control of real-world robots.
C. J. Watkins and P. Dayan, “Q-learning,” Machine learning , 1992
1992
Earlier work this paper cites.
L.-Y. Deng, “The cross-entropy method: a unified approach to combinatorial optimization, monte-carlo simulation, and machine learning,” 2006
2006
Earlier work this paper cites.
M. E. Taylor and P. Stone, “Transfer learning for reinforcement learning domains: A survey,” JMLR , 2009
2009
Earlier work this paper cites.
S. J. Pan, Q. Yang, et al. , “A survey on transfer learning,” IEEE Transactions on knowledge and data engineering , 2010
2010
Earlier work this paper cites.
S. Ross, N. Melik-Barkhudarov, K. S. Shankar, A. Wendel, D. Dey, J. A. Bagnell, and M. Hebert, “Learning monocular reactive uav control in cluttered natural environments,” in ICRA , 2013
2013
Earlier work this paper cites.
J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement learning in robotics: A survey,” IJRR , 2013
2013
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in ICLR , 2014
2014
Earlier work this paper cites.
J. Donahue, Y. Jia, O. Vinyals, J. Hoffman, N. Zhang, E. Tzeng, and T. Darrell, “Decaf: A deep convolutional activation feature for generic visual recognition,” in ICML , 2014
2014
Earlier work this paper cites.
A. Sharif Razavian, H. Azizpour, J. Sullivan, and S. Carlsson, “Cnn features off-the-shelf: an astounding baseline for recognition,” in CVPR , 2014
2014
Earlier work this paper cites.
J. Yosinski, J. Clune, Y. Bengio, and H. Lipson, “How transferable are features in deep neural networks?” in NIPS , 2014
2014
Earlier work this paper cites.
E. Tzeng, J. Hoffman, T. Darrell, and K. Saenko, “Simultaneous deep transfer across domains and tasks,” in ICCV , 2015
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al. , “Imagenet large scale visual recognition challenge,” in IJCV , 2015
2015
Earlier work this paper cites.
I. Mordatch, K. Lowrey, and E. Todorov, “Ensemble-cio: Full-body dynamic motion planning that transfers to physical humanoids,” in IROS , 2015
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al. , “Human-level control through deep reinforcement learning,” in Nature , 2015
2015
Earlier work this paper cites.
J. Oh, X. Guo, H. Lee, R. L. Lewis, and S. Singh, “Action-conditional video prediction using deep networks in atari games,” in NIPS , 2015
2015
Earlier work this paper cites.
A. Giusti, J. Guzzi, D. C. Ciresan, F.-L. He, J. P. Rodríguez, F. Fontana, M. Faessler, C. Forster, J. Schmidhuber, G. Di Caro, et al. , “A machine learning approach to visual perception of forest trails for mobile robots.” in RAL , 2016
2016
Cited alongside, same era.
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, and V. Lempitsky, “Domain-adversarial training of neural networks,” JMLR , 2016
2016
Cited alongside, same era.
S. Daftry, J. A. Bagnell, and M. Hebert, “Learning transferable policies for monocular reactive mav control,” in ISER , 2016
2016
Cited alongside, same era.
J. Fu, S. Levine, and P. Abbeel, “One-shot learning of manipulation skills with online dynamics adaptation and neural network priors,” in IROS , 2016
2016
Cited alongside, same era.
A. Bitcraze, “Crazyflie 2.0,” 2016
2016
D. Gandhi, L. Pinto, and A. Gupta, “Learning to fly by crashing,” in IROS , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Y. Li, “Deep reinforcement learning: An overview,” arXiv preprint arXiv:1701.07274 , 2017
2017
Cited alongside, same era.
F. Sadeghi and S. Levine, “Cad2rl: Real single-image flight without a single real image,” in RSS , 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
F. Zhang, J. Leitner, M. Milford, and P. Corke, “Modular deep q networks for sim-to-real transfer of visuo-motor policies,” in ACRA , 2017
2017
Cited alongside, same era.
A. Ghadirzadeh, A. Maki, D. Kragic, and M. Björkman, “Deep predictive policy training using reinforcement learning,” in IROS , 2017
2017
Cited alongside, same era.
J. P. Hanna and P. Stone, “Grounded action transformation for robot learning in simulation.” in AAAI , 2017
2017
Cited alongside, same era.
A. Rajeswaran, S. Ghotra, B. Ravindran, and S. Levine, “Epopt: Learning robust neural network policies using model ensembles,” in ICLR , 2017
2017
Cited alongside, same era.
2018
Later among the works it cites.
K. Lowrey, S. Kolev, J. Dao, A. Rajeswaran, and E. Todorov, “Reinforcement learning for non-prehensile manipulation: Transfer from simulation to physical system,” in SIMPAR , 2018
2018
Later among the works it cites.
F. Xia, A. R. Zamir, Z. He, A. Sax, J. Malik, and S. Savarese, “Gibson env: Real-world perception for embodied agents,” in CVPR , 2018
2018
Later among the works it cites.
L. Pinto, M. Andrychowicz, P. Welinder, W. Zaremba, and P. Abbeel, “Asymmetric actor critic for image-based robot learning,” in RSS , 2018
2018
Later among the works it cites.
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” in ICRA , 2018
2018
Later among the works it cites.
D. Rastogi, I. Koryakovskiy, and J. Kober, “Sample-efficient reinforcement learning via difference models,” in Machine Learning in Planning and Control of Robot Motion Workshop at ICRA , 2018
2018
Later among the works it cites.
K. Mohta, K. Sun, S. Liu, M. Watterson, B. Pfrommer, J. Svacha, Y. Mulgaonkar, C. J. Taylor, and V. Kumar, “Experiments in fast, autonomous, gps-denied quadrotor flight,” in ICRA , 2018
2018
Later among the works it cites.
A. J. Barry, P. R. Florence, and R. Tedrake, “High-speed autonomous obstacle avoidance with pushbroom stereo,” in Journal of Field Robotics , 2018
2018
Later among the works it cites.
A. Loquercio, A. I. Maqueda, C. R. del Blanco, and D. Scaramuzza, “Dronet: Learning to fly by driving,” in RA-L , 2018
2018
Later among the works it cites.
G. Kahn, A. Villaflor, B. Ding, P. Abbeel, and S. Levine, “Self-supervised deep reinforcement learning with generalized computation graphs for robot navigation,” in ICRA , 2018
2018
Later among the works it cites.
D. Quillen, E. Jang, O. Nachum, C. Finn, J. Ibarz, and S. Levine, “Deep reinforcement learning for vision-based robotic grasping: A simulated comparative evaluation of off-policy methods,” in ICRA , 2018
2018
Later among the works it cites.