Fetching the paper…
Reading the bibliography…
Deep neural network based reinforcement learning (RL) can learn appropriate visual representations for complex tasks like vision-based robotic grasping without the need for manually engineering or prior learning a perception system.
Q-learning
Christopher JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros · 2006
Earlier work this paper cites.
Domain adaptation for object recognition: An unsupervised approach
Raghuraman Gopalan, Ruonan Li, and Rama Chellappa · 2011
Earlier work this paper cites.
Data-driven grasp synthesis—a survey
Jeannette Bohg, Antonio Morales, Tamim Asfour, and Danica Kragic · 2013
Earlier work this paper cites.
Deep learning for detecting robotic grasps
Ian Lenz, Honglak Lee, and Ashutosh Saxena · 2013
Earlier work this paper cites.
Generative adversarial networks
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron C. Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Beyond the shortest path: Unsupervised domain adaptation by sampling subspaces along the spline flow
Rui Caseiro, Joao F Henriques, Pedro Martins, and Jorge Batista · 2015
Earlier work this paper cites.
Bullet physics simulation
Erwin Coumans · 2015
Earlier work this paper cites.
Learning transferable features with deep adaptation networks
Mingsheng Long, Yue Cao, Jianmin Wang, and Michael I Jordan · 2015
Earlier work this paper cites.
Visual domain adaptation: A survey of recent advances
Vishal M. Patel, Raghuraman Gopalan, Ruonan Li, and Rama Chellappa · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Unsupervised pixel-level domain adaptation with generative adversarial networks
Konstantinos Bousmalis, Nathan Silberman, David Dohan, Dumitru Erhan, and Dilip Krishnan · 2016
Earlier work this paper cites.
Domain-adversarial training of neural networks
Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor Lempitsky · 2016
Cited alongside, same era.
Cad2rl: Real single-image flight without a single real image
Fereshteh Sadeghi and Sergey Levine · 2016
Cited alongside, same era.
Instance normalization: The missing ingredient for fast stylization
Dmitry Ulyanov, Andrea Vedaldi, and Victor S. Lempitsky · 2016
Cited alongside, same era.
Donggeun Yoo, Namil Kim, Sunggyun Park, Anthony S. Paek, and In-So Kweon · 2016
Cited alongside, same era.
Using simulation and domain adaptation to improve efficiency of deep robotic grasping
Konstantinos Bousmalis, Alex Irpan, Paul Wohlhart, Yunfei Bai, Matthew Kelcey, Mrinal Kalakrishnan, Laura Downs, Julian Ibarz, Peter Pastor, Kurt Konolige, Sergey Levine, and Vincent Vanhoucke · 2017
Augmented cyclegan: Learning many-to-many mappings from unpaired data
Amjad Almahairi, Sai Rajeswar, Alessandro Sordoni, Philip Bachman, and Aaron Courville · 2018
Later among the works it cites.
Large scale gan training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, and Karen Simonyan · 2018
Later among the works it cites.
Stable prehensile pushing: In-hand manipulation with alternating sticking contacts
Nikhil Chavan-Dafle and Alberto Rodriguez · 2018
Later among the works it cites.
Addressing function approximation error in actor-critic methods
Scott Fujimoto, Herke van Hoof, and David Meger · 2018
Later among the works it cites.
Task-embedded control networks for few-shot imitation learning
Stephen James, Michael Bloesch, and Andrew J. Davison · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Cyclegan, a master of steganography
Casey Chu, Andrey Zhmoginov, and Mark Sandler · 2017
Cited alongside, same era.
Cycada: Cycle-consistent adversarial domain adaptation
Judy Hoffman, Eric Tzeng, Taesung Park, Jun-Yan Zhu, Phillip Isola, Kate Saenko, Alexei Efros, and Trevor Darrell · 2017
Cited alongside, same era.
Unsupervised cross-domain image generation
Xinru Hua, Davis Rempe, and Haotian Zhang · 2017
Cited alongside, same era.
Transferring end-to-end visuomotor control from simulation to real world for a multi-stage task
Stephen James, Andrew J. Davison, and Edward Johns · 2017
Cited alongside, same era.
Dex-net 2.0: Deep learning to plan robust grasps with synthetic point clouds and analytic grasp metrics
Jeffrey Mahler, Jacky Liang, Sherdil Niyaz, Michael Laskey, Richard Doan, Xinyu Liu, Juan Aparicio, and Ken Goldberg · 2017
Cited alongside, same era.
Sim2real view invariant visual servoing by recurrent control
Fereshteh Sadeghi, Alexander Toshev, Eric Jang, and Sergey Levine · 2017
Cited alongside, same era.
Domain randomization for transferring deep neural networks from simulation to the real world
Joshua Tobin, Rachel H Fong, Alex Ray, Jonas Schneider, Wojciech Zaremba, and Pieter Abbeel · 2017
Cited alongside, same era.
Sim-to-real via sim-to-sim: Data-efficient robotic grasping via randomized-to-canonical adaptation networks
Stephen James, Paul Wohlhart, Mrinal Kalakrishnan, Dmitry Kalashnikov, Alex Irpan, Julian Ibarz, Sergey Levine, Raia Hadsell, and Konstantinos Bousmalis · 2018
Later among the works it cites.
Scalable deep reinforcement learning for vision-based robotic manipulation
Dmitry Kalashnikov, Alex Irpan, Peter Pastor, Julian Ibarz, Alexander Herzog, Eric Jang, Deirdre Quillen, Ethan Holly, Mrinal Kalakrishnan, Vincent Vanhoucke, and Sergey Levine · 2018
Later among the works it cites.
Sim-to-real reinforcement learning for deformable object manipulation
Jan Matas, Stephen James, and Andrew J. Davison · 2018
Later among the works it cites.
Learning by playing - solving sparse reward tasks from scratch
Martin A. Riedmiller, Roland Hafner, Thomas Lampe, Michael Neunert, Jonas Degrave, Tom Van de Wiele, Volodymyr Mnih, Nicolas Heess, and Jost Tobias Springenberg · 2018
Later among the works it cites.
Learning synergies between pushing and grasping with self-supervised deep reinforcement learning
Andy Zeng, Shuran Song, Stefan Welker, Johnny Lee, Alberto Rodriguez, and Thomas Funkhouser · 2018
Later among the works it cites.
Self-attention generative adversarial networks
Han Zhang, Ian J. Goodfellow, Dimitris N. Metaxas, and Augustus Odena · 2018
Later among the works it cites.
Long-range indoor navigation with PRM-RL
Anthony Francis, Aleksandra Faust, Hao-Tien Lewis Chiang, Jasmine Hsu, J. Chase Kew, Marek Fiser, and Tsang-Wei Edward Lee · 2019
Later among the works it cites.