Fetching the paper…
Reading the bibliography…
The combination of deep neural network models and reinforcement learning algorithms can make it possible to learn policies for robotic behaviors that directly read in raw sensory inputs, such as camera images, effectively subsuming both estimation and control into one model.
Algorithms for inverse reinforcement learning
Andrew Y. Ng and Stuart J. Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y. Ng · 2004
Earlier work this paper cites.
Maximum margin planning
Nathan D. Ratliff, J. Andrew Bagnell, and Martin Zinkevich · 2006
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
Brian D. Ziebart, Andrew L. Maas, J. Andrew Bagnell, and Anind K. Dey · 2008
Earlier work this paper cites.
Active learning for reward estimation in inverse reinforcement learning
Manuel Lopes, Francisco S. Melo, and Luis Montesano · 2009
Earlier work this paper cites.
Robot motor skill coordination with em-based reinforcement learning
Petar Kormushev, Sylvain Calinon, and Darwin G Caldwell · 2010
Earlier work this paper cites.
Modeling purposeful adaptive behavior with the principle of maximum causal entropy
Brian Ziebart · 2010
Earlier work this paper cites.
Comparing action-query strategies in semi-autonomous agents
Robert Cohn, Edmund H. Durfee, and Satinder P. Singh · 2011
Earlier work this paper cites.
Learning trajectory preferences for manipulators via iterative improvement
Ashesh Jain, Brian Wojcik, Thorsten Joachims, and Ashutosh Saxena · 2013
Earlier work this paper cites.
Active reward learning
Christian Daniel, Malte Viering, Jan Metz, Oliver Kroemer, and Jan Peters · 2014
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian J. Goodfellow, and Rob Fergus · 2014
Earlier work this paper cites.
Inverse reinforcement learning algorithms and features for robot navigation in crowds: An experimental comparison
Dizan Vasquez, Billy Okal, and Kai Oliver Arras · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Trust Region Policy Optimization
John Schulman, Sergey Levine, Philipp Moritz, Michael I. Jordan, and Pieter Abbeel · 2015
Earlier work this paper cites.
Learning robot in-hand manipulation with tactile features
Herke van Hoof, Tucker Hermans, Gerhard Neumann, and Jan Peters · 2015
Cited alongside, same era.
Maximum entropy deep inverse reinforcement learning
Markus Wulfmeier, Peter Ondruska, and Ingmar Posner · 2015
Cited alongside, same era.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
Learning dexterous manipulation policies from experience and imitation
Vikash Kumar, Abhishek Gupta, Emanuel Todorov, and Sergey Levine · 2016
Cited alongside, same era.
Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours
Lerrel Pinto and Abhinav Gupta · 2016
Cited alongside, same era.
Active reward learning from critiques
Yuchen Cui and Scott Niekum · 2018
Later among the works it cites.
Visual foresight: Model-based deep reinforcement learning for vision-based robotic control
Frederik Ebert, Chelsea Finn, Sudeep Dasari, Annie Xie, Alex X. Lee, and Sergey Levine · 2018
Later among the works it cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation
Dmitry Kalashnikov, Alex Irpan, Peter Pastor, Julian Ibarz, Alexander Herzog, Eric Jang, Deirdre Quillen, Ethan Holly, Mrinal Kalakrishnan, Vincent Vanhoucke, and Sergey Levine · 2018
Later among the works it cites.
Discriminator-actor-critic: Addressing sample inefficiency and reward bias in adversarial imitation learning
Ilya Kostrikov, Kumar Krishna Agrawal, Debidatta Dwibedi, Sergey Levine, and Jonathan Tompson · 2018
Later among the works it cites.
Reinforcement learning and control as probabilistic inference: Tutorial and review
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yevgen Chebotar, Mrinal Kalakrishnan, Ali Yahya, Adrian Li, Stefan Schaal, and Sergey Levine · 2017
Cited alongside, same era.
Deep reinforcement learning from human preferences
Paul F. Christiano, Jan Leike, Tom B. Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Cited alongside, same era.
Perceptual Goal Specifications for Reinforcement Learning
Ashley D Edwards · 2017
Cited alongside, same era.
Reinforcement learning with deep energy-based policies
Tuomas Haarnoja, Haoran Tang, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Risk-sensitive inverse reinforcement learning via coherent risk models
Anirudha Majumdar, Sumeet Singh, Ajay Mandlekar, and Marco Pavone · 2017
Cited alongside, same era.
Sim-to-real robot learning from pixels with progressive nets
Andrei A. Rusu, Matej Vecerik, Thomas Rothörl, Nicolas Heess, Razvan Pascanu, and Raia Hadsell · 2017
Cited alongside, same era.
Visual closed-loop control for pouring liquids
Connor Schenck and Dieter Fox · 2017
Cited alongside, same era.
Sergey Levine · 2018
Later among the works it cites.
Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection
Sergey Levine, Peter Pastor, Alex Krizhevsky, Julian Ibarz, and Deirdre Quillen · 2018
Later among the works it cites.
Sim-to-real reinforcement learning for deformable object manipulation
Jan Matas, Stephen James, and Andrew J. Davison · 2018
Later among the works it cites.
Learning dexterous in-hand manipulation
OpenAI · 2018
Later among the works it cites.
Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
Aravind Rajeswaran, Vikash Kumar, Abhishek Gupta, Giulia Vezzani, John Schulman, Emanuel Todorov, and Sergey Levine · 2018
Later among the works it cites.
A practical approach to insertion with variable socket position using deep reinforcement learning
Mel Vecerik, Oleg Sushkov, David Barker, Thomas Rothörl, Todd Hester, and Jon Scholz · 2018
Later among the works it cites.
Few-shot goal inference for visuomotor learning and planning
Annie Xie, Avi Singh, Sergey Levine, and Chelsea Finn · 2018
Later among the works it cites.
Learning a prior over intent via meta-inverse reinforcement learning
Kelvin Xu, Ellis Ratner, Anca Dragan, Sergey Levine, and Chelsea Finn · 2018
Later among the works it cites.
mixup: Beyond empirical risk minimization
Hongyi Zhang, Moustapha Cissé, Yann N. Dauphin, and David Lopez-Paz · 2018
Later among the works it cites.