Fetching the paper…
Reading the bibliography…
Reinforcement learning can enable complex, adaptive behavior to be learned automatically for autonomous robotic platforms.
The Jackknife, the Bootstrap and Other Resampling Plans
B. Efron and R. Tibshirani · 1982
Earlier work this paper cites.
An introduction to the bootstrap
B. Efron and R. Tibshirani · 1994
Earlier work this paper cites.
Exploiting Model Uncertainty Estimates for Safe Dynamic Control Learning
J. Schneider · 1997
Earlier work this paper cites.
Lyapunov Design for Safe Reinforcement Learning
T. Perkins and A. Barto · 2002
Earlier work this paper cites.
Policy Gradient Methods for Robotics
J. Peters and S. Schaal · 2006
Earlier work this paper cites.
Autonomous Driving in Urban Environments: Boss and the Urban Challenge
C. Urmson and et. al · 2008
Earlier work this paper cites.
Viability and Predictive Control for Safe Locomotion
P. Wieber · 2008
Earlier work this paper cites.
Rectified Linear Units Improve Restricted Boltzmann Machines
V. Nair and G. Hinton · 2010
Earlier work this paper cites.
PILCO: A Model-based and Data-Efficient Approach to Policy Search
M. Deisenroth and C. Rasmussen · 2011
Earlier work this paper cites.
The Big Data Bootstrap
A. Kleiner, A. Talwalkar, P. Sarkar, and M. I. Jordan · 2012
Earlier work this paper cites.
Safe Exploration in Markov Decision Processes
T. Moldovan and P. Abbeel · 2012
Cited alongside, same era.
A Survey on Policy Search for Robotics
M. P. Deisenroth, G. Neumann, and J. Peters · 2013
Cited alongside, same era.
In Reinforcement learning in robotics: A survey , volume 32, pages 1238–1274, 2013
J. Kober, J. A. Bagnell, and J. Peters · 2013
Cited alongside, same era.
Playing Atari with Deep Reinforcement Learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and Riedmiller M · 2013
Cited alongside, same era.
Dropout: A Simple Way to Prevent Neural Networks from Overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutidnov · 2014
Cited alongside, same era.
Adam: A Method for Stochastic Optimization
D.P. Kingma and J. Ba · 2015
Safe receding horizon control for aggressive MAV flight with limited range sensing
M. Watterson and V. Kumar · 2015
Later among the works it cites.
Bayesian Optimization with Safety Constraints: Safe and Automatic Parameter Tuning in Robotics
F. Berkenkamp, A. Krause, and A. Schoellig · 2016
Later among the works it cites.
Introspective Perception: Learning to Predict Failures in Vision Systems
S. Daftry, S. Zeng, J. A. Bagnell, and M. Hebert · 2016
Later among the works it cites.
Dropout as a Bayesian Approximation: Representing Model Uncertainty in Deep Learning
Y. Gal and Z. Ghahramani · 2016
Later among the works it cites.
Improving PILCO with Bayesian Neural Network Dynamics Models
Y. Gal, R. Mcallister, and C. Rasmussen · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Continuous control with deep reinforcement learning
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2015
Cited alongside, same era.
Relaxed hover solutions for multicopters: Application to algorithmic redundancy and novel vehicles
M. Mueller and R. D’Andrea · 2015
Cited alongside, same era.
Bayesian Learning for Safe High-Speed Navigation in Unknown Environments
C. Richter, W. Vega-Brown, and N. Roy · 2015
Cited alongside, same era.
Trust Region Policy Optimization
J. Schulman, S. Levine, P. Moritz, M. I. Jordan, and P. Abbeel · 2015
Cited alongside, same era.
Guaranteed Safe Online Learning via Reachability: Tracking a Ground Target using a Quadrotor
J. Gillula and C. Tomlin
Cited in the paper.
Reducing Conservativeness in Safety Guarantees by Learning Disturbances Online: Iterated Guaranteed Safe Online Learning
J. Gillula and C. Tomlin
Cited in the paper.
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Later among the works it cites.
Funnel Libraries for Real-Time Robust Feedback Motion Planning
A. Majumdar and R. Tedrake · 2016
Later among the works it cites.
Deep Exploration via Bootstrapped DQN
I. Osband, C. Blundell, A. Pritzel, and B. Van Roy · 2016
Later among the works it cites.
PLATO: Policy Learning using Adaptive Trajectory Optimization
G. Kahn, C. Zhang, S. Levine, and P. Abbeel · 2017
Closest in time.