Fetching the paper…
Reading the bibliography…
In classical reinforcement learning, when exploring an environment, agents accept arbitrary short term loss for long term gain.
Bayesian Approach to Global Optimization , volume 37 of Mathematics and Its Applications
Jonas Mockus · 1989
Earlier work this paper cites.
Reinforcement learning: an introduction
Richard S. Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
Risk-sensitive and minimax control of discrete-time, finite-state Markov decision processes
Stefano P. Coraluppi and Steven I. Marcus · 1999
Earlier work this paper cites.
Learning with Kernels: Support Vector Machines, Regularization, Optimization, and Beyond
Bernhard Schölkopf and Alexander J. Smola · 2002
Earlier work this paper cites.
Risk-sensitive reinforcement learning applied to control under constraints
Peter Geibel and Fritz Wysotzki · 2005
Earlier work this paper cites.
Posterior consistency of Gaussian process prior for nonparametric binary regression
Subhashis Ghosal and Anindya Roy · 2006
Earlier work this paper cites.
Introduction: Mars Science Laboratory: The Next Generation of Mars Landers
Mary Kae Lockwood · 2006
Earlier work this paper cites.
Gaussian processes for machine learning
Carl Edward Rasmussen and Christopher K. I. Williams · 2006
Earlier work this paper cites.
Mars Reconnaissance Orbiter’s High Resolution Imaging Science Experiment (HiRISE)
Alfred S. McEwen, Eric M. Eliason, James W. Bergstrom, Nathan T. Bridges, Candice J. Hansen, W. Alan Delamere, John A. Grant, Virginia C. Gulick, Kenneth E. Herkenhoff, Laszlo Keszthelyi, Randolph L. Kirk, Michael T. Mellon, Steven W. Squyres, Nicolas Thomas, and Catherine M. Weitz · 2007
Cited alongside, same era.
MSL Landing Site Selection User’s Guide to Engineering Constraints, 2007
MSL · 2007
Cited alongside, same era.
Safe exploration for reinforcement learning
Alexander Hans, Daniel Schneegaß, Anton Maximilian Schäfer, and Steffen Udluft · 2008
Cited alongside, same era.
A survey of robot learning from demonstration
Brenna D. Argall, Sonia Chernova, Manuela Veloso, and Brett Browning · 2009
Cited alongside, same era.
Learning Control in Robotics
Stefan Schaal and Christopher Atkeson · 2010
Cited alongside, same era.
Gaussian process optimization in the bandit setting: no regret and experimental design
Reinforcement learning in robotics: a survey
Jens Kober, J. Andrew Bagnell, and Jan Peters · 2013
Later among the works it cites.
Reachability-based safe learning with Gaussian processes
Anayo K. Akametalu, Shahab Kaynama, Jaime F. Fisac, Melanie N. Zeilinger, Jeremy H. Gillula, and Claire J. Tomlin · 2014
Later among the works it cites.
Safe exploration techniques for reinforcement learning – an overview
Martin Pecka and Tomas Svoboda · 2014
Later among the works it cites.
Safe and robust learning control with Gaussian processes
Felix Berkenkamp and Angela P. Schoellig · 2015
Later among the works it cites.
Safe exploration for active learning with Gaussian processes
Jens Schreiter, Duy Nguyen-Tuong, Mona Eberts, Bastian Bischoff, Heiner Markert, and Marc Toussaint · 2015
Later among the works it cites.
Safe exploration for optimization with Gaussian processes
Yanan Sui, Alkis Gotovos, Joel Burdick, and Andreas Krause · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Niranjan Srinivas, Andreas Krause, Sham M. Kakade, and Matthias Seeger · 2010
Cited alongside, same era.
Safe exploration of state and action spaces in reinforcement learning
Javier Garcia and Fernando Fernández · 2012
Cited alongside, same era.
Safe exploration in Markov decision processes
Teodor Mihai Moldovan and Pieter Abbeel · 2012
Cited alongside, same era.
Later among the works it cites.
Safe controller optimization for quadrotors with Gaussian processes
Felix Berkenkamp, Angela P. Schoellig, and Andreas Krause · 2016
Closest in time.