Fetching the paper…
Reading the bibliography…
Learning from human feedback is a viable alternative to control design that does not require modelling or control expertise.
On the generalized distance in statistics
Mahalanobis, P. C. (1936) · 1936
Earlier work this paper cites.
Handbook of Mathematical Functions: With Formulas, Graphs, and Mathematical Tables
Abramowitz, M. and Stegun, I. A. (1965) · 1965
Earlier work this paper cites.
Bayesian learning for neural networks
Neal, R. M. (1995) · 1995
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P. and Ng, A. Y. (2004) · 2004
Earlier work this paper cites.
Gaussian Processes for Machine Learning
Rasmussen, C. E. and Williams, C. K. (2006) · 2006
Earlier work this paper cites.
Reinforcement learning with human teachers: Evidence of feedback and guidance with implications for learning performance
Thomaz, A. L. and Breazeal, C. (2006) · 2006
Earlier work this paper cites.
A survey of robot learning from demonstration
Argall, B. D., Chernova, S., Veloso, M., and Browning, B. (2009) · 2009
Earlier work this paper cites.
Interactive policy learning through confidence-based autonomy
Chernova, S. and Veloso, M. (2009) · 2009
Earlier work this paper cites.
Interactively shaping agents via human reinforcement: The TAMER framework
Knox, W. B. and Stone, P. (2009) · 2009
Cited alongside, same era.
Reinforcement Learning and Dynamic Programming using Function Approximators
Busoniu, L., Babuska, R., De Schutter, B., and Ernst, D. (2010) · 2010
Cited alongside, same era.
Incremental sparsification for real-time online model learning
Nguyen-Tuong, D. and Peters, J. (2010) · 2010
Cited alongside, same era.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G., and Bagnell, D. (2011) · 2011
Cited alongside, same era.
Effect of human guidance and state space size on interactive reinforcement learning
Suay, H. B. and Chernova, S. (2011) · 2011
Cited alongside, same era.
Learning sequential tasks interactively from demonstrations and own experience
Automatic Model Construction with Gaussian Processes
Duvenaud, D. (2014) · 2014
Later among the works it cites.
COACH: Learning continuous actions from corrective advice communicated by humans
Celemin, C. and Ruiz-del Solar, J. (2015) · 2015
Later among the works it cites.
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W. (2016) · 2016
Later among the works it cites.
Active incremental learning of robot movement primitives
Maeda, G., Ewerton, M., Osa, T., Busch, B., and Peters, J. (2017) · 2017
Later among the works it cites.
A fast hybrid reinforcement learning framework with human corrective feedback
Celemin, C., Ruiz-del Solar, J., and Kober, J. (2018) · 2018
Later among the works it cites.
Rainbow: Combining improvements in deep reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gräve, K. and Behnke, S. (2013) · 2013
Cited alongside, same era.
Policy shaping: Integrating human feedback with reinforcement learning
Griffith, S., Subramanian, K., Scholz, J., Isbell, C. L., and Thomaz, A. L. (2013) · 2013
Cited alongside, same era.
Robot learning from human teachers
Chernova, S. and Thomaz, A. L. (2014) · 2014
Cited alongside, same era.
Hessel, M., Modayil, J., Van Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M., and Silver, D. (2018) · 2018
Later among the works it cites.
Including uncertainty when learning from human corrections
Losey, D. P. and O’Malley, M. K. (2018) · 2018
Later among the works it cites.