What is intrinsic motivation? a typology of computational approaches
Oudeyer, P.-Y. and Kaplan, F · 2009
Earlier work this paper cites.
Model predictive control
Camacho, E. F. and Alba, C. B · 2013
Earlier work this paper cites.
Safe and efficient off-policy reinforcement learning
Original
Munos, R., Stepleton, T., Harutyunyan, A., and Bellemare, M. G · 2016
Earlier work this paper cites.
Hindsight experience replay
Original
Andrychowicz, M., Wolski, F., Ray, A., Schneider, J., Fong, R., Welinder, P., McGrew, B., Tobin, J., Abbeel, P., and Zaremba, W · 2017
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction
Pathak, D., Agrawal, P., Efros, A. A., and Darrell, T · 2017
Earlier work this paper cites.
Information theoretic mpc for model-based reinforcement learning
Williams, G., Wagener, N., Goldfain, B., Drews, P., Rehg, J. M., Boots, B., and Theodorou, E. A · 2017
Earlier work this paper cites.
Maximum a posteriori policy optimisation
Original
Abdolmaleki, A., Springenberg, J. T., Tassa, Y., Munos, R., Heess, N., and Riedmiller, M · 2018
Earlier work this paper cites.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
Chua, K., Calandra, R., McAllister, R., and Levine, S · 2018
Earlier work this paper cites.
Diversity is all you need: Learning skills without a reward function
Original
Eysenbach, B., Gupta, A., Ibarz, J., and Levine, S · 2018
Earlier work this paper cites.
Plan online, learn offline: Efficient learning and exploration via model-based control
Original
Lowrey, K., Rajeswaran, A., Kakade, S., Todorov, E., and Mordatch, I · 2018
Earlier work this paper cites.
Deepmind control suite
Original
Tassa, Y., Doron, Y., Muldal, A., Erez, T., Li, Y., Casas, D. d. L., Budden, D., Abdolmaleki, A., Merel, J., Lefrancq, A., et al · 2018
Earlier work this paper cites.
A framework for data-driven robotics
Original
Cabi, S., Colmenarejo, S. G., Novikov, A., Konyushkova, K., Reed, S. E., Jeong, R., Zolna, K., Aytar, Y., Budden, D., Vecerík, M., Sushkov, O. O., Barker, D., Scholz, J., Denil, M., de Freitas, N., and Wang, Z · 2019
Earlier work this paper cites.
Evaluating task-agnostic exploration for fixed-batch learning of arbitrary future tasks
Original
Dasagi, V., Lee, R., Bruce, J., and Leitner, J · 2019
Earlier work this paper cites.
Low-level control of a quadrotor with deep model-based reinforcement learning
Lambert, N. O., Drew, D. S., Yaconelli, J., Levine, S., Calandra, R., and Pister, K. S. J · 2019
Earlier work this paper cites.