Empowerment–an introduction
Salge, C., Glackin, C., and Polani, D. (2014) · 2014
Later among the works it cites.
Importance weighted autoencoders
Original
Burda, Y., Grosse, R., and Salakhutdinov, R. (2015) · 2015
Later among the works it cites.
Efficient Empowerment
Original
Karl, M., Bayer, J., and van der Smagt, P. (2015) · 2015
Later among the works it cites.
Variational information maximisation for intrinsically motivated reinforcement learning
Mohamed, S. and Rezende, D. J. (2015) · 2015
Later among the works it cites.
Variational inference with normalizing flows
Rezende, D. J. and Mohamed, S. (2015) · 2015
Later among the works it cites.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Original
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G. S., Davis, A., Dean, J., Devin, M., et al. (2016) · 2016
Later among the works it cites.
Unifying count-based exploration and intrinsic motivation
Bellemare, M. G., Srinivasan, S., Ostrovski, G., Schaul, T., Saxton, D., and Munos, R. (2016) · 2016
Later among the works it cites.
Openai gym
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W. (2016) · 2016
Later among the works it cites.
Variational intrinsic control
Original
Gregor, K., Rezende, D. J., and Wierstra, D. (2016) · 2016
Later among the works it cites.
Vime: Variational information maximizing exploration
Houthooft, R., Chen, X., Duan, Y., Schulman, J., De Turck, F., and Abbeel, P. (2016) · 2016
Later among the works it cites.
Improving variational autoencoders with inverse autoregressive flow
Kingma, D. P., Salimans, T., Józefowicz, R., Chen, X., Sutskever, I., and Welling, M. (2016) · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al. (2016) · 2016
Later among the works it cites.
Surprise-based intrinsic motivation for deep reinforcement learning
Original
Achiam, J. and Sastry, S. (2017) · 2017
Closest in time.
Deep variational bayes filters: Unsupervised learning of state space models from raw data
Karl, M., Soelch, M., Bayer, J., and van der Smagt, P. (2017) · 2017
Closest in time.
Intrinsic motivation and automatic curricula via asymmetric self-play
Original
Sukhbaatar, S., Kostrikov, I., Szlam, A., and Fergus, R. (2017) · 2017
Closest in time.