Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) agents are often sensitive to visual changes that were unseen in their training environments.
Predictive information
Bialek, W. and Tishby, N · 1999
Earlier work this paper cites.
The information bottleneck method
Tishby, N., Pereira, F. C., and Bialek, W · 2000
Earlier work this paper cites.
Visualizing data using t-sne
Van der Maaten, L. and Hinton, G · 2008
Earlier work this paper cites.
Deep auto-encoder neural networks in reinforcement learning
Lange, S. and Riedmiller, M · 2010
Earlier work this paper cites.
Autonomous reinforcement learning on raw visual input data in a real world application
Lange, S., Riedmiller, M., and Voigtländer, A · 2012
Earlier work this paper cites.
Learning state representations with robotic priors
Jonschkowski, R. and Brock, O · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al · 2015
Earlier work this paper cites.
From pixels to torques: Policy learning with deep dynamical models
Wahlström, N., Schön, T. B., and Desienroth, M. P · 2015
Earlier work this paper cites.
Embed to control: a locally linear latent dynamics model for control from raw images
Watter, M., Springenberg, J. T., Boedecker, J., and Riedmiller, M · 2015
Earlier work this paper cites.
End to end learning for self-driving cars
Bojarski, M., Del Testa, D., Dworakowski, D., Firner, B., Flepp, B., Goyal, P., Jackel, L. D., Monfort, M., Muller, U., Zhang, J., et al · 2016
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
Levine, S., Finn, C., Darrell, T., and Abbeel, P · 2016
Earlier work this paper cites.
The kinetics human action video dataset
Kay, W., Carreira, J., Simonyan, K., Zhang, B., Hillier, C., Vijayanarasimhan, S., Viola, F., Green, T., Back, T., Natsev, P., et al · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Earlier work this paper cites.
Paying more attention to attention: Improving the performance of convolutional neural networks via attention transfer
Zagoruyko, S. and Komodakis, N · 2017
Earlier work this paper cites.
Mutual information neural estimation
Belghazi, M. I., Baratin, A., Rajeshwar, S., Ozair, S., Bengio, Y., Courville, A., and Hjelm, D · 2018
Cited alongside, same era.
Generalization and regularization in dqn
Farebrother, J., Machado, M. C., and Bowling, M · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Haarnoja, T., Zhou, A., Abbeel, P., and Levine, S · 2018
Cited alongside, same era.
A survey of multi-view representation learning
Li, Y., Yang, M., and Zhang, Z · 2018
Cited alongside, same era.
Representation learning with contrastive predictive coding
Oord, A. v. d., Li, Y., and Vinyals, O · 2018
Cited alongside, same era.
Leveraging procedural generation to benchmark reinforcement learning
Cobbe, K., Hesse, C., Hilton, J., and Schulman, J · 2020
Later among the works it cites.
Learning robust representations via multi-view information bottleneck
Federici, M., Dutta, A., Forré, P., Kushman, N., and Akata, Z · 2020
Later among the works it cites.
The conditional entropy bottleneck
Fischer, I · 2020
Later among the works it cites.
Dream to control: Learning behaviors by latent imagination
Hafner, D., Lillicrap, T., Ba, J., and Norouzi, M · 2020
Later among the works it cites.
Data-efficient image recognition with contrastive predictive coding
Hénaff, O. J., Srinivas, A., De Fauw, J., Razavi, A., Doersch, C., Eslami, S., and van den Oord, A · 2020
Later among the works it cites.
Deep reinforcement and infomax learning
Mazoure, B., Tachet des Combes, R., DOAN, T. L., Bachman, P., and Hjelm, R. D · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tassa, Y., Doron, Y., Muldal, A., Erez, T., Li, Y., Casas, D. d. L., Budden, D., Abdolmaleki, A., Merel, J., Lefrancq, A., et al · 2018
Cited alongside, same era.
Quantifying generalization in reinforcement learning
Cobbe, K., Klimov, O., Hesse, C., Kim, T., and Schulman, J · 2019
Cited alongside, same era.
Learning latent dynamics for planning from pixels
Hafner, D., Lillicrap, T., Fischer, I., Villegas, R., Ha, D., Lee, H., and Davidson, J · 2019
Cited alongside, same era.
Learning deep representations by mutual information estimation and maximization
Hjelm, R. D., Fedorov, A., Lavoie-Marchildon, S., Grewal, K., Bachman, P., Trischler, A., and Bengio, Y · 2019
Cited alongside, same era.
Emi: Exploration with mutual information
Kim, H., Kim, J., Jeong, Y., Levine, S., and Song, H. O · 2019
Cited alongside, same era.
Multi-view reinforcement learning
Li, M., Wu, L., Jun, W., and Ammar, H. B · 2019
Cited alongside, same era.
On variational bounds of mutual information
Poole, B., Ozair, S., Van Den Oord, A., Alemi, A., and Tucker, G · 2019
Cited alongside, same era.
Later among the works it cites.
Automatic data augmentation for generalization in deep reinforcement learning
Raileanu, R., Goldstein, M., Yarats, D., Kostrikov, I., and Fergus, R · 2020
Later among the works it cites.
Contrastive behavioral similarity embeddings for generalization in reinforcement learning
Agarwal, R., Machado, M. C., Castro, P. S., and Bellemare, M. G · 2021
Closest in time.
Learning task informed abstractions
Fu, X., Yang, G., Agrawal, P., and Jaakkola, T · 2021
Closest in time.
Mastering atari with discrete world models
Hafner, D., Lillicrap, T., Norouzi, M., and Ba, J · 2021
Closest in time.
Decoupling value and policy for generalization in reinforcement learning
Raileanu, R. and Fergus, R · 2021
Closest in time.
Learning representations for pixel-based control: What matters and why?
Tomar, M., Mishra, U. A., Zhang, A., and Taylor, M. E · 2021
Closest in time.
Learning invariant representations for reinforcement learning without reconstruction
Zhang, A., McAllister, R., Calandra, R., Gal, Y., and Levine, S · 2021
Closest in time.