Fetching the paper…
Reading the bibliography…
Domain adaptation is an important open problem in deep reinforcement learning (RL).
Simple memory: A theory for archicortex
Marr, D · 1971
Earlier work this paper cites.
Configural association theory: The role of the hippocampal formation in learning, memory, and amnesia
Sutherland, Robert J and Rudy, Jerry W · 1989
Earlier work this paper cites.
Learning from delayed rewards
Watkins, Christopher John Cornish Hellaby · 1989
Earlier work this paper cites.
Long-lasting perceptual priming and semantic learning in amnesia: a case experiment
Tulving, Endel, Hayman, CA, and Macdonald, Carol A · 1991
Earlier work this paper cites.
Learning factorial codes by predictability minimization
Schmidhuber, Jürgen · 1992
Earlier work this paper cites.
Efficient dynamic programming-based learning for control
Peng, J · 1993
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Puterman, Martin L · 1994
Earlier work this paper cites.
Why there are complementary learning systems in the hippocampus and neocortex: insights from the successes and failures of connectionist models of learning and memory
McClelland, James L, McNaughton, Bruce L, and O’Reilly, Randall C · 1995
Earlier work this paper cites.
Incremental multi-step q-learning
Peng, Jing and Williams, Ronald J · 1996
Earlier work this paper cites.
Reinforcement Learning: An Introduction
Sutton, Richard S. and Barto, Andrew G · 1998
Earlier work this paper cites.
Predictive representations of state
Littman, Michael L., Sutton, Richard S., and Singh, Satinder · 2001
Earlier work this paper cites.
Modeling hippocampal and neocortical contributions to recognition memory: a complementary-learning-systems approach
Norman, Kenneth A and O’Reilly, Randall C · 2003
Earlier work this paper cites.
An experts algorithm for transfer learning
Talvitie, Erik and Singh, Satinder · 2007
Earlier work this paper cites.
Retinal image quality and postnatal visual experience during infancy
Candy, T. Rowan, Wang, Jingyun, and Ravikumar, Sowmya · 2009
Earlier work this paper cites.
Development of visual acuity and contrast sensitivity in children
Leat, Susan J., Yadav, Naveen K., and Irving, Elizabeth L · 2009
Earlier work this paper cites.
A survey on transfer learning
Pan, Sinno Jialin and Yang, Quiang · 2009
Earlier work this paper cites.
Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion
Vincent, Pascal, Larochelle, Hugo, Lajoie, Isabelle, Bengio, Yoshua, and Manzagol, Pierre-Antoine · 2010
Earlier work this paper cites.
Transforming auto-encoders
Hinton, G., Krizhevsky, A., and Wang, S. D · 2011
Earlier work this paper cites.
Model-free reinforcement learning with continuous action in practice
Degris, Thomas, Pilarski Patrick M and Sutton, Richard S · 2012
Earlier work this paper cites.
Disentangling factors of variation via generative entangling
Desjardins, G., Courville, A., and Bengio, Y · 2012
Earlier work this paper cites.
Efficient bayes-adaptive reinforcement learning using sample-based search
Guez, Arthur, Silver, David, and Dayan, Peter · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoffrey E · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Bengio, Y., Courville, A., and Vincent, P · 2013
Earlier work this paper cites.
Incremental semantically grounded learning from demonstration
Niekum, Scott, Chitta, Sachin, Barto, Andrew G, Marthi, Bhaskara, and Osentoski, Sarah · 2013
Cited alongside, same era.
High-dimensional probability estimation with deep density models
Rippel, Oren and Adams, Ryan Prescott · 2013
Cited alongside, same era.
Learning the irreducible representations of commutative lie groups
Cohen, Taco and Welling, Max · 2014
Cited alongside, same era.
Generative adversarial nets
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, Jimmy · 2014
Cited alongside, same era.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2014
Cited alongside, same era.
Model-free episodic control
Blundell, Charles, Uria, Benigno, Pritzel, Alexander, Li, Yazhe, Ruderman, Avraham, Leibo, Joel Z, Rae, Jack, Wierstra, Daan, and Hassabis, Demis · 2016
Later among the works it cites.
Infogan: Interpretable representation learning by information maximizing generative adversarial nets
Chen, Xi, Duan, Yan, Houthooft, Rein, Schulman, John, Sutskever, Ilya, and Abbeel, Pieter · 2016
Later among the works it cites.
Learning transferable policies for monocular reactive mav control
Daftry, Shreyansh, Bagnell, J. Andrew, and Hebert, Martial · 2016
Later among the works it cites.
Generating images with perceptual similarity metrics based on deep networks
Dosovitskiy, Alexey and Brox, Thomas · 2016
Later among the works it cites.
Towards deep symbolic reinforcement learning
Garnelo, Marta, Arulkumaran, Kai, and Shanahan, Murray · 2016
Later among the works it cites.
Bayesian representation learning with oracle constraints
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning neural network policies with guided policy search under unknown dynamics
Levine, Sergey and Abbeel, Pieter · 2014
Cited alongside, same era.
Learning to disentangle factors of variation with manifold interaction
Reed, Scott, Sohn, Kihyuk, Zhang, Yuting, and Lee, Honglak · 2014
Cited alongside, same era.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, Danilo J., Mohamed, Shakir, and Wierstra, Daan · 2014
Cited alongside, same era.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Abadi, Martin, Agarwal, Ashish, and et al, Paul Barham · 2015
Cited alongside, same era.
Discovering hidden factors of variation in deep networks
Cheung, Brian, Levezey, Jesse A., Bansal, Arjun K., and Olshausen, Bruno A · 2015
Cited alongside, same era.
Transformation properties of learned visual representations
Cohen, T. and Welling, M · 2015
Cited alongside, same era.
Karaletsos, Theofanis, Belongie, Serge, and Rätsch, Gunnar · 2016
Later among the works it cites.
Building machines that learn and think like people
Lake, Brenden M., Ullman, Tomer D., Tenenbaum, Joshua B., and Gershman, Samuel J · 2016
Later among the works it cites.
Autoencoding beyond pixels using a learned similarity metric
Larsen, Anders Boesen Lindbo, Sønderby, Søren Kaae, Larochelle, Hugo, and Winther, Ole · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
Mnih, Volodymyr, Badia, Adrià Puigdomènech, Mirza, Mehdi, Graves, Alex, Lillicrap, Timothy P., Harley, Tim, Silver, David, and Kavukcuoglu, Koray · 2016
Later among the works it cites.
Context encoders: Feature learning by inpainting
Pathak, Deepak, Krähenbühl, Philipp, Donahue, Jeff, Darrell, Trevor, and Efros, Alexei A · 2016
Later among the works it cites.
A survey of inductive biases for factorial Representation-Learning
Ridgeway, Karl · 2016
Later among the works it cites.
Sim-to-real robot learning from pixels with progressive nets
Rusu, Andrei A., Vecerik, Matej, Rothörl, Thomas, Heess, Nicolas, Pascanu, Razvan, and Hadsell, Raia · 2016
Later among the works it cites.
High-dimensional continuous control using generalized advantage estimation
Schulman, John, Moritz, Philipp, Levine, Sergey, Jordan, Michael, and Abbeel, Pieter · 2016
Later among the works it cites.
Adapting deep visuomotor representations with weak pairwise constraints
Tzeng, Eric, Devin, Coline, Hoffman, Judy, Finn, Chelsea, Abbeel, Pieter, Levine, Sergey, Saenko, Kate, and Darrell, Trevor · 2016
Later among the works it cites.
Understanding visual concepts with continuation learning
Whitney, William F., Chang, Michael, Kulkarni, Tejas, and Tenenbaum, Joshua B · 2016
Later among the works it cites.
Generalizing skills with semi-supervised reinforcement learning
Finn, Chelsea, Yu, Tianhe, Fu, Justin, Abbeel, Pieter, and Levine, Sergey · 2017
Closest in time.
Learning invariant feature spaces to transfer skills with reinforcement learning
Gupta, Abhishek, Devin, Coline, Liu, YuXuan, Abbeel, Pieter, and Levine, Sergey · 2017
Closest in time.
Beta-vae: Learning basic visual concepts with a constrained variational framework
Higgins, Irina, Matthey, Loic, Pal, Arka, Burgess, Christopher, Glorot, Xavier, Botvinick, Matthew, Mohamed, Shakir, and Lerchner, Alexander · 2017
Closest in time.
Reinforcement learning with unsupervised auxiliary tasks
Jaderberg, Max, Mnih, Volodymyr, Czarnecki, Wojciech Marian, Schaul, Tom, Leibo, Joel Z, Silver, David, and Kavukcuoglu, Koray · 2017
Closest in time.
Unsupervised learning of state representations for multiple tasks
Raffin, Antonin, Höfer, Sebastian, Jonschkowski, Rico, Brock, Oliver, and Stulp, Freek · 2017
Closest in time.
Attend, adapt and transfer: Attentive deep architecture for adaptive transfer from multiple sources in the same domain
Rajendran, Janarthanan, Lakshminarayanan, Aravind, Khapra, Mitesh M., P, Prasanna, and Ravindran, Balaraman · 2017
Closest in time.
Domain randomization for transferring deep neural networks from simulation to the real world
Tobin, Josh, Fong, Rachel, Ray, Alex, Schneider, Jonas, Zaremba, Wojciech, and Abbeel, Pieter · 2017
Closest in time.
Improving generative adversarial networks with denoising feature matching
Warde-Farley, David and Bengio, Yoshua · 2017
Closest in time.