Fetching the paper…
Reading the bibliography…
A generally intelligent learner should generalize to more complex tasks than it has previously encountered, but the two common paradigms in machine learning -- either training a separate learner per task or training a single learner for all tasks -- both have difficulty with such generalization because they do not leverage the compositional structure of the task distribution.
The sciences of the artificial
Herbert A Simon · 1969
Earlier work this paper cites.
Non-holographic associative memory
David J Willshaw, O Peter Buneman, and Hugh Christopher Longuet-Higgins · 1969
Earlier work this paper cites.
Correlation matrix memories
Teuvo Kohonen · 1972
Earlier work this paper cites.
Neural networks and physical systems with emergent collective computational abilities
John J Hopfield · 1982
Earlier work this paper cites.
Evolutionary principles in self-referential learning, or on learning how to learn: the meta-meta-… hook
Jürgen Schmidhuber · 1987
Earlier work this paper cites.
Connectionism and cognitive architecture: A critical analysis
Jerry A Fodor and Zenon W Pylyshyn · 1988
Earlier work this paper cites.
The science of design: Creating the artificial
Herbert A Simon · 1988
Earlier work this paper cites.
The adaptive character of thought
John Robert Anderson · 1990
Earlier work this paper cites.
Towards compositional learning with dynamic neural networks
Jürgen Schmidhuber · 1990
Earlier work this paper cites.
Principles of object perception
Elizabeth S Spelke · 1990
Earlier work this paper cites.
Adaptive mixtures of local experts
Robert A Jacobs, Michael I Jordan, Steven J Nowlan, and Geoffrey E Hinton · 1991
Earlier work this paper cites.
Principles of metareasoning
Stuart Russell and Eric Wefald · 1991
Earlier work this paper cites.
Learning to control fast-weight memories: An alternative to dynamic recurrent networks
Jürgen Schmidhuber · 1992
Earlier work this paper cites.
Baby born talking—describes heaven
S Pinker · 1994
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, Patrick Haffner, et al · 1998
Earlier work this paper cites.
Rethinking eliminative connectionism
Gary F Marcus · 1998
Earlier work this paper cites.
Relational reinforcement learning
Sašo Džeroski, Luc De Raedt, and Kurt Driessens · 2001
Earlier work this paper cites.
The compositionality papers
Jerry A Fodor and Ernest Lepore · 2002
Earlier work this paper cites.
Recent advances in hierarchical reinforcement learning
Andrew G Barto and Sridhar Mahadevan · 2003
Earlier work this paper cites.
Abstraction and reformulation in artificial intelligence
Robert C Holte and Berthe Y Choueiry · 2003
Earlier work this paper cites.
Hierarchically organized behavior and its neural foundations: a reinforcement learning perspective
Matthew M Botvinick, Yael Niv, and Andrew C Barto · 2009
Earlier work this paper cites.
Ultimate cognition à la gödel
Jürgen Schmidhuber · 2009
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Vinod Nair and Geoffrey E Hinton · 2010
Earlier work this paper cites.
Reconstructing constructivism: Causal models, bayesian learning mechanisms, and the theory theory
Alison Gopnik and Henry M Wellman · 2012
Earlier work this paper cites.
Self-delimiting neural networks
Jürgen Schmidhuber · 2012
Earlier work this paper cites.
Learning to learn
Sebastian Thrun and Lorien Pratt · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Yoshua Bengio, Aaron Courville, and Pascal Vincent · 2013
Earlier work this paper cites.
Bootstrap learning via modular concept discovery
Eyal Dechter, Jonathan Malmaud, Ryan P Adams, and Joshua B Tenenbaum · 2013
Earlier work this paper cites.
Models of information processing in the brain
James A Anderson and Geoffrey E Hinton · 2014
Earlier work this paper cites.
The Architecture of Cognition: Rethinking Fodor and Pylyshyn’s Systematicity Challenge
Paco Calvo and John Symons · 2014
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Alex Graves, Greg Wayne, and Ivo Danihelka · 2014
Earlier work this paper cites.
Selecting computations: Theory and applications
Nicholas Hay, Stuart Russell, David Tolpin, and Solomon Eyal Shimony · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Optimal behavioral hierarchy
Alec Solway, Carlos Diuk, Natalia Córdova, Debbie Yee, Andrew G Barto, Yael Niv, and Matthew M Botvinick · 2014
Earlier work this paper cites.
Striving for simplicity: The all convolutional net
Jost Tobias Springenberg, Alexey Dosovitskiy, Thomas Brox, and Martin Riedmiller · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Earlier work this paper cites.
Wojciech Zaremba and Ilya Sutskever · 2014
Cited alongside, same era.
15 why theories of concepts should not ignore the problem of acquisition
Susan Carey · 2015
Cited alongside, same era.
Learning to transduce with unbounded memory
Edward Grefenstette, Karl Moritz Hermann, Mustafa Suleyman, and Phil Blunsom · 2015
Cited alongside, same era.
Rational use of cognitive resources: Levels of analysis between the computational and the algorithmic
Thomas L Griffiths, Falk Lieder, and Noah D Goodman · 2015
Cited alongside, same era.
Spatial transformer networks
Max Jaderberg, Karen Simonyan, Andrew Zisserman, et al · 2015
Cited alongside, same era.
Inferring algorithmic patterns with stack-augmented recurrent nets
Armand Joulin and Tomas Mikolov · 2015
Learning to select computations
Falk Lieder, Frederick Callaway, Sayan Gul, Paul M Krueger, and Thomas L Griffiths · 2017
Later among the works it cites.
Inverse compositional spatial transformer networks
Chen-Hsuan Lin and Simon Lucey · 2017
Later among the works it cites.
Gradient episodic memory for continual learning
David Lopez-Paz et al · 2017
Later among the works it cites.
Zero-shot task generalization with multi-task deep reinforcement learning
Junhyuk Oh, Satinder Singh, Honglak Lee, and Pushmeet Kohli · 2017
Later among the works it cites.
Learning independent causal mechanisms
Giambattista Parascandolo, Mateo Rojas-Carulla, Niki Kilbertus, and Bernhard Schölkopf · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Łukasz Kaiser and Ilya Sutskever · 2015
Cited alongside, same era.
Deep convolutional inverse graphics network
Tejas D Kulkarni, William F Whitney, Pushmeet Kohli, and Josh Tenenbaum · 2015
Cited alongside, same era.
Karol Kurach, Marcin Andrychowicz, and Ilya Sutskever · 2015
Cited alongside, same era.
Human-level concept learning through probabilistic program induction
Brenden M Lake, Ruslan Salakhutdinov, and Joshua B Tenenbaum · 2015
Cited alongside, same era.
Neural programmer-interpreters
Scott Reed and Nando De Freitas · 2015
Cited alongside, same era.
End-to-end memory networks
Sainbayar Sukhbaatar, Jason Weston, Rob Fergus, et al · 2015
Cited alongside, same era.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Later among the works it cites.
A simple neural network module for relational reasoning
Adam Santoro, David Raposo, David GT Barrett, Mateusz Malinowski, Razvan Pascanu, Peter Battaglia, and Timothy Lillicrap · 2017
Later among the works it cites.
Gated fast weights for on-the-fly neural program generation
Imanol Schlag and Jürgen Schmidhuber · 2017
Later among the works it cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Independently controllable features
Valentin Thomas, Jules Pondard, Emmanuel Bengio, Marc Sarfati, Philippe Beaudoin, Marie-Jean Meurs, Joelle Pineau, Doina Precup, and Yoshua Bengio · 2017
Later among the works it cites.
Feudal networks for hierarchical reinforcement learning
Alexander Sasha Vezhnevets, Simon Osindero, Tom Schaul, Nicolas Heess, Max Jaderberg, David Silver, and Koray Kavukcuoglu · 2017
Later among the works it cites.
Neural task programming: Learning to generalize across hierarchical tasks
Danfei Xu, Suraj Nair, Yuke Zhu, Julian Gao, Animesh Garg, Li Fei-Fei, and Silvio Savarese · 2017
Later among the works it cites.
Ferran Alet, Tomás Lozano-Pérez, and Leslie P Kaelbling · 2018
Closest in time.
Systematic generalization: What is required and can it be learned?
Dzmitry Bahdanau, Shikhar Murty, Michael Noukhovitch, Thien Huu Nguyen, Harm de Vries, and Aaron Courville · 2018
Closest in time.
Relational inductive biases, deep learning, and graph networks
Peter W Battaglia, Jessica B Hamrick, Victor Bapst, Alvaro Sanchez-Gonzalez, Vinicius Zambaldi, Mateusz Malinowski, Andrea Tacchetti, David Raposo, Adam Santoro, Ryan Faulkner, et al · 2018
Closest in time.
Leveraging grammar and reinforcement learning for neural program synthesis
Rudy Bunel, Matthew Hausknecht, Jacob Devlin, Rishabh Singh, and Pushmeet Kohli · 2018
Closest in time.
Learning libraries of subroutines for neurally–guided bayesian program induction
Kevin Ellis, Lucas Morales, Mathias Sablé-Meyer, Armando Solar-Lezama, and Josh Tenenbaum · 2018
Closest in time.
Synthesizing programs for images using reinforced adversarial learning
Yaroslav Ganin, Tejas Kulkarni, Igor Babuschkin, S.M. Ali Eslami, and Orial Vinyals · 2018
Closest in time.
Recasting gradient-based meta-learning as hierarchical bayes
Erin Grant, Chelsea Finn, Sergey Levine, Trevor Darrell, and Thomas Griffiths · 2018
Closest in time.
David Ha and Jürgen Schmidhuber · 2018
Closest in time.
Towards a definition of disentangled representations
Irina Higgins, David Amos, David Pfau, Sebastien Racaniere, Loic Matthey, Danilo Rezende, and Alexander Lerchner · 2018
Closest in time.
Matrix capsules with em routing
Geoffrey Hinton, Nicholas Frosst, and Sara Sabour · 2018
Closest in time.
Modular networks: Learning to decompose neural computation
Louis Kirsch, Julius Kunze, and David Barber · 2018
Closest in time.
Memorize or generalize? searching for a compositional rnn in a haystack
Adam Liška, Germán Kruszewski, and Marco Baroni · 2018
Closest in time.
Rearranging the familiar: Testing compositional generalization in recurrent networks
Joao Loula, Marco Baroni, and Brenden M Lake · 2018
Closest in time.
The algebraic mind: Integrating connectionism and cognitive science
Gary F Marcus · 2018
Closest in time.
A simple neural attentive meta-learner
Nikhil Mishra, Mostafa Rohaninejad, Xi Chen, and Pieter Abbeel · 2018
Closest in time.
Data-efficient hierarchical reinforcement learning
Ofir Nachum, Shane Gu, Honglak Lee, and Sergey Levine · 2018
Closest in time.
Visual reinforcement learning with imagined goals
Ashvin V Nair, Vitchyr Pong, Murtaza Dalal, Shikhar Bahl, Steven Lin, and Sergey Levine · 2018
Closest in time.
On first-order meta-learning algorithms
Alex Nichol, Joshua Achiam, and John Schulman · 2018
Closest in time.
Routing networks: Adaptive selection of non-linear functions for multi-task learning
Clemens Rosenbaum, Tim Klinger, and Matthew Riemer · 2018
Closest in time.
Representational efficiency outweighs action efficiency in human program induction
S Sanborn, D Bourgin, M Chang, and T Griffiths · 2018
Closest in time.
Graph networks as learnable physics engines for inference and control
Alvaro Sanchez-Gonzalez, Nicolas Heess, Jost Tobias Springenberg, Josh Merel, Martin Riedmiller, Raia Hadsell, and Peter Battaglia · 2018
Closest in time.
Learning to reason with third order tensor products
Imanol Schlag and Jürgen Schmidhuber · 2018
Closest in time.
Aravind Srinivas, Allan Jabri, Pieter Abbeel, Sergey Levine, and Chelsea Finn · 2018
Closest in time.
Houdini: Lifelong learning as program synthesis
Lazar Valkov, Dipak Chaudhari, Akash Srivastava, Charles Sutton, and Swarat Chaudhuri · 2018
Closest in time.
Relational neural expectation maximization: Unsupervised discovery of objects and their interactions
Sjoerd van Steenkiste, Michael Chang, Klaus Greff, and Jürgen Schmidhuber · 2018
Closest in time.
Doing more with less: Meta-reasoning and meta-learning in humans and machines
Thomas L Griffiths, Fred Callaway, Michael B Chang, Erin Grant, Paul M Krueger, and Falk Lieder · 2019
Closest in time.