Fetching the paper…
Reading the bibliography…
World models are self-supervised predictive models of how the world evolves.
An experimental study of apparent behavior
Fritz Heider and Marianne Simmel · 1944
Earlier work this paper cites.
The origins of intelligence in children
J. Piaget · 1952
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff · 1978
Earlier work this paper cites.
Developing theories of mind
Janet W Astington, Paul L Harris, and David R Olson · 1990
Earlier work this paper cites.
Query by committee
H Sebastian Seung, Manfred Opper, and Haim Sompolinsky · 1992
Earlier work this paper cites.
The child’s theory of mind
Henry M Wellman · 1992
Earlier work this paper cites.
Taking the intentional stance at 12 months of age
György Gergely, Zoltán Nádasdy, Gergely Csibra, and Szilvia Bíró · 1995
Earlier work this paper cites.
The many faces of configural processing
Daphne Maurer, Richard Le Grand, and Catherine J Mondloch · 2002
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
Pierre-Yves Oudeyer, Frdric Kaplan, and Verena V Hafner · 2007
Earlier work this paper cites.
An analysis of model-based interval estimation for markov decision processes
Alexander L Strehl and Michael L Littman · 2008
Earlier work this paper cites.
Formal theory of creativity, fun, and intrinsic motivation (1990 – 2010)
J. Schmidhuber · 2010
Earlier work this paper cites.
Active Learning , volume 18
Burr Settles · 2011
Earlier work this paper cites.
The goldilocks effect: Human infants allocate attention to visual sequences that are neither too simple nor too complex
Celeste Kidd, Steven T Piantadosi, and Richard N Aslin · 2012
Earlier work this paper cites.
Infants’ perception of chasing
Willem E Frankenhuis, Bailey House, H Clark Barrett, and Scott P Johnson · 2013
Earlier work this paper cites.
Attention to eyes is present but in decline in 2–6-month-old infants later diagnosed with autism
Warren Jones and Ami Klin · 2013
Earlier work this paper cites.
Android+ sphero: teaching mobile computing and robotics in a single course
Stan Kurkovsky · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Earlier work this paper cites.
Intrinsically motivated learning of real-world sensorimotor skills with developmental constraints
Pierre-Yves Oudeyer, Adrien Baranes, and Frédéric Kaplan · 2013
Earlier work this paper cites.
The autism diagnostic observation schedule, module 4: revised algorithm and standardized severity scores
Vanessa Hus and Catherine Lord · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Variational information maximisation for intrinsically motivated reinforcement learning
Shakir Mohamed and Danilo Jimenez Rezende · 2015
Cited alongside, same era.
Incentivizing exploration in reinforcement learning with deep predictive models
Bradly Stadie, Sergey Levine, and Pieter Abbeel · 2015
Cited alongside, same era.
Observing the unexpected enhances infants’ learning and exploration
Aimee E Stahl and Lisa Feigenson · 2015
Cited alongside, same era.
Deep visual foresight for planning robot motion
Chelsea Finn and Sergey Levine · 2017
Later among the works it cites.
Automated curriculum learning for neural networks
Alex Graves, Marc G Bellemare, Jacob Menick, Remi Munos, and Koray Kavukcuoglu · 2017
Later among the works it cites.
Count-based exploration with neural density models
Georg Ostrovski, Marc G Bellemare, Aäron van den Oord, and Rémi Munos · 2017
Later among the works it cites.
Curiosity-driven exploration by self-supervised prediction
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell · 2017
Later among the works it cites.
Relational inductive biases, deep learning, and graph networks
Peter W Battaglia, Jessica B Hamrick, Victor Bapst, Alvaro Sanchez-Gonzalez, Vinicius Zambaldi, Mateusz Malinowski, Andrea Tacchetti, David Raposo, Adam Santoro, Ryan Faulkner, Caglar Gulcehre, Francis Song, Andrew Ballard, Justin Gilmer, George Dahl, Ashish Vaswani, Kelsey Allen, Charles Nash, Victoria Langston, Chris Dyer, Nicolas Heess, Daan Wierstra, Pushmeet Kohli, Matt Botvinick, Oriol Vinyals, Yujia Li, and Razvan Pascanu · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Interaction networks for learning about objects, relations and physics
Peter Battaglia, Razvan Pascanu, Matthew Lai, Danilo Jimenez Rezende, et al · 2016
Cited alongside, same era.
Unifying count-based exploration and intrinsic motivation
Marc Bellemare, Sriram Srinivasan, Georg Ostrovski, Tom Schaul, David Saxton, and Remi Munos · 2016
Cited alongside, same era.
A compositional object-based approach to learning physical dynamics
Michael B Chang, Tomer Ullman, Antonio Torralba, and Joshua B Tenenbaum · 2016
Cited alongside, same era.
Density estimation using real NVP
Laurent Dinh, Jascha Sohl-Dickstein, and Samy Bengio · 2016
Cited alongside, same era.
Unsupervised learning for physical interaction through video prediction
Chelsea Finn, Ian Goodfellow, and Sergey Levine · 2016
Cited alongside, same era.
Explicit information for category-orthogonal object properties increases along the ventral stream
Ha Hong, Daniel LK Yamins, Najib J Majaj, and James J DiCarlo · 2016
Cited alongside, same era.
Vime: Variational information maximizing exploration
Rein Houthooft, Xi Chen, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel · 2016
Cited alongside, same era.
Later among the works it cites.
David Ha and Jürgen Schmidhuber · 2018
Later among the works it cites.
Learning to play with intrinsically-motivated self-aware agents
Nick Haber, Damian Mrowca, Stephanie Wang, Li Fei-Fei, and Daniel LK Yamins · 2018
Later among the works it cites.
Challenging common assumptions in the unsupervised learning of disentangled representations
Francesco Locatello, Stefan Bauer, Mario Lucic, Gunnar Rätsch, Sylvain Gelly, Bernhard Schölkopf, and Olivier Bachem · 2018
Later among the works it cites.
Flexible neural representation for physics prediction
Damian Mrowca, Chengxu Zhuang, Elias Wang, Nick Haber, Li Fei-Fei, Joshua B Tenenbaum, and Daniel L K Yamins · 2018
Later among the works it cites.
Predrnn++: Towards a resolution of the deep-in-time dilemma in spatiotemporal predictive learning
Yunbo Wang, Zhifeng Gao, Mingsheng Long, Jianmin Wang, and Philip S Yu · 2018
Later among the works it cites.
A survey on intrinsic motivation in reinforcement learning
Arthur Aubret, Laetitia Matignon, and Salima Hassas · 2019
Later among the works it cites.
Common knowledge, coordination, and strategic mentalizing in human social life
Julian De Freitas, Kyle Thomas, Peter DeScioli, and Steven Pinker · 2019
Later among the works it cites.
Learning dynamics model in reinforcement learning by incorporating the long term future
Nan Rosemary Ke, Amanpreet Singh, Ahmed Touati, Anirudh Goyal, Yoshua Bengio, Devi Parikh, and Dhruv Batra · 2019
Later among the works it cites.
Adapting behaviour via intrinsic reward: A survey and empirical study
Cam Linke, Nadia M Ady, Martha White, Thomas Degris, and Adam White · 2019
Later among the works it cites.
Self-supervised exploration via disagreement
Deepak Pathak, Dhiraj Gandhi, and Abhinav Gupta · 2019
Later among the works it cites.
Modeling expectation violation in intuitive physics with coarse probabilistic object representations
Kevin Smith, Lingjie Mei, Shunyu Yao, Jiajun Wu, Elizabeth Spelke, Josh Tenenbaum, and Tomer Ullman · 2019
Later among the works it cites.
Model imitation for model-based reinforcement learning
Yueh-Hua Wu, Ting-Han Fan, Peter J. Ramadge, and Hao Su · 2019
Later among the works it cites.