Fetching the paper…
Reading the bibliography…
World models for environments with many objects face a combinatorial explosion of states: as the number of objects increases, the number of possible arrangements grows exponentially.
The hungarian method for the assignment problem
Harold W Kuhn · 1955
Earlier work this paper cites.
Efficient solution algorithms for factored mdps
Carlos Guestrin, Daphne Koller, Ronald Parr, and Shobha Venkataraman · 2003
Earlier work this paper cites.
The cross entropy method for fast policy search
Shie Mannor, Reuven Y. Rubinstein, and Yohai Gat · 2003
Earlier work this paper cites.
An Algebraic Approach to Abstraction in Reinforcement Learning
Balaraman Ravindran · 2004
Earlier work this paper cites.
A new model for learning in graph domains
M. Gori, G. Monfardini, and F. Scarselli · 2005
Earlier work this paper cites.
A causal approach to hierarchical decomposition of factored mdps
Anders Jonsson and Andrew G. Barto · 2005
Earlier work this paper cites.
Planning algorithms
Steven M LaValle · 2006
Earlier work this paper cites.
Defining object types and options using mdp homomorphisms
Alicia P. Wolfe · 2006
Earlier work this paper cites.
Decision tree methods for finding reusable MDP homomorphisms
Alicia P. Wolfe and Andrew G. Barto · 2006
Earlier work this paper cites.
An object-oriented representation for efficient reinforcement learning
Carlos Diuk, Andre Cohen, and Michael L. Littman · 2008
Earlier work this paper cites.
PILCO: A model-based and data-efficient approach to policy search
Marc Peter Deisenroth and Carl Edward Rasmussen · 2011
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin A. Riedmiller, Andreas Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Earlier work this paper cites.
Embed to control: A locally linear latent dynamics model for control from raw images
Manuel Watter, Jost Tobias Springenberg, Joschka Boedecker, and Martin A. Riedmiller · 2015
Cited alongside, same era.
Lei Jimmy Ba, Jamie Ryan Kiros, and Geoffrey E. Hinton · 2016
Cited alongside, same era.
Interaction networks for learning about objects, relations and physics
Peter W. Battaglia, Razvan Pascanu, Matthew Lai, Danilo Jimenez Rezende, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Deep variational bayes filters: Unsupervised learning of state space models from raw data
Maximilian Karl, Maximilian Soelch, Justin Bayer, and Patrick van der Smagt · 2017
Cited alongside, same era.
Robust locally-linear controllable embedding
Ershad Banijamali, Rui Shu, Mohammad Ghavamzadeh, Hung Hai Bui, and Ali Ghodsi · 2018
Cited alongside, same era.
Entity abstraction in visual model-based reinforcement learning
Rishi Veerapaneni, John D. Co-Reyes, Michael Chang, Michael Janner, Chelsea Finn, Jiajun Wu, Joshua B. Tenenbaum, and Sergey Levine · 2019
Later among the works it cites.
Object-centric forward modeling for model predictive control
Yufei Ye, Dhiraj Gandhi, Abhinav Gupta, and Shubham Tulsiani · 2019
Later among the works it cites.
Object detection with deep learning: A review
Zhong-Qiu Zhao, Peng Zheng, Shou-tao Xu, and Xindong Wu · 2019
Later among the works it cites.
On the binding problem in artificial neural networks
Klaus Greff, Sjoerd van Steenkiste, and Jürgen Schmidhuber · 2020
Later among the works it cites.
Better set representations for relational reasoning
Qian Huang, Horace He, Abhay Singh, Yan Zhang, Ser-Nam Lim, and Austin R. Benson · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Efficient model-based deep reinforcement learning with variational state tabulation
Dane Corneil, Wulfram Gerstner, and Johanni Brea · 2018
Cited alongside, same era.
Neural relational inference for interacting systems
Thomas Kipf, Ethan Fetaya, Kuan-Chieh Wang, Max Welling, and Richard Zemel · 2018
Cited alongside, same era.
State representation learning for control: An overview
Timothée Lesort, Natalia Díaz Rodríguez, Jean-François Goudou, and David Filliat · 2018
Cited alongside, same era.
Visual reinforcement learning with imagined goals
Ashvin Nair, Vitchyr Pong, Murtaza Dalal, Shikhar Bahl, Steven Lin, and Sergey Levine · 2018
Cited alongside, same era.
Graph networks as learnable physics engines for inference and control
Alvaro Sanchez-Gonzalez, Nicolas Heess, Jost Tobias Springenberg, Josh Merel, Martin Riedmiller, Raia Hadsell, and Peter Battaglia · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Cited alongside, same era.
PHYRE: A new benchmark for physical reasoning
Anton Bakhtin, Laurens van der Maaten, Justin Johnson, Laura Gustafson, and Ross B. Girshick · 2019
Cited alongside, same era.
Thomas N. Kipf, Elise van der Pol, and Max Welling · 2020
Later among the works it cites.
Structured object-aware physics prediction for video modeling and planning
Jannik Kossen, Karl Stelzner, Marcel Hussing, Claas Voelcker, and Kristian Kersting · 2020
Later among the works it cites.
Prediction, consistency, curvature: Representation learning for locally-linear control
Nir Levine, Yinlam Chow, Rui Shu, Ang Li, Mohammad Ghavamzadeh, and Hung Bui · 2020
Later among the works it cites.
Towards practical multi-object manipulation using relational reinforcement learning
Richard Li, Allan Jabri, Trevor Darrell, and Pulkit Agrawal · 2020
Later among the works it cites.
Object-centric learning with slot attention
Francesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran, Georg Heigold, Jakob Uszkoreit, Alexey Dosovitskiy, and Thomas Kipf · 2020
Later among the works it cites.
Policy learning in se (3) action spaces
Dian Wang, Colin Kohler, and Robert Platt · 2020
Later among the works it cites.
Action priors for large action spaces in robotics
Ondrej Biza, Dian Wang, Robert Platt Jr., Jan-Willem van de Meent, and Lawson L. S. Wong · 2021
Later among the works it cites.
Learning long-term visual dynamics with region proposal interaction networks
Haozhi Qi, Xiaolong Wang, Deepak Pathak, Yi Ma, and Jitendra Malik · 2021
Later among the works it cites.