Fetching the paper…
Reading the bibliography…
In many reinforcement learning tasks, the agent has to learn to interact with many objects of different types and generalize to unseen combinations and numbers of objects.
Efficient reinforcement learning in factored mdps
Michael Kearns and Daphne Koller · 1999
Earlier work this paper cites.
Stochastic dynamic programming with factored representations
Craig Boutilier, Richard Dearden, and Moisés Goldszmidt · 2000
Earlier work this paper cites.
Dynamic bayesian networks: Representation, inference and learning
Kevin Murphy · 2002
Earlier work this paper cites.
Envelope-based planning in relational mdps
Natalia Gardiol and Leslie Kaelbling · 2003
Earlier work this paper cites.
Reinforcement learning with factored states and actions
Brian Sallans and Geoffrey E Hinton · 2004
Earlier work this paper cites.
A survey of reinforcement learning in relational domains
Martijn Van Otterlo · 2005
Earlier work this paper cites.
An object-oriented representation for efficient reinforcement learning
Carlos Diuk, Andre Cohen, and Michael L Littman · 2008
Earlier work this paper cites.
Model predictive control
Eduardo F Camacho and Carlos Bordons Alba · 2013
Earlier work this paper cites.
A physics-based model prior for object-oriented mdps
Jonathan Scholz, Martin Levihn, Charles Isbell, and David Wingate · 2014
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Off-policy model-based learning under unknown factored dynamics
Assaf Hallak, François Schnitzler, Timothy Mann, and Shie Mannor · 2015
Earlier work this paper cites.
Attend, infer, repeat: Fast scene understanding with generative models
SM Eslami, Nicolas Heess, Theophane Weber, Yuval Tassa, David Szepesvari, Geoffrey E Hinton, et al · 2016
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Neural relational inference for interacting systems
Thomas Kipf, Ethan Fetaya, Kuan-Chieh Wang, Max Welling, and Richard Zemel · 2018
Cited alongside, same era.
Multi-goal reinforcement learning: Challenging robotics environments and request for research
Matthias Plappert, Marcin Andrychowicz, Alex Ray, Bob McGrew, Bowen Baker, Glenn Powell, Jonas Schneider, Josh Tobin, Maciek Chociej, Peter Welinder, et al · 2018
Cited alongside, same era.
Learning latent dynamics for planning from pixels
Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson · 2019
Cited alongside, same era.
Self-supervised visual reinforcement learning with object-centric representations
Andrii Zadaianchuk, Maximilian Seitzer, and Georg Martius · 2021
Later among the works it cites.
Recurrent independent mechanisms
Anirudh Goyal, Alex Lamb, Jordan Hoffmann, Shagun Sodhani, Sergey Levine, Yoshua Bengio, and Bernhard Schölkopf · 2021
Later among the works it cites.
Self-supervised reinforcement learning with independently controllable subgoals
Andrii Zadaianchuk, Georg Martius, and Fanny Yang · 2022
Later among the works it cites.
Compositional multi-object reinforcement learning with linear relation networks
Davide Mambelli, Frederik Träuble, Stefan Bauer, Bernhard Schölkopf, and Francesco Locatello · 2022
Later among the works it cites.
Toward compositional generalization in object-oriented world modeling
Linfeng Zhao, Lingzhi Kong, Robin Walters, and Lawson LS Wong · 2022
Later among the works it cites.
Policy architectures for compositional generalization in control
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nicholas Watters, Loic Matthey, Matko Bosnjak, Christopher P Burgess, and Alexander Lerchner · 2019
Cited alongside, same era.
Bayesian reinforcement learning in factored pomdps
Sammie Katt, Frans A. Oliehoek, and Christopher Amato · 2019
Cited alongside, same era.
Multi-object search using object-oriented pomdps
Arthur Wandzel, Yoonseon Oh, Michael Fishman, Nishanth Kumar, Lawson L.S. Wong, and Stefanie Tellex · 2019
Cited alongside, same era.
Curiosity-driven multi-criteria hindsight experience replay
John Banister Lanier · 2019
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination
Danijar Hafner, Timothy Lillicrap, Jimmy Ba, and Mohammad Norouzi · 2020
Cited alongside, same era.
Towards practical multi-object manipulation using relational reinforcement learning
Richard Li, Allan Jabri, Trevor Darrell, and Pulkit Agrawal · 2020
Cited alongside, same era.
Structured object-aware physics prediction for video modeling and planning
Jannik Kossen, Karl Stelzner, Marcel Hussing, Claas Voelcker, and Kristian Kersting · 2020
Cited alongside, same era.
Object-centric learning with slot attention
Francesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran, Georg Heigold, Jakob Uszkoreit, Alexey Dosovitskiy, and Thomas Kipf · 2020
Cited alongside, same era.
Allan Zhou, Vikash Kumar, Chelsea Finn, and Aravind Rajeswaran · 2022
Later among the works it cites.
AdaRL: What, where, and how to adapt in transfer reinforcement learning
Biwei Huang, Fan Feng, Chaochao Lu, Sara Magliacane, and Kun Zhang · 2022
Later among the works it cites.
Factored adaptation for non-stationary reinforcement learning
Fan Feng, Biwei Huang, Kun Zhang, and Sara Magliacane · 2022
Later among the works it cites.
Illiterate DALL-e learns to compose
Gautam Singh, Fei Deng, and Sungjin Ahn · 2022
Later among the works it cites.
Transformers are sample efficient world models
Vincent Micheli, Eloi Alonso, and François Fleuret · 2023
Closest in time.
Unsupervised object interaction learning with counterfactual dynamics models
Jongwook Choi, Sungtae Lee, Xinyu Wang, Sungryull Sohn, and Honglak Lee · 2023
Closest in time.
Hierarchical abstraction for combinatorial generalization in object rearrangement
Michael Chang, Alyssa Li Dayan, Franziska Meier, Thomas L. Griffiths, Sergey Levine, and Amy Zhang · 2023
Closest in time.
An investigation into pre-training object-centric representations for reinforcement learning
Jaesik Yoon, Yi-Fu Wu, Heechul Bae, and Sungjin Ahn · 2023
Closest in time.
Interaction-based disentanglement of entities for object-centric world models
Akihiro Nakano, Masahiro Suzuki, and Yutaka Matsuo · 2023
Closest in time.