Fetching the paper…
Reading the bibliography…
Understanding the world in terms of objects and the possible interplays with them is an important cognition ability, especially in robotics manipulation, where many tasks require robot-object interactions.
Curious model-building control systems
J. Schmidhuber · 1991
Earlier work this paper cites.
Nearest neighbor estimates of entropy
H. Singh, N. Misra, V. Hnizdo, A. Fedorowicz, and E. Demchuk · 2003
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
P. Oudeyer, F. Kaplan, and V. V. Hafner · 2007
Earlier work this paper cites.
An object-oriented representation for efficient reinforcement learning
C. Diuk, A. Cohen, and M. L. Littman · 2008
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches, 2014
K. Cho, B. van Merrienboer, D. Bahdanau, and Y. Bengio · 2014
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Earlier work this paper cites.
Concrete problems in ai safety, 2016
D. Amodei, C. Olah, J. Steinhardt, P. Christiano, J. Schulman, and D. Mané · 2016
Earlier work this paper cites.
Data-efficient deep reinforcement learning for dexterous manipulation, 2017
I. Popov et al · 2017
Earlier work this paper cites.
A theory of how columns in the neocortex enable learning the structure of the world
J. Hawkins, S. Ahmad, and Y. Cui · 2017
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction, 2017
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell · 2017
Earlier work this paper cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation
D. Kalashnikov, A. Irpan, P. Pastor, J. Ibarz, A. Herzog, E. Jang, D. Quillen, E. Holly, M. Kalakrishnan, V. Vanhoucke, and S. Levine · 2018
Earlier work this paper cites.
Addressing function approximation error in actor-critic methods, 2018
S. Fujimoto, H. van Hoof, and D. Meger · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Earlier work this paper cites.
World models
D. Ha and J. Schmidhuber · 2018
Earlier work this paper cites.
The developing infant creates a curriculum for statistical learning
L. B. Smith, S. Jayaraman, E. Clerkin, and C. Yu · 2018
Earlier work this paper cites.
Large-scale study of curiosity-driven learning, 2018
Y. Burda, H. Edwards, D. Pathak, A. Storkey, T. Darrell, and A. A. Efros · 2018
Earlier work this paper cites.
Solving rubik’s cube with a robot hand
OpenAI, I. Akkaya, M. Andrychowicz, M. Chociej, M. Litwin, B. McGrew, A. Petron, A. Paino, M. Plappert, G. Powell, R. Ribas, J. Schneider, N. A. Tezak, J. Tworek, P. Welinder, L. Weng, Q. Yuan, W. Zaremba, and L. M. Zhang · 2019
Earlier work this paper cites.
Self‐generated variability in object images predicts vocabulary growth
L. Slone, L. Smith, and C. Yu · 2019
Earlier work this paper cites.
Self-supervised exploration via disagreement, 2019
D. Pathak, D. Gandhi, and A. Gupta · 2019
Earlier work this paper cites.
Monet: Unsupervised scene decomposition and representation, 2019
C. P. Burgess, L. Matthey, N. Watters, R. Kabra, I. Higgins, M. Botvinick, and A. Lerchner · 2019
Cited alongside, same era.
Learning latent dynamics for planning from pixels
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2019
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination, 2020
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi · 2020
Cited alongside, same era.
Mastering atari, go, chess and shogi by planning with a learned model
J. Schrittwieser, I. Antonoglou, T. Hubert, K. Simonyan, L. Sifre, S. Schmitt, A. Guez, E. Lockhart, D. Hassabis, T. Graepel, T. Lillicrap, and D. Silver · 2020
Cited alongside, same era.
Planning to explore via self-supervised world models
R. Sekar, O. Rybkin, K. Daniilidis, P. Abbeel, D. Hafner, and D. Pathak · 2020
Cited alongside, same era.
Object-centric learning with slot attention, 2020
Urlb: Unsupervised reinforcement learning benchmark, 2021
M. Laskin, D. Yarats, H. Liu, K. Lee, A. Zhan, K. Lu, C. Cang, L. Pinto, and P. Abbeel · 2021
Later among the works it cites.
The distracting control suite – a challenging benchmark for reinforcement learning from pixels, 2021
A. Stone, O. Ramirez, K. Konolige, and R. Jonschkowski · 2021
Later among the works it cites.
Daydreamer: World models for physical robot learning, 2022
P. Wu, A. Escontrela, D. Hafner, K. Goldberg, and P. Abbeel · 2022
Later among the works it cites.
Masked world models for visual control, 2022
Y. Seo, D. Hafner, H. Liu, F. Liu, S. James, K. Lee, and P. Abbeel · 2022
Later among the works it cites.
Faulty reward functions in the wild
J. Clark and D. Amodei · 2022
Later among the works it cites.
Specification gaming: the flip side of ai ingenuity
V. Krakovna et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Locatello, D. Weissenborn, T. Unterthiner, A. Mahendran, G. Heigold, J. Uszkoreit, A. Dosovitskiy, and T. Kipf · 2020
Cited alongside, same era.
Multi-object representation learning with iterative variational inference, 2020
K. Greff, R. L. Kaufman, R. Kabra, N. Watters, C. Burgess, D. Zoran, L. Matthey, M. Botvinick, and A. Lerchner · 2020
Cited alongside, same era.
Contrastive learning of structured world models, 2020
T. Kipf, E. van der Pol, and M. Welling · 2020
Cited alongside, same era.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, R. Martín-Martín, A. Joshi, S. Nasiriany, and Y. Zhu · 2020
Cited alongside, same era.
A simple framework for contrastive learning of visual representations, 2020
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton · 2020
Cited alongside, same era.
AW-opt: Learning robotic skills with imitation andreinforcement at scale
Y. Lu, K. Hausman, Y. Chebotar, M. Yan, E. Jang, A. Herzog, T. Xiao, A. Irpan, M. Khansari, D. Kalashnikov, and S. Levine · 2021
Cited alongside, same era.
Beyond pick-and-place: Tackling robotic stacking of diverse shapes
A. X. Lee, C. Devin, Y. Zhou, T. Lampe, K. Bousmalis, J. T. Springenberg, A. Byravan, A. Abdolmaleki, N. Gileadi, D. Khosid, C. Fantacci, J. E. Chen, A. S. Raju, R. Jeong, M. Neunert, A. Laurens, S. Saliceti, F. Casarini, M. A. Riedmiller, R. Hadsell, and F. Nori · 2021
Cited alongside, same era.
Temporal difference learning for model predictive control, 2022
N. Hansen, X. Wang, and H. Su · 2022
Later among the works it cites.
Curiosity-driven exploration via latent bayesian surprise, 2022
P. Mazzaglia, O. Catal, T. Verbelen, and B. Dhoedt · 2022
Later among the works it cites.
Generalization and robustness implications in object-centric learning, 2022
A. Dittadi, S. Papa, M. D. Vita, B. Schölkopf, O. Winther, and F. Locatello · 2022
Later among the works it cites.
Mastering the unsupervised reinforcement learning benchmark from pixels
S. Rajeswar, P. Mazzaglia, T. Verbelen, A. Piché, B. Dhoedt, A. Courville, and A. Lacoste · 2023
Closest in time.
Mastering diverse domains through world models
D. Hafner, J. Pasukonis, J. Ba, and T. Lillicrap · 2023
Closest in time.
Multi-view masked world models for visual robotic manipulation, 2023
Y. Seo, J. Kim, S. James, K. Lee, J. Shin, and P. Abbeel · 2023
Closest in time.
Interaction-based disentanglement of entities for object-centric world models
A. Nakano, M. Suzuki, and Y. Matsuo · 2023
Closest in time.
Maniskill2: A unified benchmark for generalizable manipulation skills, 2023
J. Gu, F. Xiang, X. Li, Z. Ling, X. Liu, T. Mu, Y. Tang, S. Tao, X. Wei, Y. Yao, X. Yuan, P. Xie, Z. Huang, R. Chen, and H. Su · 2023
Closest in time.
Segment anything, 2023
A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W.-Y. Lo, P. Dollár, and R. Girshick · 2023
Closest in time.
Slotformer: Unsupervised visual dynamics simulation with object-centric models, 2023
Z. Wu, N. Dvornik, K. Greff, T. Kipf, and A. Garg · 2023
Closest in time.
A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W.-Y. Lo, et al · 2023
Closest in time.
Y. Cheng, L. Li, Y. Xu, X. Li, Z. Yang, W. Wang, and Y. Yang · 2023
Closest in time.