Fetching the paper…
Reading the bibliography…
A generalist robot equipped with learned skills must be able to perform many tasks in many different environments.
J. Gibson,
1979
Earlier work this paper cites.
L. P. Kaelbling, “Learning to achieve goals,” in
1993
Earlier work this paper cites.
H. Benbrahim and J. A. Franklin, “Biped dynamic walking using reinforcement learning,”
1997
Earlier work this paper cites.
N. Kohl and P. Stone, “Machine Learning for Fast Quadrupedal Locomotion,” in
2004
Earlier work this paper cites.
N. Chentanez, A. G. Barto, and S. P. Singh, “Intrinsically motivated reinforcement learning,” in
2005
Earlier work this paper cites.
J. Kober and J. Peter, “Policy search for motor primitives in robotics,” in
2008
Earlier work this paper cites.
S. Hart and R. Grupen, “Learning Generalizable Control Programs,” in
2010
Earlier work this paper cites.
J. Peters, K. Mülling, and Y. Altün, “Relative Entropy Policy Search,” in
2010
Earlier work this paper cites.
M. P. Deisenroth and C. E. Rasmussen, “PILCO: A model-based and data-efficient approach to policy search,” in
2011
Earlier work this paper cites.
A. Baranes and P.-Y. Oudeyer, “Active Learning of Inverse Models with Intrinsically Motivated Goal Exploration in Robots,”
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in
2012
Earlier work this paper cites.
M. Lopes, T. Lang, M. Toussaint, and P.-Y. Oudeyer, “Exploration in model-based reinforcement learning by empirically estimating learning progress,” in
2012
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing Atari with Deep Reinforcement Learning,” in
2013
Earlier work this paper cites.
D. Abel, G. Barth-Maron, J. Macglashan, and S. Tellex, “Toward Affordance-Aware Planning,” in
2014
Earlier work this paper cites.
K. Berger,
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative Adversarial Nets,” in
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-Encoding Variational Bayes,” in
2014
Earlier work this paper cites.
T. Schaul, D. Horgan, K. Gregor, and D. Silver, “Universal Value Function Approximators,” in
2015
Earlier work this paper cites.
K. Sohn, X. Yan, and H. Lee, “Learning Structured Output Representation using Deep Conditional Generative Models,” in
2015
Cited alongside, same era.
P. Agrawal, A. Nair, P. Abbeel, J. Malik, and S. Levine, “Learning to Poke by Poking: Experiential Learning of Intuitive Physics,” in
2016
Cited alongside, same era.
M. Bellemare, S. Srinivasan, G. Ostrovski, T. Schaul, D. Saxton, and R. Munos, “Unifying count-based exploration and intrinsic motivation,” in
2016
Cited alongside, same era.
Y. Duan, J. Schulman, X. Chen, P. L. Bartlett, I. Sutskever, and P. Abbeel, “RL$
2016
Cited alongside, same era.
R. Houthooft, X. Chen, Y. Duan, J. Schulman, F. De Turck, and P. Abbeel, “Variational Information Maximizing Exploration,” in
2016
Cited alongside, same era.
M. Hassanin, S. Khan, and M. Tahtali, “Visual Affordance and Function Understanding: A Survey, Tech. Rep. 1, 2018
2018
Later among the works it cites.
D. Held, X. Geng, C. Florensa, and P. Abbeel, “Automatic Goal Generation for Reinforcement Learning Agents,” in
2018
Later among the works it cites.
O. Nachum, S. S. Gu, H. Lee, and S. Levine, “Data-Efficient Hierarchical Reinforcement Learning,” in
2018
Later among the works it cites.
A. Nair, V. Pong, M. Dalal, S. Bahl, S. Lin, and S. Levine, “Visual Reinforcement Learning with Imagined Goals,” in
2018
Later among the works it cites.
A. Péré, S. Forestier, O. Sigaud, and P.-Y. Oudeyer, “Unsupervised Learning of Goal Spaces for Intrinsically Motivated Goal Exploration,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-End Training of Deep Visuomotor Policies,”
2016
Cited alongside, same era.
H. Min, A. Yi, R. Luo, J. Zhu, and S. Bi, “Affordance Research in Developmental Robotics: A Survey,”
2016
Cited alongside, same era.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, and D. Hassabis, “Mastering the game of Go with deep neural networks and tree search,”
2016
Cited alongside, same era.
B. C. Stadie, S. Levine, and P. Abbeel, “Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models,” in
2016
Cited alongside, same era.
A. van den Oord, N. Kalchbrenner, O. Vinyals, L. Espeholt, A. Graves, and K. Kavukcuoglu, “Conditional Image Generation with PixelCNN Decoders,” in
2016
Cited alongside, same era.
M. Andrychowicz, F. Wolski, A. Ray, J. Schneider, R. Fong, P. Welinder, B. Mcgrew, J. Tobin, P. Abbeel, and W. Zaremba, “Hindsight Experience Replay,” in
2017
Cited alongside, same era.
J. Donahue, P. Krähenbühl, and T. Darrell, “Adversarial Feature Learning,” in
2017
Cited alongside, same era.
N. Yamanobe, W. Wan, I. G. Ramirez-Alpizar, D. Petit, T. Tsuji, S. Akizuki, M. Hashimoto, K. Nagata, and K. Harada, “A brief review of affordance in robotic manipulation research,”
2018
Later among the works it cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding,” in
2019
Later among the works it cites.
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, and P. Sermanet, “Learning Latent Plans from Play,” in
2019
Later among the works it cites.
A. Nair, S. Bahl, A. Khazatsky, V. Pong, G. Berseth, and S. Levine, “Contextual Imagined Goals for Self-Supervised Robotic Learning,” in
2019
Later among the works it cites.
K. Rakelly, A. Zhou, D. Quillen, C. Finn, and S. Levine, “Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables,” in
2019
Later among the works it cites.
D. Warde-Farley, T. Van De Wiele, T. Kulkarni, C. Ionescu, S. Hansen, and M. Volodymyr, “Unsupervised Control Through Non-Parametric Discriminative Rewards,” in
2019
Later among the works it cites.
K. Khetarpal, Z. Ahmed, G. Comanici, D. Abel, and D. Precup, “What can I do here? A Theory of Affordances in Reinforcement Learning,” in
2020
Later among the works it cites.
A. Nair, M. Dalal, A. Gupta, and S. Levine, “Accelerating Online Reinforcement Learning with Offline Datasets,” jun 2020
2020
Later among the works it cites.
V. H. Pong, M. Dalal, S. Lin, A. Nair, S. Bahl, and S. Levine, “Skew-Fit: State-Covering Self-Supervised Reinforcement Learning,” in
2020
Later among the works it cites.
2020
Later among the works it cites.
C. Colas, T. Karch, O. Sigaud, and P.-Y. Oudeyer, “Intrinsically Motivated Goal-Conditioned Reinforcement Learning: a Short Survey,” Tech. Rep., 2021
2021
Closest in time.
E. Coumans and Y. Bai, “Pybullet, a python module for physics simulation for games, robotics and machine learning,”
2021
Closest in time.
D. Xu, A. Mandlekar, R. Martín-Martín, Y. Zhu, S. Savarese, and L. Fei-Fei, “Deep Affordance Foresight: Planning Through What Can Be Done in the Future,” in
2021
Closest in time.