Fetching the paper…
Reading the bibliography…
Planning at a higher level of abstraction instead of low level torques improves the sample efficiency in reinforcement learning, and computational efficiency in classical planning.
STRIPS: A new approach to the application of theorem proving to problem solving
R. E. Fikes and N. J. Nilsson · 1971
Earlier work this paper cites.
Forward models: Supervised learning with a distal teacher
M. I. Jordan and D. E. Rumelhart · 1992
Earlier work this paper cites.
Feudal reinforcement learning
P. Dayan and G. E. Hinton · 1993
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
R. S. Sutton, D. Precup, and S. Singh · 1999
Earlier work this paper cites.
Recent advances in hierarchical reinforcement learning
A. G. Barto and S. Mahadevan · 2003
Earlier work this paper cites.
Probabilistic robotics
S. Thrun, W. Burgard, and D. Fox · 2005
Earlier work this paper cites.
Planning Algorithms
S. M. LaValle · 2006
Earlier work this paper cites.
Dynamic movement primitives-a framework for motor control in humans and humanoid robotics
S. Schaal · 2006
Earlier work this paper cites.
Using motion primitives in probabilistic sample-based planning for humanoid robots
K. Hauser, T. Bretl, K. Harada, and J.-C. Latombe · 2008
Earlier work this paper cites.
Robot programming by demonstration
A. Billard, S. Calinon, R. Dillmann, and S. Schaal · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2009
Earlier work this paper cites.
Dynamical movement primitives: learning attractor models for motor behaviors
A. J. Ijspeert, J. Nakanishi, H. Hoffmann, P. Pastor, and S. Schaal · 2013
Earlier work this paper cites.
People watching: Human actions as a cue for single view geometry
D. F. Fouhey, V. Delaitre, A. Gupta, A. A. Efros, I. Laptev, and J. Sivic · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Cad2rl: Real single-image flight without a single real image
F. Sadeghi and S. Levine · 2016
Cited alongside, same era.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Cited alongside, same era.
Learning to poke by poking: Experiential learning of intuitive physics
P. Agrawal, A. V. Nair, P. Abbeel, J. Malik, and S. Levine · 2016
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
E. Jang, S. Gu, and B. Poole · 2016
Cited alongside, same era.
3D semantic parsing of large-scale indoor spaces
I. Armeni, O. Sener, A. R. Zamir, H. Jiang, I. Brilakis, M. Fischer, and S. Savarese · 2016
Cited alongside, same era.
Matterport3D: Learning from RGB-D data in indoor environments
A. Chang, A. Dai, T. Funkhouser, M. Halber, M. Niessner, M. Savva, S. Song, A. Zeng, and Y. Zhang · 2017
Later among the works it cites.
Sfv: Reinforcement learning of physical skills from videos
X. B. Peng, A. Kanazawa, J. Malik, P. Abbeel, and S. Levine · 2018
Later among the works it cites.
Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen · 2018
Later among the works it cites.
Playing hard exploration games by watching youtube
Y. Aytar, T. Pfaff, D. Budden, T. L. Paine, Z. Wang, and N. de Freitas · 2018
Later among the works it cites.
Behavioral cloning from observation
F. Torabi, G. Warnell, and P. Stone · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Xu, Y. Gao, F. Yu, and T. Darrell · 2017
Cited alongside, same era.
Target-driven visual navigation in indoor scenes using deep reinforcement learning
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi · 2017
Cited alongside, same era.
Learning to navigate in complex environments
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu, et al · 2017
Cited alongside, same era.
Cognitive mapping and planning for visual navigation
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik · 2017
Cited alongside, same era.
Curiosity-driven exploration by self-supervised prediction
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell · 2017
Cited alongside, same era.
Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets
K. Hausman, Y. Chebotar, S. Schaal, G. Sukhatme, and J. J. Lim · 2017
Cited alongside, same era.
One-shot visual imitation learning via meta-learning
C. Finn, T. Yu, T. Zhang, P. Abbeel, and S. Levine · 2017
Cited alongside, same era.
A. D. Edwards, H. Sahni, Y. Schroeker, and C. L. Isbell · 2018
Later among the works it cites.
Zero-shot visual imitation
D. Pathak, P. Mahmoudieh, G. Luo, P. Agrawal, D. Chen, Y. Shentu, E. Shelhamer, J. Malik, A. A. Efros, and T. Darrell · 2018
Later among the works it cites.
One-shot imitation from observing humans via domain-adaptive meta-learning
T. Yu, C. Finn, A. Xie, S. Dasari, T. Zhang, P. Abbeel, and S. Levine · 2018
Later among the works it cites.
Visual memory for robust path following
A. Kumar*, S. Gupta*, D. Fouhey, S. Levine, and J. Malik · 2018
Later among the works it cites.
Deep visual teach and repeat on path networks
T. Swedish and R. Raskar · 2018
Later among the works it cites.
On evaluation of embodied navigation agents
P. Anderson, A. Chang, D. S. Chaplot, A. Dosovitskiy, S. Gupta, V. Koltun, J. Kosecka, J. Malik, R. Mottaghi, M. Savva, and A. Zamir · 2018
Later among the works it cites.
Divis: Domain invariant visual servoing for collision-free goal reaching
F. Sadeghi · 2019
Closest in time.
Diversity is all you need: Learning skills without a reward function
B. Eysenbach, A. Gupta, J. Ibarz, and S. Levine · 2019
Closest in time.
Pyrobot: An open-source robotics framework for research and benchmarking
A. Murali, T. Chen, K. V. Alwala, D. Gandhi, L. Pinto, S. Gupta, and A. Gupta · 2019
Closest in time.