Fetching the paper…
Reading the bibliography…
Long-term planning poses a major difficulty to many reinforcement learning algorithms.
The ff planning system: Fast plan generation through heuristic search
J. Hoffmann and B. Nebel · 2001
Earlier work this paper cites.
The metric-ff planning system: Translating“ignoring delete lists”to numeric state variables
J. Hoffmann · 2003
Earlier work this paper cites.
Hierarchical task and motion planning in the now
L. P. Kaelbling and T. Lozano-Pérez · 2011
Earlier work this paper cites.
Moveit![ros topics]
S. Chitta, I. Sucan, and S. Cousins · 2012
Earlier work this paper cites.
The open motion planning library
I. A. Sucan, M. Moll, and L. E. Kavraki · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Learning phrase representations using rnn encoder–decoder for statistical machine translation
K. Cho, B. van Merriënboer, Ç. Gülçehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
Policy search for multi-robot coordination under uncertainty
C. Amato, G. Konidaris, A. Anders, G. Cruz, J. P. How, and L. P. Kaelbling · 2016
Earlier work this paper cites.
C. Beattie, J. Z. Leibo, D. Teplyashin, T. Ward, M. Wainwright, H. Küttler, A. Lefrancq, S. Green, V. Valdés, A. Sadik, et al · 2016
Earlier work this paper cites.
Openai gym, 2016
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Earlier work this paper cites.
Guided search for task and motion plans using learned heuristics
R. Chitnis, D. Hadfield-Menell, A. Gupta, S. Srivastava, E. Groshev, C. Lin, and P. Abbeel · 2016
Earlier work this paper cites.
Vizdoom: A doom-based ai research platform for visual reinforcement learning
M. Kempka, M. Wydmuch, G. Runc, J. Toczek, and W. Jaśkowski · 2016
Earlier work this paper cites.
Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation
T. D. Kulkarni, K. Narasimhan, A. Saeedi, and J. Tenenbaum · 2016
Earlier work this paper cites.
Deeper depth prediction with fully convolutional residual networks
I. Laina, C. Rupprecht, V. Belagiannis, F. Tombari, and N. Navab · 2016
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Cognitive mapping and planning for visual navigation
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik · 2017
Cited alongside, same era.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Cited alongside, same era.
AI2-THOR: An Interactive 3D Environment for Visual AI
E. Kolve, R. Mottaghi, D. Gordon, Y. Zhu, A. Gupta, and A. Farhadi · 2017
Cited alongside, same era.
Learning to navigate in complex environments
P. W. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. J. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu, D. Kumaran, and R. Hadsell · 2017
Vision-and-Language Navigation: Interpreting visually-grounded navigation instructions in real environments
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. van den Hengel · 2018
Later among the works it cites.
Hyp-despot: A hybrid parallel algorithm for online planning under uncertainty
P. Cai, Y. Luo, D. Hsu, and W. S. Lee · 2018
Later among the works it cites.
Gated-attention architectures for task-oriented language grounding
D. S. Chaplot, K. M. Sathyendra, R. K. Pasumarthi, D. Rajagopal, and R. Salakhutdinov · 2018
Later among the works it cites.
Embodied question answering
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra · 2018
Later among the works it cites.
Neural Modular Control for Embodied Question Answering
A. Das, G. Gkioxari, S. Lee, D. Parikh, and D. Batra · 2018
Later among the works it cites.
Iqa: Visual question answering in interactive environments
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Curiosity-driven exploration by self-supervised prediction
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell · 2017
Cited alongside, same era.
Mastering the game of go without human knowledge
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al · 2017
Cited alongside, same era.
A deep hierarchical approach to lifelong learning in minecraft
C. Tessler, S. Givony, T. Zahavy, D. J. Mankowitz, and S. Mannor · 2017
Cited alongside, same era.
Visual semantic planning using deep successor representations
Y. Zhu, D. Gordon, E. Kolve, D. Fox, L. Fei-Fei, A. Gupta, R. Mottaghi, and A. Farhadi · 2017
Cited alongside, same era.
Target-driven visual navigation in indoor scenes using deep reinforcement learning
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi · 2017
Cited alongside, same era.
On evaluation of embodied navigation agents
P. Anderson, A. Chang, D. S. Chaplot, A. Dosovitskiy, S. Gupta, V. Koltun, J. Kosecka, J. Malik, R. Mottaghi, M. Savva, and A. R. Zamir · 2018
Cited alongside, same era.
D. Gordon, A. Kembhavi, M. Rastegari, J. Redmon, D. Fox, and A. Farhadi · 2018
Later among the works it cites.
From skills to symbols: Learning symbolic representations for abstract high-level planning
G. Konidaris, L. P. Kaelbling, and T. Lozano-Perez · 2018
Later among the works it cites.
Fusion++: Volumetric object-level slam
J. McCormac, R. Clark, M. Bloesch, A. Davison, and S. Leutenegger · 2018
Later among the works it cites.
Shifting the baseline: Single modality performance on visual navigation & qa
J. Thomason, D. Gordon, and Y. Bisk · 2018
Later among the works it cites.
Building generalizable agents with a realistic and rich 3d environment
Y. Wu, Y. Wu, G. Gkioxari, and Y. Tian · 2018
Later among the works it cites.
Gibson env: Real-world perception for embodied agents
F. Xia, A. R. Zamir, Z. He, A. Sax, J. Malik, and S. Savarese · 2018
Later among the works it cites.
Chalet: Cornell house agent learning environment
C. Yan, D. Misra, A. Bennnett, A. Walsman, Y. Bisk, and Y. Artzi · 2018
Later among the works it cites.