Fetching the paper…
Reading the bibliography…
This paper presents a reinforcement learning method for object goal navigation (ObjNav) where an agent navigates in 3D indoor environments to reach a target object based on long-term observations of objects and scenes.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in Proceedings of The 33rd International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, M. F. Balcan and K. Q. Weinberger, Eds., vol. 48. New York, New York, USA: PMLR, 20–22 Jun 2016, pp. 1928–1937
1937
Earlier work this paper cites.
R. Pascanu, T. Mikolov, and Y. Bengio, “On the difficulty of training recurrent neural networks,” ICML , no. PART 3, pp. 2347–2355, 2013
2013
Earlier work this paper cites.
T. Mikolov, K. Chen, G. Corrado, and J. Dean, “Efficient estimation of word representations in vector space,” ICLR Workshop , pp. 1–12, 2013
2013
Earlier work this paper cites.
A. Graves, G. Wayne, and I. Danihelka, “Neural turing machines,” arXiv , vol. abs/1410.5401, 2014
2014
Earlier work this paper cites.
R. Jozefowicz, W. Zaremba, and I. Sutskever, “An empirical exploration of Recurrent Network architectures,” ICML , vol. 3, pp. 2332–2340, 2015
2015
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, and D. Hassabis, “Mastering the game of go with deep neural networks and tree search,” Nature , vol. 529, pp. 484–503, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
J. Oh, V. Chockalingam, S. Singh, and H. Lee, “Control of memory, active perception, and action in minecraft,” ICML , vol. 6, pp. 4067–4089, 2016
2016
Earlier work this paper cites.
A. Graves, G. Wayne, M. Reynolds, T. Harley, I. Danihelka, A. Grabska, S. G. Colmenarejo, E. Grefenstette, T. Ramalho, J. Agapiou, A. P. Badia, K. M. Hermann, Y. Zwols, G. Ostrovski, A. Cain, H. King, C. Summerfield, P. Blunsom, K. Kavukcuoglu, and D. Hassabis, “Hybrid computing using a neural network with dynamic external memory,” Nature , vol. 538, no. 7626, pp. 471–476, 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” CVPR , pp. 770–778, 2016
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Nips , pp. 5999–6009, 2017
2017
Earlier work this paper cites.
A. Pritzel, B. Uria, S. Srinivasan, A. P. Badia, O. Vinyals, D. Hassabis, D. Wierstra, and C. Blundell, “Neural episodic control,” ICML , vol. 6, pp. 4320–4331, 2017
2017
Cited alongside, same era.
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi, “Target-driven visual navigation in indoor scenes using deep reinforcement learning,” in ICRA , jul 2017, pp. 3357–3364
2017
Cited alongside, same era.
E. Kolve, R. Mottaghi, W. Han, E. VanderBilt, L. Weihs, A. Herrasti, D. Gordon, Y. Zhu, A. Gupta, and A. Farhadi, “AI2-THOR: An Interactive 3D Environment for Visual AI,” arXiv , 2017
2017
Cited alongside, same era.
2018
Cited alongside, same era.
R. Druon, Y. Yoshiyasu, A. Kanezaki, and A. Watt, “Visual object search by learning spatial context,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 1279–1286, 2020
2020
Later among the works it cites.
A. P. Badia, B. Piot, S. Kapturowski, P. Sprechmann, A. Vitvitskyi, Z. D. Guo, and C. Blundell, “Agent57: Outperforming the Atari human benchmark,” in ICML , ser. PMLR, H. D. III and A. Singh, Eds., vol. 119. PMLR, 13–18 Jul 2020, pp. 507–517
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
A. Khan, C. Zhang, N. Atanasov, K. Karydis, V. Kumar, and D. D. Lee, “Memory augmented control networks,” in ICLR , 2018
2018
Cited alongside, same era.
X. Ye, Z. Lin, H. Li, S. Zheng, and Y. Yang, “Active Object Perceiver: Recognition-Guided Policy Learning for Object Searching on Mobile Robots,” IROS , pp. 6857–6863, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
K. Fang, A. Toshev, L. Fei-Fei, and S. Savarese, “Scene memory transformer for embodied agents in long-horizon tasks,” CVPR , vol. 2019-June, pp. 538–547, 2019
2019
Cited alongside, same era.
M. Wortsman, K. Ehsani, M. Rastegari, A. Farhadi, and R. Mottaghi, “Learning to learn how to learn: Self-adaptive visual navigation using meta-learning,” 2019 IEEE/CVF CVPR , pp. 6743–6752, 2019
2019
Cited alongside, same era.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in ECCV , 2020, pp. 213–229
2020
Later among the works it cites.
G. Zhu*, Z. Lin*, G. Yang, and C. Zhang, “Episodic reinforcement learning with associative memory,” in ICLR , 2020
2020
Later among the works it cites.
H. Du, X. Yu, and L. Zheng, “Learning Object Relation Graph and Tentative Policy for Visual Navigation,” ECCV , pp. 19–34, 2020
2020
Later among the works it cites.
H. Du, X. Yu, and L. Zheng, “VTNet: Visual transformer network for object goal navigation,” in ICLR , 2021
2021
Later among the works it cites.
B. Mayo, T. Hazan, and A. Tal, “Visual navigation with spatial attention,” in CVPR , 2021
2021
Later among the works it cites.
S. Khan, M. Naseer, M. Hayat, S. W. Zamir, F. S. Khan, and M. Shah, “Transformers in vision: A survey,” 2021
2021
Later among the works it cites.