Fetching the paper…
Reading the bibliography…
"Embodied visual navigation" problem requires an agent to navigate in a 3D environment mainly rely on its first-person observation.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Harley, T. P. Lillicrap, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in ICML , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
H. Sakoe and S. Chiba, “Dynamic programming algorithm optimization for spoken word recognition,” IEEE transactions on acoustics, speech, and signal processing , vol. 26, no. 1, pp. 43–49, 1978
1978
Earlier work this paper cites.
J. Thomason, D. Gordon, and Y. Bisk, “Shifting the baseline: Single modality performance on visual navigation & qa,” in NAACL , 2019, pp. 1977–1983
1983
Earlier work this paper cites.
M. Kabuka and A. Arenas, “Position verification of a mobile robot using standard pattern,” ICRA , vol. 3, no. 6, pp. 505–516, 1987
1987
Earlier work this paper cites.
C. Thorpe, M. H. Hebert, T. Kanade, and S. A. Shafer, “Vision and navigation for the carnegie-mellon navlab,” TPAMI , vol. 10, no. 3, pp. 362–373, 1988
1988
Earlier work this paper cites.
J. Borenstein and Y. Koren, “Real-time obstacle avoidance for fast mobile robots in cluttered environments,” in ICRA . IEEE, 1990, pp. 572–577
1990
Earlier work this paper cites.
A. Kosaka and A. Kak, “Fast vision-guided mobile robot navigation using model-based reasoning and prediction of uncertainties,” in IROS , vol. 3, 1992, pp. 2177–2186
1992
Earlier work this paper cites.
P. Dayan and G. E. Hinton, “Feudal reinforcement learning,” in NeurIPS , vol. 5, 1992, pp. 271–278
1992
Earlier work this paper cites.
M. L. Littman, “Markov games as a framework for multi-agent reinforcement learning,” in ICML , 1994, pp. 157–163
1994
Earlier work this paper cites.
D. J. Berndt and J. Clifford, “Using dynamic time warping to find patterns in time series,” in AAAIWS’94 Proceedings of the 3rd International Conference on Knowledge Discovery and Data Mining , 1994, pp. 359–370
1994
Earlier work this paper cites.
M. Tan, “Multi-agent reinforcement learning: independent vs. cooperative agents,” in ICML , 1997, pp. 487–494
1997
Earlier work this paper cites.
H. Haddad, M. Khatib, S. Lacroix, and R. Chatila, “Reactive navigation in outdoor environments using potential fields,” in IROS , vol. 2, 1998, pp. 1232–1237
1998
Earlier work this paper cites.
A. J. Davison and D. W. Murray, “Mobile robot localisation using active vision,” in ECCV . Springer, 1998, pp. 809–825
1998
Earlier work this paper cites.
S. Thrun, “Probabilistic algorithms in robotics,” Ai Magazine , vol. 21, no. 4, pp. 93–109, 2000
2000
Earlier work this paper cites.
E. J. Keogh and M. J. Pazzani, “Scaling up dynamic time warping for datamining applications,” in KDD , 2000, pp. 285–289
2000
Earlier work this paper cites.
P. Dourish, Where the Action Is: The Foundations of Embodied Interaction , 2001
2001
Earlier work this paper cites.
G. N. DeSouza and A. C. Kak, “Vision for mobile robot navigation: A survey,” TPAMI , vol. 24, no. 2, pp. 237–267, 2002
2002
Earlier work this paper cites.
S. Se, D. G. Lowe, and J. J. Little, “Mobile robot localization and mapping with uncertainty using scale-invariant visual landmarks,” The International Journal of Robotics Research , vol. 21, no. 8, pp. 735–758, 2002
2002
Earlier work this paper cites.
S. Thrun, “Probabilistic robotics,” Communications of the ACM , vol. 45, no. 3, pp. 52–57, 2002
2002
Earlier work this paper cites.
S. Lenser and M. Veloso, “Visual sonar: fast obstacle avoidance using monocular vision,” in IROS , vol. 1, 2003, pp. 886–891
2003
Earlier work this paper cites.
C. F. Olson, L. H. Matthies, M. Schoppers, and M. W. Maimone, “Rover navigation using stereo ego-motion,” Robotics and Autonomous Systems , vol. 43, no. 4, pp. 215–229, 2003
2003
Earlier work this paper cites.
Davison, “Real-time simultaneous localisation and mapping with a single camera,” in ICCV , vol. 2, 2003, pp. 1403–1410
2003
Earlier work this paper cites.
A. G. Barto and S. Mahadevan, “Recent advances in hierarchical reinforcement learning,” Discrete Event Dynamic Systems , vol. 13, no. 1, pp. 41–77, 2003
2003
Earlier work this paper cites.
A. Remazeilles, F. Chaumette, and P. Gros, “Robot motion control from a visual memory,” in ICRA , vol. 5, 2004, pp. 4695–4700
2004
Earlier work this paper cites.
S. Thrun, Probabilistic Robotics , 2005
2005
Earlier work this paper cites.
E. Royer, J. Bom, M. Dhome, B. Thuilot, M. Lhuillier, and F. Marmoiton, “Outdoor autonomous navigation using monocular vision,” in IROS , 2005, pp. 1253–1258
2005
Earlier work this paper cites.
L. D. Jackel, E. Krotkov, M. Perschbacher, J. Pippine, and C. Sullivan, “The darpa lagr program: Goals, challenges, methodology, and phase i results,” Journal of Field robotics , vol. 23, no. 11-12, pp. 945–973, 2006
2006
Earlier work this paper cites.
H. Durrant-Whyte and T. Bailey, “Simultaneous localization and mapping: part i,” IEEE robotics & automation magazine , vol. 13, no. 2, pp. 99–110, 2006
2006
Earlier work this paper cites.
U. Muller, J. Ben, E. Cosatto, B. Flepp, and Y. L. Cun, “Off-road obstacle avoidance through end-to-end learning,” in NeurIPS . Citeseer, 2006, pp. 739–746
2006
Earlier work this paper cites.
S. M. LaValle, Planning algorithms . Cambridge university press, 2006
2006
Earlier work this paper cites.
L. Jeni, Z. Istenes, P. Korondi, and H. Hashimoto, “Hierarchical reinforcement learning for robot navigation using the intelligent space concept,” in 2007 11th International Conference on Intelligent Engineering Systems . IEEE, 2007, pp. 149–153
2007
Earlier work this paper cites.
R. Hadsell, P. Sermanet, J. Ben, A. Erkan, M. Scoffier, K. Kavukcuoglu, U. Muller, and Y. LeCun, “Learning long-range vision for autonomous off-road driving,” Journal of Field Robotics , vol. 26, no. 2, pp. 120–144, 2009
2009
Earlier work this paper cites.
T. Jaksch, R. Ortner, and P. Auer, “Near-optimal regret bounds for reinforcement learning,” JMLR , vol. 11, no. 51, pp. 1563–1600, 2010
2010
Earlier work this paper cites.
M. Fisher, D. Ritchie, M. Savva, T. Funkhouser, and P. Hanrahan, “Example-based synthesis of 3d object arrangements,” international conference on computer graphics and interactive techniques , vol. 31, no. 6, p. 135, 2012
2012
Earlier work this paper cites.
A. Vakanski, I. Mantegh, A. Irish, and F. Janabi-Sharifi, “Trajectory learning for robot programming by demonstration using hidden markov model and dynamic time warping,” IEEE Transactions on Systems, Man, and Cybernetics , vol. 42, no. 4, pp. 1039–1052, 2012
2012
Earlier work this paper cites.
A. Geiger, P. Lenz, and R. Urtasun, “Are we ready for autonomous driving? the kitti vision benchmark suite,” in CVPR . IEEE, 2012, pp. 3354–3361
2012
Earlier work this paper cites.
N. Malone and V. Kypraios, “Dd t07 - designing a situational judgment test to evaluate lll-structured problem solving in a virtual learning environment,” 2012
2012
Earlier work this paper cites.
T. Kruse, A. K. Pandey, R. Alami, and A. Kirsch, “Human-aware robot navigation: A survey,” Robotics and Autonomous Systems , vol. 61, no. 12, pp. 1726–1743, 2013
2013
Earlier work this paper cites.
S. Ross, N. Melik-Barkhudarov, K. S. Shankar, A. Wendel, D. Dey, J. A. Bagnell, and M. Hebert, “Learning monocular reactive uav control in cluttered natural environments,” in ICRA . IEEE, 2013, pp. 1765–1772
2013
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. J. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in ECCV , 2014, pp. 740–755
2014
Earlier work this paper cites.
A. Khosla, B. An, J. J. Lim, and A. Torralba, “Looking beyond the visible scene,” in CVPR , 2014, pp. 3710–3717
2014
Earlier work this paper cites.
J. Engel, T. Schöps, and D. Cremers, “Lsd-slam: Large-scale direct monocular slam,” in ECCV . Springer, 2014, pp. 834–849
2014
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” NeurIPS , vol. 27, pp. 3104–3112, 2014
2014
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,” The JMLR , vol. 15, no. 1, pp. 1929–1958, 2014
2014
Earlier work this paper cites.
S. Tellex, R. A. Knepper, A. Li, D. Rus, and N. Roy, “Asking for help using inverse semantics,” in Robotics: Science and Systems 2014 , vol. 10, 2014
2014
Earlier work this paper cites.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep learning,” Nature , vol. 521, no. 7553, pp. 436–444, 2015
2015
Earlier work this paper cites.
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh, “Vqa: Visual question answering,” in ICCV , 2015, pp. 2425–2433
2015
Earlier work this paper cites.
C. Doersch, A. Gupta, and A. A. Efros, “Unsupervised visual representation learning by context prediction,” in ICCV , 2015, pp. 1422–1430
2015
Earlier work this paper cites.
T. Schaul, D. Horgan, K. Gregor, and D. Silver, “Universal value function approximators,” in ICML , 2015, pp. 1312–1320
2015
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in ICLR , 2015
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
C. Xie, S. Patil, T. Moldovan, S. Levine, and P. Abbeel, “Model-based reinforcement learning with parametrized physical models and optimism-driven exploration,” in ICRA , 2016, pp. 504–511
2016
Earlier work this paper cites.
A. Handa, V. Patraucean, S. Stent, and R. Cipolla, “Scenenet: An annotated model generator for indoor scene understanding,” in ICRA , 2016, pp. 5737–5743
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola, “Stacked attention networks for image question answering,” in CVPR , 2016, pp. 21–29
2016
Earlier work this paper cites.
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” in ICLR , 2016
2016
Earlier work this paper cites.
A. Mousavian, H. Pirsiavash, and J. Kosecka, “Joint semantic segmentation and depth estimation with deep convolutional networks,” in 3DV , 2016, pp. 611–619
2016
Earlier work this paper cites.
M. Noroozi and P. Favaro, “Unsupervised learning of visual representations by solving jigsaw puzzles,” in ECCV , 2016, pp. 69–84
2016
Earlier work this paper cites.
R. Zhang, P. Isola, and A. A. Efros, “Colorful image colorization,” in ECCV , 2016, pp. 649–666
2016
Earlier work this paper cites.
A. Dosovitskiy and V. Koltun, “Learning to act by predicting the future,” in ICLR , 2016
2016
Earlier work this paper cites.
M. Kempka, M. Wydmuch, G. Runc, J. Toczek, and W. Jaśkowski, “Vizdoom: A doom-based ai research platform for visual reinforcement learning,” in CIG . IEEE, 2016, pp. 1–8
2016
Earlier work this paper cites.
M. Burri, J. Nikolic, P. Gohl, T. Schneider, J. Rehder, S. Omari, M. W. Achtelik, and R. Siegwart, “The euroc micro aerial vehicle datasets,” The International Journal of Robotics Research , vol. 35, no. 10, pp. 1157–1163, 2016
2016
Earlier work this paper cites.
R. Jonschkowski and O. Brock, “End-to-end learnable histogram filters,” 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
B. Shi, X. Wang, P. Lyu, C. Yao, and X. Bai, “Robust scene text recognition with automatic rectification,” in CVPR , 2016, pp. 4168–4176
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
W. Zhang, K. Liu, W. Zhang, Y. Zhang, and J. Gu, “Deep neural networks for wireless localization in indoor and outdoor environments,” Neurocomputing , vol. 194, pp. 279–287, 2016
2016
Earlier work this paper cites.
A. A. Rusu, S. G. Colmenarejo, C. Gulcehre, G. Desjardins, J. Kirkpatrick, R. Pascanu, V. Mnih, K. Kavukcuoglu, and R. Hadsell, “Policy distillation,” in ICLR , 2016
2016
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” Communications of The ACM , vol. 60, no. 6, pp. 84–90, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi, “Target-driven visual navigation in indoor scenes using deep reinforcement learning,” in ICRA , 2017, pp. 3357–3364
2017
Earlier work this paper cites.
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu, “Reinforcement learning with unsupervised auxiliary tasks,” in ICLR 2017 : ICLR 2017 , 2017
2017
Earlier work this paper cites.
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu, D. Kumaran, and R. Hadsell, “Learning to navigate in complex environments,” in ICLR , 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Song, F. Yu, A. Zeng, A. X. Chang, M. Savva, and T. Funkhouser, “Semantic scene completion from a single depth image,” in CVPR , 2017, pp. 190–198
2017
Earlier work this paper cites.
A. Chang, A. Dai, T. Funkhouser, M. Halber, M. Niebner, M. Savva, S. Song, A. Zeng, and Y. Zhang, “Matterport3d: Learning from rgb-d data in indoor environments,” in 3DV , 2017, pp. 667–676
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma, M. S. Bernstein, and L. Fei-Fei, “Visual genome: Connecting language and vision using crowdsourced dense image annotations,” IJCV , vol. 123, no. 1, pp. 32–73, 2017
2017
Cited alongside, same era.
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik, “Cognitive mapping and planning for visual navigation,” in CVPR , 2017, pp. 7272–7281
2017
Cited alongside, same era.
D. K. Misra, J. Langford, and Y. Artzi, “Mapping instructions and visual observations to actions with reinforcement learning.” in EMNLP , 2017, pp. 1004–1015
2017
Cited alongside, same era.
H. Tan and M. Bansal, “Lxmert: Learning cross-modality encoder representations from transformers,” in EMNLP , 2019, pp. 5099–5110
2019
Later among the works it cites.
X. Li, C. Li, Q. Xia, Y. Bisk, A. Celikyilmaz, J. Gao, N. A. Smith, and Y. Choi, “Robust navigation with language pretraining and stochastic sampling,” in EMNLP , 2019, pp. 1494–1499
2019
Later among the works it cites.
E. Wijmans, S. Datta, O. Maksymets, A. Das, G. Gkioxari, S. Lee, I. Essa, D. Parikh, and D. Batra, “Embodied Question Answering in Photorealistic Environments with Point Cloud Perception,” in CVPR , 2019
2019
Later among the works it cites.
H. Luo, G. Lin, Z. Liu, F. Liu, Z. Tang, and Y. Yao, “Segeqa: Video segmentation based visual attention for embodied question answering,” in ICCV , 2019, pp. 9667–9676
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Tamar, Y. Wu, G. Thomas, S. Levine, and P. Abbeel, “Value iteration networks,” in IJCAI , 2017, pp. 4949–4953
2017
Cited alongside, same era.
Y. Zhu, D. Gordon, E. Kolve, D. Fox, L. Fei-Fei, A. Gupta, R. Mottaghi, and A. Farhadi, “Visual semantic planning using deep successor representations,” in ICCV , 2017, pp. 483–492
2017
Cited alongside, same era.
S. Brahmbhatt and J. Hays, “Deepnav: Learning to navigate large cities,” in CVPR , 2017, pp. 3087–3096
2017
Cited alongside, same era.
G.-H. Liu, A. Siravuru, S. P. Selvaraj, M. M. Veloso, and G. Kantor, “Learning end-to-end multimodal sensor policies for autonomous navigation.” CoRL , pp. 249–261, 2017
2017
Cited alongside, same era.
J. Engel, V. Koltun, and D. Cremers, “Direct sparse odometry,” TPAMI , vol. 40, no. 3, pp. 611–625, 2017
2017
Cited alongside, same era.
R. Mur-Artal and J. D. Tardós, “Orb-slam2: An open-source slam system for monocular, stereo, and rgb-d cameras,” IEEE Transactions on Robotics , vol. 33, no. 5, pp. 1255–1262, 2017
2017
Cited alongside, same era.
K. Tateno, F. Tombari, I. Laina, and N. Navab, “Cnn-slam: Real-time dense monocular slam with learned depth prediction,” in CVPR , 2017, pp. 6243–6252
2017
Cited alongside, same era.
E. Parisotto and R. Salakhutdinov, “Neural map: Structured memory for deep reinforcement learning.” in ICLR , 2017
2017
Cited alongside, same era.
2019
Later among the works it cites.
C. Cangea, E. Belilovsky, P. Liò, and A. C. Courville, “Videonavqa: Bridging the gap between visual and embodied question answering.” in ViGIL@NeurIPS , 2019, p. 280
2019
Later among the works it cites.
J. Thomason, A. Padmakumar, J. Sinapov, N. Walker, Y. Jiang, H. Yedidsion, J. Hart, P. Stone, and R. J. Mooney, “Improving grounded natural language understanding through human-robot dialog,” in ICRA . IEEE, 2019, pp. 6934–6941
2019
Later among the works it cites.
K. Nguyen, D. Dey, C. Brockett, and B. Dolan, “Vision-based navigation with language-based assistance via imitation learning with indirect intervention,” in CVPR , 2019, pp. 12 527–12 537
2019
Later among the works it cites.
U. Jain, L. Weihs, E. Kolve, M. Rastegari, S. Lazebnik, A. Farhadi, A. G. Schwing, and A. Kembhavi, “Two body problem: Collaborative visual task completion,” in CVPR , 2019, pp. 6689–6699
2019
Later among the works it cites.
M. Marge, S. Nogar, C. J. Hayes, S. M. Lukin, J. Bloecker, E. Holder, and C. R. Voss, “A research platform for multi-robot dialogue with humans,” in Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics (Demonstrations) , 2019, pp. 132–137
2019
Later among the works it cites.
B. Liu, L. Wang, and M. Liu, “Lifelong federated reinforcement learning: A learning architecture for navigation in cloud robotic systems,” in IEEE Robotics and Automation Letters , vol. 4, no. 4, 2019, pp. 4555–4562
2019
Later among the works it cites.
D. Gordon, A. Kadian, D. Parikh, J. Hoffman, and D. Batra, “Splitnet: Sim2sim and task2task transfer for embodied visual navigation,” in ICCV , 2019, pp. 1022–1031
2019
Later among the works it cites.
W. Yuan, K. Hang, D. Kragic, M. Y. Wang, and J. A. Stork, “End-to-end nonprehensile rearrangement with deep reinforcement learning and simulation-to-reality transfer,” Robotics and Autonomous Systems , vol. 119, pp. 119–134, 2019
2019
Later among the works it cites.
F. Zhu, L. Zhu, and Y. Yang, “Sim-real joint reinforcement transfer for 3d indoor navigation,” in CVPR , 2019, pp. 11 388–11 397
2019
Later among the works it cites.
B. Qin, Y. Gao, and Y. Bai, “Sim-to-real: Six-legged robot control with deep reinforcement learning and curriculum learning,” in ICRAE , 2019
2019
Later among the works it cites.
Z. Yang, Z. Dai, Y. Yang, J. G. Carbonell, R. Salakhutdinov, and Q. V. Le, “Xlnet: Generalized autoregressive pretraining for language understanding,” in NeurIPS , vol. 32, 2019, pp. 5753–5763
2019
Later among the works it cites.
J. Lee, W. Yoon, S. Kim, D. Kim, S. Kim, C. H. So, and J. Kang, “Biobert: a pre-trained biomedical language representation model for biomedical text mining.” Bioinformatics , vol. 36, no. 4, pp. 1234–1240, 2019
2019
Later among the works it cites.
S. Gupta, V. Tolani, J. Davidson, S. Levine, R. Sukthankar, and J. Malik, “Cognitive mapping and planning for visual navigation,” IJCV , vol. 128, no. 5, pp. 1311–1330, 2020
2020
Later among the works it cites.
W. Qi, R. T. Mullapudi, S. Gupta, and D. Ramanan, “Learning to move with affordance maps,” in ICLR , 2020
2020
Later among the works it cites.
D. S. Chaplot, D. Gandhi, S. Gupta, A. Gupta, and R. Salakhutdinov, “Learning to explore using active neural slam,” in ICLR , 2020
2020
Later among the works it cites.
L. Yan, D. Liu, Y. Song, and C. Yu, “Multimodal aggregation approach for memory vision-voice indoor navigation with meta-learning,” in IROS , 2020, pp. 5847–5854
2020
Later among the works it cites.
M. Deitke, W. Han, A. Herrasti, A. Kembhavi, E. Kolve, R. Mottaghi, J. Salvador, D. Schwenk, E. VanderBilt, M. Wallingford et al. , “Robothor: An open simulation-to-real embodied ai platform,” in CVPR , 2020, pp. 3164–3174
2020
Later among the works it cites.
F. Xia, W. B. Shen, C. Li, P. Kasimbeg, M. E. Tchapmi, A. Toshev, R. Martin-Martin, and S. Savarese, “Interactive gibson benchmark: A benchmark for interactive navigation in cluttered environments,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 713–720, 2020
2020
Later among the works it cites.
S. Wani, S. Patel, U. Jain, A. X. Chang, and M. Savva, “Multion: Benchmarking semantic map memory using multi-object navigation,” in NeurIPS , vol. 33, 2020
2020
Later among the works it cites.
A. Ku, P. Anderson, R. Patel, E. Ie, and J. Baldridge, “Room-across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding,” in EMNLP , 2020, pp. 4392–4412
2020
Later among the works it cites.
Y. Hong, C. Rodriguez, Q. Wu, and S. Gould, “Sub-instruction aware vision-and-language navigation,” in EMNLP , 2020, pp. 3360–3376
2020
Later among the works it cites.
Y. Qi, Q. Wu, P. Anderson, X. Wang, W. Y. Wang, C. Shen, and A. van den Hengel, “Reverie: Remote embodied visual referring expression in real indoor environments,” in CVPR , 2020, pp. 9982–9991
2020
Later among the works it cites.
C. Chen, U. Jain, C. Schissler, S. V. A. Gari, Z. Al-Halah, V. K. Ithapu, P. Robinson, and K. Grauman, “Soundspaces: Audio-visual navigation in 3d environments,” in ECCV , 2020, pp. 17–36
2020
Later among the works it cites.
W. Zhu, H. Hu, J. Chen, Z. Deng, V. Jain, E. Ie, and F. Sha, “Babywalk: Going farther in vision-and-language navigation by taking baby steps,” in ACL , 2020, pp. 2539–2556
2020
Later among the works it cites.
Y. Li and J. Kosecka, “Learning view and target invariant visual servoing for navigation,” in ICRA , 2020, pp. 658–664
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. Sun, X. Wang, Z. Liu, J. Miller, A. A. Efros, and M. Hardt, “Test-time training with self-supervision for generalization under distribution shifts,” in ICML , 2020, pp. 9229–9248
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
V. Dean, S. Tulsiani, and A. Gupta, “See, hear, explore: Curiosity via audio-visual association,” in NeurIPS , vol. 33, 2020
2020
Later among the works it cites.
N. Yang, L. v. Stumberg, R. Wang, and D. Cremers, “D3vo: Deep depth, deep pose and deep uncertainty for monocular visual odometry,” in CVPR , 2020, pp. 1281–1292
2020
Later among the works it cites.
D. S. Chaplot, R. Salakhutdinov, A. Gupta, and S. Gupta, “Neural topological slam for visual navigation,” in CVPR , 2020, pp. 12 875–12 884
2020
Later among the works it cites.
D. S. Chaplot, D. P. Gandhi, A. Gupta, and R. R. Salakhutdinov, “Object goal navigation using goal-oriented semantic exploration,” in NeurIPS , vol. 33, 2020
2020
Later among the works it cites.
Y. Zhu, F. Zhu, Z. Zhan, B. Lin, J. Jiao, X. Chang, and X. Liang, “Vision-dialog navigation by exploring cross-modal memory,” in CVPR , 2020, pp. 10 730–10 739
2020
Later among the works it cites.
H. Wang, Q. Wu, and C. Shen, “Soft expert reward learning for vision-and-language navigation,” in ECCV , 2020, pp. 126–141
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. Hong, C. Rodriguez, Y. Qi, Q. Wu, and S. Gould, “Language and visual entity relationship graph for agent navigation,” in NeurIPS , vol. 33, 2020
2020
Later among the works it cites.
A. Majumdar, A. Shrivastava, S. Lee, P. Anderson, D. Parikh, and D. Batra, “Improving vision-and-language navigation with image-text pairs from the web,” in ECCV , 2020, pp. 259–274
2020
Later among the works it cites.
W. Hao, C. Li, X. Li, L. Carin, and J. Gao, “Towards learning a generic agent for vision-and-language navigation via pre-training,” in CVPR , 2020, pp. 13 137–13 146
2020
Later among the works it cites.
Y. Wu, L. Jiang, and Y. Yang, “Revisiting embodiedqa: A simple baseline and beyond,” IEEE Transactions on Image Processing , vol. 29, pp. 3984–3992, 2020
2020
Later among the works it cites.
S. Tan, W. Xiang, H. Liu, D. Guo, and F. Sun, “Multi-agent embodied question answering in interactive environments,” in ECCV , 2020, pp. 663–678
2020
Later among the works it cites.
D. Nilsson, A. Pirinen, E. Gärtner, and C. Sminchisescu, “Embodied visual active learning for semantic segmentation.” in AAAI , 2020, pp. 2373–2383
2020
Later among the works it cites.
H. R. Roman, Y. Bisk, J. Thomason, A. Celikyilmaz, and J. Gao, “Rmm: A recursive mental model for dialog navigation,” in EMNLP , 2020, pp. 1732–1745
2020
Later among the works it cites.
M. Hahn, J. Krantz, D. Batra, D. Parikh, J. M. Rehg, S. Lee, and P. Anderson, “Where are you? localization from embodied dialog,” in EMNLP , 2020, pp. 806–822
2020
Later among the works it cites.
S. Banerjee, J. Thomason, and J. J. Corso, “The robotslang benchmark: Dialog-guided robot localization and navigation.” arXiv: Robotics , 2020
2020
Later among the works it cites.
T. Manderson, J. C. G. Higuera, S. Wapnick, J.-F. Tremblay, F. Shkurti, D. Meger, and G. Dudek, “Vision-based goal-conditioned policies for underwater navigation in the presence of obstacles,” in RSS , vol. 16, 2020
2020
Later among the works it cites.
A. Francis, A. Faust, H.-T. L. Chiang, J. Hsu, J. C. Kew, M. Fiser, and T.-W. E. Lee, “Long-range indoor navigation with prm-rl,” IEEE Transactions on Robotics , vol. 36, no. 4, pp. 1115–1134, 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
K. Lee, Y. Seo, S. Lee, H. Lee, and J. Shin, “Context-aware dynamics model for generalization in model-based reinforcement learning,” in ICML , vol. 1, 2020, pp. 5757–5766
2020
Later among the works it cites.
E. Wijmans, A. Kadian, A. Morcos, S. Lee, I. Essa, D. Parikh, M. Savva, and D. Batra, “Dd-ppo: Learning near-perfect pointgoal navigators from 2.5 billion frames,” in Eighth ICLR , 2020
2020
Later among the works it cites.
H. Wang, W. Wang, T. Shu, W. Liang, and J. Shen, “Active visual information gathering for vision-language navigation,” in ECCV , 2020, pp. 307–322
2020
Later among the works it cites.
S. K. Ramakrishnan, Z. Al-Halah, and K. Grauman, “Occupancy anticipation for efficient exploration and navigation,” in ECCV , 2020, pp. 400–418
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Li, X. Wang, S. Tang, H. Shi, F. Wu, Y. Zhuang, and W. Y. Wang, “Unsupervised reinforcement learning of transferable meta-skills for embodied navigation,” in CVPR , 2020, pp. 12 123–12 132
2020
Later among the works it cites.
D. S. Chaplot, L. Lee, R. Salakhutdinov, D. Parikh, and D. Batra, “Embodied multimodal multitask learning,” in IJCAI , vol. 3, 2020, pp. 2442–2448
2020
Later among the works it cites.
X. E. Wang, V. Jain, E. Ie, W. Y. Wang, Z. Kozareva, and S. Ravi, “Environment-agnostic multitask learning for natural language grounded navigation,” in ECCV , 2020, pp. 413–430
2020
Later among the works it cites.
W. Zhao, J. P. Queralta, and T. Westerlund, “Sim-to-real transfer in deep reinforcement learning for robotics: a survey,” in SSCI , 2020, pp. 737–744
2020
Later among the works it cites.
H. F. Bassani, R. A. Delgado, J. N. de O. Lima Junior, H. R. Medeiros, P. H. M. Braga, and A. Tapp, “Learning to play soccer by reinforcement and applying sim-to-real to compete in the real world,” CoRR , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
S. D. Morad, R. Mecca, R. P. K. Poudel, S. Liwicki, and R. Cipolla, “Embodied visual navigation with automatic curriculum learning in real environments,” in IEEE Robotics and Automation Letters , vol. 6, no. 2, 2021, pp. 683–690
2021
Closest in time.
Q. Wu, K. Xu, J. Wang, M. Xu, X. Gong, and D. Manocha, “Reinforcement learning-based visual navigation with information-theoretic regularization,” in IEEE Robotics and Automation Letters , vol. 6, no. 2, 2021, pp. 731–738
2021
Closest in time.
C. Chen, S. Majumder, Z. Al-Halah, R. Gao, S. K. Ramakrishnan, and K. Grauman, “Learning to set waypoints for audio-visual navigation,” in ICLR , 2021
2021
Closest in time.
2021
Closest in time.
R. Bigazzi, F. Landi, M. Cornia, S. Cascianelli, L. Baraldi, and R. Cucchiara, “Explore and explain: Self-supervised navigation and recounting,” in ICPR , 2021, pp. 1152–1159
2021
Closest in time.
Y. Hong, Q. Wu, Y. Qi, C. Rodriguez-Opazo, and S. Gould, “Vln bert: A recurrent vision-and-language bert for navigation,” in CVPR , 2021, pp. 1643–1653
2021
Closest in time.
Y. Deng, D. Guo, X. Guo, N. Zhang, H. Liu, and F. Sun, “Mqa: Answering the question via robotic manipulation,” in RSS , vol. 17, 2021
2021
Closest in time.
G. Berseth, D. Geng, C. M. Devin, N. Rhinehart, C. Finn, D. Jayaraman, and S. Levine, “Smirl: Surprise minimizing reinforcement learning in unstable environments,” in ICLR , 2021
2021
Closest in time.
Y. Zhu, Y. Weng, F. Zhu, X. Liang, Q. Ye, Y. Lu, and J. jiao, “Self-motivated communication agent for real-world vision-dialog navigation,” in ICCV , 2021
2021
Closest in time.
C. Wang, J. Miguel Buenaposada, R. Zhu, and S. Lucey, “Learning depth from monocular videos using direct methods,” in CVPR , 2018, pp. 2022–2030
2030
Closest in time.
R. Sim and J. J. Little, “Autonomous vision-based exploration and mapping using hybrid maps and rao-blackwellised particle filters,” in IROS . IEEE, 2006, pp. 2082–2089
2089
Closest in time.