Fetching the paper…
Reading the bibliography…
In this paper we propose a new framework - MoViLan (Modular Vision and Language) for execution of visually grounded natural language instructions for day to day indoor household tasks.
On the representation and estimation of spatial uncertainty
Randall C Smith and Peter Cheeseman · 1986
Earlier work this paper cites.
Simultaneous map building and localization for an autonomous mobile robot
John J Leonard and Hugh F Durrant-Whyte · 1991
Earlier work this paper cites.
Modelling navigational knowledge by route graphs
Steffen Werner, Bernd Krieg-Brückner, and Theo Herrmann · 2000
Earlier work this paper cites.
Learning for semantic parsing with statistical machine translation
Yuk Wah Wong and Raymond Mooney · 2006
Earlier work this paper cites.
Probabilistic robotics. sebastian thrun, wolfram burgard, and dieter fox.(2005, mit press.) 647 pages, 2008
Josh Bongard · 2008
Earlier work this paper cites.
A tutorial on graph-based slam
Giorgio Grisetti, Rainer Kümmerle, Cyrill Stachniss, and Wolfram Burgard · 2010
Earlier work this paper cites.
Following directions using statistical machine translation
Cynthia Matuszek, Dieter Fox, and Karl Koscher · 2010
Earlier work this paper cites.
Learning to interpret natural language navigation instructions from observations
David L Chen and Raymond J Mooney · 2011
Earlier work this paper cites.
Large-scale semantic mapping and reasoning with heterogeneous modalities
Andrzej Pronobis and Patric Jensfelt · 2012
Earlier work this paper cites.
Definition of semantic maps for outdoor robotic tasks
Dagmar Lang, Susanne Friedmann, Marcel Häselich, and Dietrich Paulus · 2014
Earlier work this paper cites.
Visual simultaneous localization and mapping: a survey
Jorge Fuentes-Pacheco, José Ruiz-Ascencio, and Juan Manuel Rendón-Mancha · 2015
Earlier work this paper cites.
What to talk about and how? selective generation using lstms with coarse-to-fine alignment
Hongyuan Mei, Mohit Bansal, and Matthew R Walter · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Deep learning
Ian Goodfellow, Yoshua Bengio, and Aaron Courville · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Listen, attend, and walk: Neural mapping of navigational instructions to action sequences
Hongyuan Mei, Mohit Bansal, and Matthew R Walter · 2016
Earlier work this paper cites.
Bidirectional attention flow for machine comprehension
Minjoon Seo, Aniruddha Kembhavi, Ali Farhadi, and Hannaneh Hajishirzi · 2016
Earlier work this paper cites.
Value iteration networks
Aviv Tamar, Yi Wu, Garrett Thomas, Sergey Levine, and Pieter Abbeel · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al · 2016
Cited alongside, same era.
Dynamic memory networks for visual and textual question answering
Caiming Xiong, Stephen Merity, and Richard Socher · 2016
Cited alongside, same era.
End-to-end navigation in unknown environments using neural networks
Arbaaz Khan, Clark Zhang, Nikolay Atanasov, Konstantinos Karydis, Daniel D Lee, and Vijay Kumar · 2017
Cited alongside, same era.
Ai2-thor: An interactive 3d environment for visual ai
Eric Kolve, Roozbeh Mottaghi, Winson Han, Eli VanderBilt, Luca Weihs, Alvaro Herrasti, Daniel Gordon, Yuke Zhu, Abhinav Gupta, and Ali Farhadi · 2017
Xiaoxue Zang, Ashwini Pokle, Marynel Vázquez, Kevin Chen, Juan Carlos Niebles, Alvaro Soto, and Silvio Savarese · 2018
Later among the works it cites.
Chasing ghosts: Instruction following as bayesian state tracking
Peter Anderson, Ayush Shrivastava, Devi Parikh, Dhruv Batra, and Stefan Lee · 2019
Later among the works it cites.
A behavioral approach to visual navigation with graph localization networks
Kevin Chen, Juan Pablo de Vicente, Gabriel Sepulveda, Fei Xia, Alvaro Soto, Marynel Vázquez, and Silvio Savarese · 2019
Later among the works it cites.
Bert for joint intent classification and slot filling
Qian Chen, Zhu Zhuo, and Wen Wang · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Mapping instructions and visual observations to actions with reinforcement learning
Dipendra Misra, John Langford, and Yoav Artzi · 2017
Cited alongside, same era.
Pointnet: Deep learning on point sets for 3d classification and segmentation
Charles R Qi, Hao Su, Kaichun Mo, and Leonidas J Guibas · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Neural slam: Learning to explore with external memory
Jingwei Zhang, Lei Tai, Joschka Boedecker, Wolfram Burgard, and Ming Liu · 2017
Cited alongside, same era.
Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
Peter Anderson, Qi Wu, Damien Teney, Jake Bruce, Mark Johnson, Niko Sünderhauf, Ian Reid, Stephen Gould, and Anton van den Hengel · 2018
Cited alongside, same era.
Mapnet: An allocentric spatial memory for mapping environments
Joao F Henriques and Andrea Vedaldi · 2018
Cited alongside, same era.
Lisa Lee, Emilio Parisotto, Devendra Singh Chaplot, Eric Xing, and Ruslan Salakhutdinov · 2018
Cited alongside, same era.
Niko Sünderhauf · 2019
Later among the works it cites.
Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation
Xin Wang, Qiuyuan Huang, Asli Celikyilmaz, Jianfeng Gao, Dinghan Shen, Yuan-Fang Wang, William Yang Wang, and Lei Zhang · 2019
Later among the works it cites.
Cross-modal self-attention network for referring image segmentation
Linwei Ye, Mrigank Rochan, Zhi Liu, and Yang Wang · 2019
Later among the works it cites.
Trajectory planning and tracking for autonomous vehicle based on state lattice and model predictive control
Chaoyong Zhang, Duanfeng Chu, Shidong Liu, Zejian Deng, Chaozhong Wu, and Xiaocong Su · 2019
Later among the works it cites.
Canet: Class-agnostic segmentation networks with iterative refinement and attentive few-shot learning
Chi Zhang, Guosheng Lin, Fayao Liu, Rui Yao, and Chunhua Shen · 2019
Later among the works it cites.
Object interaction
Allen AI · 2020
Later among the works it cites.
Learning to explore using active neural slam
Devendra Singh Chaplot, Dhiraj Gandhi, Saurabh Gupta, Abhinav Gupta, and Ruslan Salakhutdinov · 2020
Later among the works it cites.
Graphconv
Dgl · 2020
Later among the works it cites.
Lingunet
lil lab · 2020
Later among the works it cites.
Jointbert
monologg · 2020
Later among the works it cites.
Alfred: A benchmark for interpreting grounded instructions for everyday tasks
Mohit Shridhar, Jesse Thomason, Daniel Gordon, Yonatan Bisk, Winson Han, Roozbeh Mottaghi, Luke Zettlemoyer, and Dieter Fox · 2020
Later among the works it cites.
Vision-dialog navigation by exploring cross-modal memory
Yi Zhu, Fengda Zhu, Zhaohuan Zhan, Bingqian Lin, Jianbin Jiao, Xiaojun Chang, and Xiaodan Liang · 2020
Later among the works it cites.