Fetching the paper…
Reading the bibliography…
Intelligent embodied agents (e.g.
General Evaluation for Instruction Conditioned Navigation using Dynamic Time Warping, 2019
Gabriel Ilharco, Vihan Jain, Alexander Ku, Eugene Ie, and Jason Baldridge · 1907
Earlier work this paper cites.
DD-PPO: Learning near-perfect pointgoal navigators from 2.5 billion frames
Erik Wijmans, Abhishek Kadian, Ari Morcos, Stefan Lee, Irfan Essa, Devi Parikh, Manolis Savva, and Dhruv Batra · 1911
Earlier work this paper cites.
ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks
Mohit Shridhar, Jesse Thomason, Daniel Gordon, Yonatan Bisk, Winson Han, Roozbeh Mottaghi, Luke Zettlemoyer, and Dieter Fox · 1912
Earlier work this paper cites.
Hybrid position/force control of manipulators
M. H. Raibert and J. J. Craig · 1981
Earlier work this paper cites.
Shakey the robot
Artificial Intellgence Center · 1984
Earlier work this paper cites.
Manipulability of Robotic Mechanisms
Tsuneo Yoshikawa · 1985
Earlier work this paper cites.
Using occupancy grids for mobile robot perception and navigation
A. Elfes · 1989
Earlier work this paper cites.
Fast-marching level-set methods for three-dimensional photolithography development
James A Sethian · 1996
Earlier work this paper cites.
The dynamic window approach to collision avoidance
Dieter Fox, Wolfram Burgard, and Sebastian Thrun · 1997
Earlier work this paper cites.
A frontier-based approach for autonomous exploration
Brian Yamauchi · 1997
Earlier work this paper cites.
Integrating topological and metric maps for mobile robot navigation: A statistical approach
Sebastian Thrun, Jens-Steffen Gutmann, Dieter Fox, Wolfram Burgard, Benjamin Kuipers, et al · 1998
Earlier work this paper cites.
Feature correspondence: A markov chain monte carlo approach
Frank Dellaert, Steven Seitz, Sebastian Thrun, and Charles Thorpe · 2000
Earlier work this paper cites.
A real-time algorithm for mobile robot mapping with applications to multi-robot and 3D mapping
Sebastian Thrun, Wolfram Burgard, and Dieter Fox · 2000
Earlier work this paper cites.
Topological simultaneous localization and mapping (SLAM): toward exact localization without explicit localization
Howie Choset and Keiji Nagatani · 2001
Earlier work this paper cites.
Robust Monte Carlo localization for mobile robots
Sebastian Thrun, Dieter Fox, Wolfram Burgard, and Frank Dellaert · 2001
Earlier work this paper cites.
Combining topological and metric: A natural integration for simultaneous localization and map building
Nicola Tomatis, Illah Nourbakhsh, and Roland Siegwart · 2001
Earlier work this paper cites.
3D Dynamic Scene Graphs: Actionable Spatial Perception with Places, Objects, and Humans
A. Rosinol, A. Gupta, M. Abate, J. Shi, and L. Carlone · 2002
Earlier work this paper cites.
Probabilistic robotics
Sebastian Thrun · 2002
Earlier work this paper cites.
Robotic mapping: a survey
Sebastian Thrun · 2003
Earlier work this paper cites.
Selective neural representation of objects relevant for navigation
Gabriele Janzen and Miranda Van Turennout · 2004
Earlier work this paper cites.
An introduction to factor graphs
H-A Loeliger · 2004
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
David G Lowe · 2004
Earlier work this paper cites.
Embodied artificial intelligence: Trends and challenges
Rolf Pfeifer and Fumiya Iida · 2004
Earlier work this paper cites.
Do humans integrate routes into a cognitive map? map-versus landmark-based navigation of novel shortcuts
Patrick Foo, William H Warren, Andrew Duchon, and Michael J Tarr · 2005
Earlier work this paper cites.
Multi-hierarchical semantic maps for mobile robotics
Cipriano Galindo, Alessandro Saffiotti, Silvia Coradeschi, Pär Buschka, Juan-Antonio Fernandez-Madrigal, and Javier González · 2005
Earlier work this paper cites.
Information Gain-based Exploration Using Rao-Blackwellized Particle Filters
C. Stachniss, Giorgio Grisetti, and Wolfram Burgard · 2005
Earlier work this paper cites.
ObjectNav revisited: On evaluation of embodied agents navigating to objects
Dhruv Batra, Aaron Gokaslan, Aniruddha Kembhavi, Oleksandr Maksymets, Roozbeh Mottaghi, Manolis Savva, Alexander Toshev, and Erik Wijmans · 2006
Earlier work this paper cites.
Changhao Chen, Bing Wang, Chris Xiaoxuan Lu, Niki Trigoni, and Andrew Markham · 2006
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, and action in route instructions
Matt MacMahon, Brian Stankiewicz, and Benjamin Kuipers · 2006
Earlier work this paper cites.
The graph SLAM algorithm with applications to large-scale mapping of urban structures
Sebastian Thrun and Michael Montemerlo · 2006
Earlier work this paper cites.
Wifi-slam using gaussian process latent variable models
Brian Ferris, Dieter Fox, and Neil D Lawrence · 2007
Earlier work this paper cites.
Supervised semantic labeling of places using information extracted from sensor data
Oscar Martinez Mozos, Rudolph Triebel, Patric Jensfelt, Axel Rottmann, and Wolfram Burgard · 2007
Earlier work this paper cites.
6D SLAM—3D mapping outdoor environments
Andreas Nüchter, Kai Lingemann, Joachim Hertzberg, and Hartmut Surmann · 2007
Earlier work this paper cites.
Speeded-up robust features (surf)
Herbert Bay, Andreas Ess, Tinne Tuytelaars, and Luc Van Gool · 2008
Earlier work this paper cites.
Curious george: An attentive semantic robot
David Meger, Per-Erik Forssén, Kevin Lai, Scott Helmer, Sancho McCann, Tristram Southey, Matthew Baumann, James J Little, and David G Lowe · 2008
Earlier work this paper cites.
Towards semantic maps for mobile robots
Andreas Nüchter and Joachim Hertzberg · 2008
Earlier work this paper cites.
Bayesian space conceptualization and place classification for semantic maps in mobile robotics
Shrihari Vasudevan and Roland Siegwart · 2008
Earlier work this paper cites.
Conceptual spatial representations for indoor mobile robots
Hendrik Zender, O Martínez Mozos, Patric Jensfelt, G-JM Kruijff, and Wolfram Burgard · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Localization, mapping, and planning in 3D environments
Nathaniel Fairfield · 2009
Earlier work this paper cites.
Appearance-based loop detection from 3D laser data using the normal distributions transform
Martin Magnusson, Henrik Andreasson, Andreas Nuchter, and Achim J. Lilienthal · 2009
Earlier work this paper cites.
Robust 3D-mapping with time-of-flight cameras
Stefan May, David Dröschel, Stefan Fuchs, Dirk Holz, and Andreas Nüchter · 2009
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2010
Earlier work this paper cites.
Following directions using statistical machine translation
Cynthia Matuszek, Dieter Fox, and Karl Koscher · 2010
Earlier work this paper cites.
Multi-modal semantic place classification
Andrzej Pronobis, Oscar Martinez Mozos, Barbara Caputo, and Patric Jensfelt · 2010
Earlier work this paper cites.
Search in the real world: Active visual object search based on spatial relations
Alper Aydemir, Kristoffer Sjöö, John Folkesson, Andrzej Pronobis, and Patric Jensfelt · 2011
Earlier work this paper cites.
Semantic structure from motion
Sid Yingze Bao and Silvio Savarese · 2011
Earlier work this paper cites.
Rearrangement: A challenge for embodied ai
Dhruv Batra, Angel X Chang, Sonia Chernova, Andrew J Davison, Jia Deng, Vladlen Koltun, Sergey Levine, Jitendra Malik, Igor Mordatch, Roozbeh Mottaghi, et al · 2011
Earlier work this paper cites.
Learning to Interpret Natural Language Navigation Instructions from Observations
David Chen and Raymond Mooney · 2011
Earlier work this paper cites.
Towards semantic SLAM using a monocular camera
Javier Civera, Dorian Gálvez-López, Luis Riazuelo, Juan D Tardós, and Jose Maria Martinez Montiel · 2011
Earlier work this paper cites.
Manhattan scene understanding using monocular, stereo, and 3d features
Alex Flint, David Murray, and Ian Reid · 2011
Earlier work this paper cites.
Following and interpreting narrated guided tours
Sachithra Hemachandra, Thomas Kollar, Nicholas Roy, and Seth Teller · 2011
Earlier work this paper cites.
Efficient, generalized indoor wifi graphslam
Joseph Huang, David Millman, Morgan Quigley, David Stavens, Sebastian Thrun, and Alok Aggarwal · 2011
Earlier work this paper cites.
Navigation in hybrid metric-topological maps
Kurt Konolige, Eitan Marder-Eppstein, and Bhaskara Marthi · 2011
Earlier work this paper cites.
ORB: An efficient alternative to SIFT or SURF
Ethan Rublee, Vincent Rabaud, Kurt Konolige, and Gary Bradski · 2011
Earlier work this paper cites.
3-D scene analysis via sequenced predictions over points and regions
Xuehan Xiong, Daniel Munoz, J Andrew Bagnell, and Martial Hebert · 2011
Earlier work this paper cites.
Semantic structure from motion with points, regions, and objects
Sid Yingze Bao, Mohit Bagra, Yu-Wei Chao, and Silvio Savarese · 2012
Earlier work this paper cites.
Performance of histogram descriptors for the classification of 3D laser range data in urban environments
Jens Behley, Volker Steinhage, and Armin B Cremers · 2012
Earlier work this paper cites.
From objects to landmarks: The function of visual location information in spatial navigation
Edgar Chan, Oliver Baumann, Mark Bellgrove, and Jason Mattingley · 2012
Earlier work this paper cites.
Neural systems for landmark-based wayfinding in humans
Russell A Epstein and Lindsay K Vass · 2012
Earlier work this paper cites.
Joint 2d-3d temporally consistent semantic segmentation of street scenes
Georgios Floros and Bastian Leibe · 2012
Earlier work this paper cites.
Kintinuous: Spatially extended kinectfusion
M Kaess, M Fallon, H Johannsson, and JJ Leonard · 2012
Earlier work this paper cites.
Fast object localization and pose estimation in heavy clutter for robotic bin picking
Ming-Yu Liu, Oncel Tuzel, Ashok Veeraraghavan, Yuichi Taguchi, Tim K. Marks, and Rama Chellappa · 2012
Earlier work this paper cites.
Learning to Parse Natural Language Commands to a Robot Control System
Cynthia Matuszek, Evan V. Herbst, Luke Zettlemoyer, and Dieter Fox · 2012
Earlier work this paper cites.
Large-scale semantic mapping and reasoning with heterogeneous modalities
Andrzej Pronobis and Patric Jensfelt · 2012
Earlier work this paper cites.
Semantic mapping using object-class segmentation of RGB-D images
Jörg Stückler, Nenad Biresev, and Sven Behnke · 2012
Earlier work this paper cites.
Parsing outdoor scenes from streamed 3d laser data using online clustering and incremental belief updates
Rudolph Triebel, Rohan Paul, Daniela Rus, and Paul Newman · 2012
Earlier work this paper cites.
Contextually guided semantic labeling and search for three-dimensional point clouds
Abhishek Anand, Hema Swetha Koppula, Thorsten Joachims, and Ashutosh Saxena · 2013
Earlier work this paper cites.
Active visual object search in unknown environments using uncertain semantics
Alper Aydemir, Andrzej Pronobis, Moritz Göbelbecker, and Patric Jensfelt · 2013
Earlier work this paper cites.
Active SLAM and Exploration with Particle Filters Using Kullback-Leibler Divergence
Luca Carlone, Jingjing Du, Miguel Efrain Kaouk Ng, Basilio Bona, and Marina Indri · 2013
Earlier work this paper cites.
Imitation learning for natural language direction following through unknown environments
Felix Duvallet, Thomas Kollar, and Anthony Stentz · 2013
Earlier work this paper cites.
3-D mapping with an RGB-D camera
Felix Endres, Jürgen Hess, Jürgen Sturm, Daniel Cremers, and Wolfram Burgard · 2013
Earlier work this paper cites.
Joint detection, tracking and mapping by semantic bundle adjustment
Nicola Fioraio and Luigi Di Stefano · 2013
Earlier work this paper cites.
Joint 3D scene reconstruction and class segmentation
Christian Hane, Christopher Zach, Andrea Cohen, Roland Angst, and Marc Pollefeys · 2013
Earlier work this paper cites.
Efficient 3-d scene analysis from streaming data
Hanzhang Hu, Daniel Munoz, J Andrew Bagnell, and Martial Hebert · 2013
Earlier work this paper cites.
Semantic mapping and navigation: A Bayesian approach
Dong Wook Ko, Chuho Yi, and Il Hong Suh · 2013
Earlier work this paper cites.
Learning to parse natural language commands to a robot control system
Cynthia Matuszek, Evan Herbst, Luke Zettlemoyer, and Dieter Fox · 2013
Earlier work this paper cites.
Slam++: Simultaneous localisation and mapping at the level of objects
Renato F Salas-Moreno, Richard A Newcombe, Hauke Strasdat, Paul HJ Kelly, and Andrew J Davison · 2013
Earlier work this paper cites.
Mesh based semantic modelling for indoor and outdoor scenes
Julien PC Valentin, Sunando Sengupta, Jonathan Warrell, Ali Shahrokni, and Philip HS Torr · 2013
Earlier work this paper cites.
Learning Semantic Maps from Natural Language Descriptions
Matthew R Walter, Sachithra Hemachandra, Bianca Homberg, Stefanie Tellex, and Seth J Teller · 2013
Earlier work this paper cites.
LSD-SLAM: Large-scale direct monocular SLAM
Jakob Engel, Thomas Schöps, and Daniel Cremers · 2014
Earlier work this paper cites.
Learning spatial-semantic representations from natural language descriptions and scene classifications
Sachithra Hemachandra, Matthew R Walter, Stefanie Tellex, and Seth Teller · 2014
Earlier work this paper cites.
Dense 3d semantic mapping of indoor scenes from rgb-d images
Alexander Hermans, Georgios Floros, and Bastian Leibe · 2014
Earlier work this paper cites.
Joint semantic segmentation and 3d reconstruction from monocular video
Abhijit Kundu, Yin Li, Frank Dellaert, Fuxin Li, and James M Rehg · 2014
Earlier work this paper cites.
Unsupervised feature learning for 3d scene labeling
Kevin Lai, Liefeng Bo, and Dieter Fox · 2014
Earlier work this paper cites.
Prior-assisted propagation of spatial information for object search
Malte Lorbach, Sebastian Höfer, and Oliver Brock · 2014
Earlier work this paper cites.
Multi-resolution surfel maps for efficient dense 3D modeling and tracking
Jörg Stückler and Sven Behnke · 2014
Earlier work this paper cites.
A framework for learning semantic maps from grounded natural language descriptions
Matthew R Walter, Sachithra Hemachandra, Bianca Homberg, Stefanie Tellex, and Seth Teller · 2014
Earlier work this paper cites.
A fast, modular scene understanding system using context-aware object detection
Cesar Cadena, Anthony Dick, and Ian D Reid · 2015
Earlier work this paper cites.
Learning models for following natural language directions in unknown environments
Sachithra Hemachandra, Felix Duvallet, Thomas M Howard, Nicholas Roy, Anthony Stentz, and Matthew R Walter · 2015
Earlier work this paper cites.
Simultaneous localization and mapping with infinite planes
Michael Kaess · 2015
Earlier work this paper cites.
Semantic mapping for mobile robotics tasks: A survey
Ioannis Kostavelis and Antonios Gasteratos · 2015
Earlier work this paper cites.
ORB-SLAM: A versatile and accurate monocular SLAM system
Raul Mur-Artal, Jose Maria Martinez Montiel, and Juan D Tardos · 2015
Earlier work this paper cites.
Monocular slam supported object recognition
Sudeep Pillai and John Leonard · 2015
Earlier work this paper cites.
Faster R-CNN: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Semantic octree: Unifying recognition, reconstruction and representation via an octree constrained higher order mrf
Sunando Sengupta and Paul Sturgess · 2015
Earlier work this paper cites.
Incremental dense semantic stereo fusion for large-scale semantic scene reconstruction
Vibhav Vineet, Ondrej Miksik, Morten Lidegaard, Matthias Nießner, Stuart Golodetz, Victor A Prisacariu, Olaf Kähler, David W Murray, Shahram Izadi, Patrick Pérez, et al · 2015
Earlier work this paper cites.
ElasticFusion: Dense SLAM without a pose graph
Thomas Whelan, Stefan Leutenegger, Renato F Salas-Moreno, Ben Glocker, and Andrew J Davison · 2015
Earlier work this paper cites.
3d semantic parsing of large-scale indoor spaces
Iro Armeni, Ozan Sener, Amir R Zamir, Helen Jiang, Ioannis Brilakis, Martin Fischer, and Silvio Savarese · 2016
Earlier work this paper cites.
Simultaneous localization and mapping: Present, future, and the robust-perception age
Cesar Cadena, Luca Carlone, Henry Carrillo, Yasir Latif, Davide Scaramuzza, José Neira, Ian D Reid, and John J Leonard · 2016
Earlier work this paper cites.
Inferring maps and behaviors from natural language instructions
Felix Duvallet, Matthew R Walter, Thomas Howard, Sachithra Hemachandra, Jean Oh, Seth Teller, Nicholas Roy, and Anthony Stentz · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
A review of spatial reasoning and interaction for real-world robotics
Christian Landsiedel, Verena Rieser, Matthew Walter, and Dirk Wollherr · 2016
Earlier work this paper cites.
Researchdoom and cocodoom: Learning computer vision with games
Aravindh Mahendran, Hakan Bilen, João F Henriques, and Andrea Vedaldi · 2016
Earlier work this paper cites.
Listen, attend, and walk: Neural mapping of navigational instructions to action sequences
Hongyuan Mei, Mohit Bansal, and Matthew Walter · 2016
Earlier work this paper cites.
ElasticFusion: Real-time dense SLAM and light source estimation
Thomas Whelan, Renato F Salas-Moreno, Ben Glocker, Andrew J Davison, and Stefan Leutenegger · 2016
Earlier work this paper cites.
A dataset for developing and benchmarking active vision
Phil Ammirato, Patrick Poirson, Eunbyung Park, Jana Košecká, and Alexander C Berg · 2017
Cited alongside, same era.
Probabilistic data association for semantic SLAM
Sean L Bowman, Nikolay Atanasov, Kostas Daniilidis, and George J Pappas · 2017
Cited alongside, same era.
Past, present, and future of simultaneous localization and mapping: Toward the robust-perception age
Cesar Cadena, Luca Carlone, Henry Carrillo, Yasir Latif, Davide Scaramuzza, José Neira, Ian Reid, and John J Leonard · 2017
Cited alongside, same era.
Matterport3D: Learning from RGB-D data in indoor environments
Angel Chang, Angela Dai, Thomas Funkhouser, Maciej Halber, Matthias Niebner, Manolis Savva, Shuran Song, Andy Zeng, and Yinda Zhang · 2017
Cited alongside, same era.
Bundlefusion: Real-time globally consistent 3d reconstruction using on-the-fly surface reintegration
Angela Dai, Matthias Nießner, Michael Zollhöfer, Shahram Izadi, and Christian Theobalt · 2017
Cited alongside, same era.
Panoptic nerf: 3d-to-2d label transfer for panoptic urban scene segmentation
Xiao Fu, Shangzhan Zhang, Tianrun Chen, Yichong Lu, Lanyun Zhu, Xiaowei Zhou, Andreas Geiger, and Yiyi Liao · 2022
Later among the works it cites.
Open-vocabulary object detection via vision and language knowledge distillation
Xiuye Gu, Tsung-Yi Lin, Weicheng Kuo, and Yin Cui · 2022
Later among the works it cites.
Hydra: A real-time spatial perception system for 3D scene graph construction and optimization
Nathan Hughes, Yun Chang, and Luca Carlone · 2022
Later among the works it cites.
Simple but effective: CLIP embeddings for embodied AI
Apoorv Khandelwal, Luca Weihs, Roozbeh Mottaghi, and Aniruddha Kembhavi · 2022
Later among the works it cites.
Instance-specific image goal navigation: Training embodied agents to find object instances
Jacob Krantz, Stefan Lee, Jitendra Malik, Dhruv Batra, and Devendra Singh Chaplot · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Factor graphs for robot perception
Frank Dellaert, Michael Kaess, et al · 2017
Cited alongside, same era.
Cognitive mapping and planning for visual navigation
Saurabh Gupta, James Davidson, Sergey Levine, Rahul Sukthankar, and Jitendra Malik · 2017
Cited alongside, same era.
AI2-THOR: An Interactive 3D Environment for Visual AI
Eric Kolve, Roozbeh Mottaghi, Winson Han, Eli VanderBilt, Luca Weihs, Alvaro Herrasti, Matt Deitke, Kiana Ehsani, Daniel Gordon, Yuke Zhu, Aniruddha Kembhavi, Abhinav Kumar Gupta, and Ali Farhadi · 2017
Cited alongside, same era.
Semanticfusion: Dense 3d semantic mapping with convolutional neural networks
John McCormac, Ankur Handa, Andrew Davison, and Stefan Leutenegger · 2017
Cited alongside, same era.
Visual-inertial monocular SLAM with map reuse
Raúl Mur-Artal and Juan D Tardós · 2017
Cited alongside, same era.
Curiosity-driven exploration by self-supervised prediction
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell · 2017
Cited alongside, same era.
Pointnet++: Deep hierarchical feature learning on point sets in a metric space
Charles Ruizhongtai Qi, Li Yi, Hao Su, and Leonidas J Guibas · 2017
Cited alongside, same era.
Later among the works it cites.
Language-driven semantic segmentation
Boyi Li, Kilian Q Weinberger, Serge Belongie, Vladlen Koltun, and Rene Ranftl · 2022
Later among the works it cites.
A survey of transformers
Tianyang Lin, Yuxin Wang, Xiangyang Liu, and Xipeng Qiu · 2022
Later among the works it cites.
Stubborn: A strong baseline for indoor object navigation
Haokuan Luo, Albert Yue, Zhang-Wei Hong, and Pulkit Agrawal · 2022
Later among the works it cites.
Curiosity-driven exploration via latent bayesian surprise
Pietro Mazzaglia, Ozan Catal, Tim Verbelen, and Bart Dhoedt · 2022
Later among the works it cites.
Stronger together: Air-ground robotic collaboration using semantics
Ian D Miller, Fernando Cladera, Trey Smith, Camillo Jose Taylor, and Vijay Kumar · 2022
Later among the works it cites.
ISDF: Real-time neural signed distance fields for robot perception
Joseph Ortiz, Alexander Clegg, Jing Dong, Edgar Sucar, David Novotny, Michael Zollhoefer, and Mustafa Mukadam · 2022
Later among the works it cites.
TEACh: Task-driven Embodied Agents that Chat
Aishwarya Padmakumar, Jesse Thomason, Ayush Shrivastava, Patrick Lange, Anjali Narayan-Chen, Spandana Gella, Robinson Piramuthu, and Dilek Hakkani-Tur Gokhan Tur and · 2022
Later among the works it cites.
Enough is enough: Towards autonomous uncertainty-driven stopping criteria
Julio A Placed and José A Castellanos · 2022
Later among the works it cites.
Towards Accurate Loop Closure Detection in Semantic SLAM With 3D Semantic Covisibility Graphs
Zhentian Qian, Jie Fu, and Jing Xiao · 2022
Later among the works it cites.
Visual slam: What are the current trends and what to expect?
Ali Tourani, Hriday Bavle, Jose Luis Sanchez-Lopez, and Holger Voos · 2022
Later among the works it cites.
A simple approach for visual room rearrangement: 3D mapping and semantic search
Brandon Trabucco, Gunnar A Sigurdsson, Robinson Piramuthu, Gaurav S Sukhatme, and Ruslan Salakhutdinov · 2022
Later among the works it cites.
The revisiting problem in simultaneous localization and mapping: A survey on visual loop closure detection
Konstantinos A Tsintotas, Loukas Bampis, and Antonios Gasteratos · 2022
Later among the works it cites.
Language Understanding for Field and Service Robots in a Priori Unknown Environments
Matthew R. Walter, Siddharth Patki, Andrea F. Daniele, Ethan Fahnestock, Felix Duvallet, Sachithra Hemachandra, Jean Oh, Anthony Stentz, Nicholas Roy, and Thomas M. Howard · 2022
Later among the works it cites.
Clip-nerf: Text-and-image driven manipulation of neural radiance fields
Can Wang, Menglei Chai, Mingming He, Dongdong Chen, and Jing Liao · 2022
Later among the works it cites.
Habitat challenge 2022
Karmesh Yadav, Santhosh Kumar Ramakrishnan, John Turner, Aaron Gokaslan, Oleksandr Maksymets, Rishabh Jain, Ram Ramrakhya, Angel X Chang, Alexander Clegg, Manolis Savva, Eric Undersander, Devendra Singh Chaplot, and Dhruv Batra · 2022
Later among the works it cites.
Extract free dense labels from CLIP
Chong Zhou, Chen Change Loy, and Bo Dai · 2022
Later among the works it cites.
NICE-SLAM: Neural implicit scalable encoding for SLAM
Zihan Zhu, Songyou Peng, Viktor Larsson, Weiwei Xu, Hujun Bao, Zhaopeng Cui, Martin R Oswald, and Marc Pollefeys · 2022
Later among the works it cites.
Active slam: A review on last decade
Muhammad Farhan Ahmed, Khayyam Masood, Vincent Fremont, and Isabelle Fantoni · 2023
Later among the works it cites.
BEVBert: Multimodal Map Pre-training for Language-guided Navigation
Dong An, Yuankai Qi, Yangguang Li, Yan Huang, Liang Wang, Tieniu Tan, and Jing Shao · 2023
Later among the works it cites.
A review of high-definition map creation methods for autonomous driving
Zhibin Bao, Sabir Hossain, Haoxiang Lang, and Xianke Lin · 2023
Later among the works it cites.
S-graphs+: Real-time localization and mapping leveraging hierarchical representations
Hriday Bavle, Jose Luis Sanchez-Lopez, Muhammad Shaheer, Javier Civera, and Holger Voos · 2023
Later among the works it cites.
Matthew Chang, Theophile Gervet, Mukul Khanna, Sriram Yenamandra, Dhruv Shah, So Yeon Min, Kavit Shah, Chris Paxton, Saurabh Gupta, Dhruv Batra, et al · 2023
Later among the works it cites.
PaLM-E: An Embodied Multimodal Language Model
Danny Driess, F. Xia, Mehdi S. M. Sajjadi, Corey Lynch, Aakanksha Chowdhery, Brian Ichter, Ayzaan Wahid, Jonathan Tompson, Quan Ho Vuong, Tianhe Yu, Wenlong Huang, Yevgen Chebotar, Pierre Sermanet, Daniel Duckworth, Sergey Levine, Vincent Vanhoucke, Karol Hausman, Marc Toussaint, Klaus Greff, Andy Zeng, Igor Mordatch, and Peter R. Florence · 2023
Later among the works it cites.
Semantic and topological mapping using intersection identification
Scott Fredriksson, Akshit Saradagi, and George Nikolakopoulos · 2023
Later among the works it cites.
Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation
Samir Yitzhak Gadre, Mitchell Wortsman, Gabriel Ilharco, Ludwig Schmidt, and Shuran Song · 2023
Later among the works it cites.
3D scene reconstruction and mapping with real time human detection for search and rescue robotics
JS Gautham, Akash Sharma, Samiappan Dhanalakshmi, and Kumar Ramamoorthy · 2023
Later among the works it cites.
Navigating to objects in the real world
Theophile Gervet, Soumith Chintala, Dhruv Batra, Jitendra Malik, and Devendra Singh Chaplot · 2023
Later among the works it cites.
ManiSkill2: A Unified Benchmark for Generalizable Manipulation Skills
Jiayuan Gu, Fanbo Xiang, Xuanlin Li, Zhan Ling, Xiqiang Liu, Tongzhou Mu, Yihe Tang, Stone Tao, Xinyue Wei, Yunchao Yao, Xiaodi Yuan, Pengwei Xie, Zhiao Huang, Rui Chen, and Hao Su · 2023
Later among the works it cites.
ConceptFusion: Open-set Multimodal 3D Mapping
Krishna Murthy Jatavallabhula, Alihusein Kuwajerwala, Qiao Gu, Mohd Omama, Tao Chen, Shuang Li, Ganesh Iyer, Soroush Saryazdi, Nikhil Keetha, Ayush Tewari, Joshua B. Tenenbaum, Celso Miguel de Melo, Madhava Krishna, Liam Paull, Florian Shkurti, and Antonio Torralba · 2023
Later among the works it cites.
LeRF: Language embedded radiance fields
Justin Kerr, Chung Min Kim, Ken Goldberg, Angjoo Kanazawa, and Matthew Tancik · 2023
Later among the works it cites.
Topological semantic graph memory for image-goal navigation
Nuri Kim, Obin Kwon, Hwiyeon Yoo, Yunho Choi, Jeongho Park, and Songhwai Oh · 2023
Later among the works it cites.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Renderable neural radiance map for visual navigation
Obin Kwon, Jeongho Park, and Songhwai Oh · 2023
Later among the works it cites.
Junnan Li, Dongxu Li, Silvio Savarese, and Steven Hoi · 2023
Later among the works it cites.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2023
Later among the works it cites.
Feature-realistic neural fusion for real-time, open set scene understanding
Kirill Mazur, Edgar Sucar, and Andrew J Davison · 2023
Later among the works it cites.
OpenAI · 2023
Later among the works it cites.
Dinov2: Learning robust visual features without supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, et al · 2023
Later among the works it cites.
OpenScene: 3D scene understanding with open vocabularies
Songyou Peng, Kyle Genova, Chiyu "Max" Jiang, Andrea Tagliasacchi, Marc Pollefeys, and Thomas Funkhouser · 2023
Later among the works it cites.
Visual SLAM integration with semantic segmentation and deep learning: A review
Huayan Pu, Jun Luo, Gang Wang, Tao Huang, and Hongliang Liu · 2023
Later among the works it cites.
MOPA: Modular Object Navigation with PointGoal Agents
Sonia Raychaudhuri, Tommaso Campari, Unnat Jain, Manolis Savva, and Angel X. Chang · 2023
Later among the works it cites.
Nerf-slam: Real-time dense monocular slam with neural radiance fields
Antoni Rosinol, John J Leonard, and Luca Carlone · 2023
Later among the works it cites.
CLIP-Fields: Weakly supervised semantic fields for robotic memory
Nur Muhammad Mahi Shafiullah, Chris Paxton, Lerrel Pinto, Soumith Chintala, and Arthur Szlam · 2023
Later among the works it cites.
LM-Nav: Robotic navigation with large pre-trained models of language, vision, and action
Dhruv Shah, Blazej Osinski, Brian Ichter, and Sergey Levine · 2023
Later among the works it cites.
Distilled feature fields enable few-shot language-guided manipulation
William Shen, Ge Yang, Alan Yu, Jansen Wong, Leslie Pack Kaelbling, and Phillip Isola · 2023
Later among the works it cites.
Vision-based dirt distribution mapping using deep learning
Ishneet Sukhvinder Singh, ID Wijegunawardana, SM Bhagya P Samarakoon, MA Viraj J Muthugala, and Mohan Rajesh Elara · 2023
Later among the works it cites.
A systematic literature review on long-term localization and mapping for mobile robots
Ricardo B Sousa, Héber M Sobreira, and António Paulo Moreira · 2023
Later among the works it cites.
Language-enhanced RNR-map: Querying renderable neural radiance field maps with natural language
Francesco Taioli, Federico Cunico, Federico Girella, Riccardo Bologna, Alessandro Farinelli, and Marco Cristani · 2023
Later among the works it cites.
Habitat-Matterport 3D Semantics dataset
Karmesh Yadav, Ram Ramrakhya, Santhosh Kumar Ramakrishnan, Theo Gervet, John Turner, Aaron Gokaslan, Noah Maestre, Angel Xuan Chang, Dhruv Batra, Manolis Savva, et al · 2023
Later among the works it cites.
VLFM: Vision-language frontier maps for zero-shot semantic navigation
Naoki Yokoyama, Sehoon Ha, Dhruv Batra, Jiuguang Wang, and Bernadette Bucher · 2023
Later among the works it cites.
ESC: Exploration with soft commonsense constraints for zero-shot object navigation
Kaiwen Zhou, Kaizhi Zheng, Connor Pryor, Yilin Shen, Hongxia Jin, Lise Getoor, and Xin Eric Wang · 2023
Later among the works it cites.
ETPNav: Evolving topological planning for vision-language navigation in continuous environments
Dong An, Hanqing Wang, Wenguan Wang, Zun Wang, Yan Huang, Keji He, and Liang Wang · 2024
Later among the works it cites.
SLAM Handbook
Luca Carlone, Ayoung Kim, Frank Dellaert, Timothy Barfoot, and Daniel Cremers · 2024
Later among the works it cites.
Unimate and Beyond: Exploring the Genesis of Industrial Robotics
Ovidiu-Aurelian Detesan and Iuliana Fabiola Moholea · 2024
Later among the works it cites.
Multi-level neural scene graphs for dynamic urban environments
Tobias Fischer, Lorenzo Porzi, Samuel Rota Bulo, Marc Pollefeys, and Peter Kontschieder · 2024
Later among the works it cites.
Robohop: Segment-based topological map representation for open-world visual navigation
Sourav Garg, Krishan Rana, Mehdi Hosseinzadeh, Lachlan Mares, Niko Sünderhauf, Feras Dayoub, and Ian Reid · 2024
Later among the works it cites.
Dylan Goetting, Himanshu Gaurav Singh, and Antonio Loquercio · 2024
Later among the works it cites.
Collaborative dynamic 3d scene graphs for automated driving
Elias Greve, Martin Büchner, Niclas Vödisch, Wolfram Burgard, and Abhinav Valada · 2024
Later among the works it cites.
Conceptgraphs: Open-vocabulary 3d scene graphs for perception and planning
Qiao Gu, Ali Kuwajerwala, Sacha Morin, Krishna Murthy Jatavallabhula, Bipasha Sen, Aditya Agarwal, Corban Rivera, William Paul, Kirsty Ellis, Rama Chellappa, et al · 2024
Later among the works it cites.
World models for autonomous driving: An initial survey
Yanchen Guan, Haicheng Liao, Zhenning Li, Jia Hu, Runze Yuan, Guohui Zhang, and Chengzhong Xu · 2024
Later among the works it cites.
Semantic Gaussians: Open-Vocabulary Scene Understanding with 3D Gaussian Splatting
Jun Guo, Xiaojian Ma, Yue Fan, Huaping Liu, and Qing Li · 2024
Later among the works it cites.
Foundations of spatial perception for robotics: Hierarchical representations and real-time systems
Nathan Hughes, Yun Chang, Siyi Hu, Rajat Talak, Rumaia Abdulhai, Jared Strader, and Luca Carlone · 2024
Later among the works it cites.
Goat-bench: A benchmark for multi-modal lifelong navigation
Mukul Khanna, Ram Ramrakhya, Gunjan Chhablani, Sriram Yenamandra, Theophile Gervet, Matthew Chang, Zsolt Kira, Devendra Singh Chaplot, Dhruv Batra, and Roozbeh Mottaghi · 2024
Later among the works it cites.
Embodied AI with Large Language Models: A Survey and New HRI Framework
Ming-Yi Lin, Ou-Wen Lee, and Chih-Ying Lu · 2024
Later among the works it cites.
Grounding dino: Marrying dino with grounded pre-training for open-set object detection
Shilong Liu, Zhaoyang Zeng, Tianhe Ren, Feng Li, Hao Zhang, Jie Yang, Qing Jiang, Chunyuan Li, Jianwei Yang, Hang Su, et al · 2024
Later among the works it cites.
InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment
Yuxing Long, Wenzhe Cai, Hongcheng Wang, Guanqi Zhan, and Hao Dong · 2024
Later among the works it cites.
Estimating Map Completeness in Robot Exploration
Matteo Luperto, Marco Maria Ferrara, Giacomo Boracchi, and Francesco Amigoni · 2024
Later among the works it cites.
Clio: Real-time task-driven open-set 3d scene graphs
Dominic Maggio, Yun Chang, Nathan Hughes, Matthew Trang, Dan Griffith, Carlyn Dougherty, Eric Cristofalo, Lukas Schmid, and Luca Carlone · 2024
Later among the works it cites.
OpenEQA: Embodied Question Answering in the Era of Foundation Models
Arjun Majumdar, Anurag Ajay, Xiaohan Zhang, Pranav Putta, Sriram Yenamandra, Mikael Henaff, Sneha Silwal, Paul Mcvay, Oleksandr Maksymets, Sergio Arnaud, Karmesh Yadav, Qiyang Li, Ben Newman, Mohit Sharma, Vincent-Pierre Berges, Shiqi Zhang, Pulkit Agrawal, Yonatan Bisk, Dhruv Batra, Mrinal Kalakrishnan, Franziska Meier, Chris Paxton, Alexander Sax, and Aravind Rajeswaran · 2024
Later among the works it cites.
QueSTMaps: Queryable Semantic Topological Maps for 3D Scene Understanding
Yash Mehan, Kumaraditya Gupta, Rohit Jayanti, Anirudh Govil, Sourav Garg, and Madhava Krishna · 2024
Later among the works it cites.
LangSplat: 3D language Gaussian splatting
Minghan Qin, Wanhua Li, Jiawei Zhou, Haoqian Wang, and Hanspeter Pfister · 2024
Later among the works it cites.
Learning generalizable feature fields for mobile manipulation
Ri-Zhao Qiu, Yafei Hu, Ge Yang, Yuchen Song, Yang Fu, Jianglong Ye, Jiteng Mu, Ruihan Yang, Nikolay Atanasov, Sebastian Scherer, et al · 2024
Later among the works it cites.
Explore until Confident: Efficient Exploration for Embodied Question Answering
Allen Z Ren, Jaden Clark, Anushri Dixit, Masha Itkina, Anirudha Majumdar, and Dorsa Sadigh · 2024
Later among the works it cites.
Collaborative Instance Navigation: Leveraging Agent Self-Dialogue to Minimize User Input
Francesco Taioli, Edoardo Zorzi, Gianni Franchi, Alberto Castellini, Alessandro Farinelli, Marco Cristani, and Yiming Wang · 2024
Later among the works it cites.
Mobileclip: Fast image-text models through multi-modal reinforced training
Pavan Kumar Anasosalu Vasu, Hadi Pouransari, Fartash Faghri, Raviteja Vemulapalli, and Oncel Tuzel · 2024
Later among the works it cites.
A survey of visual SLAM in dynamic environment: The evolution from geometric to semantic approaches
Yanan Wang, Yaobin Tian, Jiawei Chen, Kun Xu, and Xilun Ding · 2024
Later among the works it cites.
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
Abdelrhman Werby, Chenguang Huang, Martin Büchner, Abhinav Valada, and Wolfram Burgard · 2024
Later among the works it cites.
Embodied navigation with multi-modal information: A survey from tasks to methodology
Yuchen Wu, Pengcheng Zhang, Meiying Gu, Jin Zheng, and Xiao Bai · 2024
Later among the works it cites.
SED: A simple encoder-decoder for open-vocabulary semantic segmentation
Bin Xie, Jiale Cao, Jin Xie, Fahad Shahbaz Khan, and Yanwei Pang · 2024
Later among the works it cites.
EmbodiedSAM: Online Segment Any 3D Thing in Real Time
Xiuwei Xu, Huangxing Chen, Linqing Zhao, Ziwei Wang, Jie Zhou, and Jiwen Lu · 2024
Later among the works it cites.
3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning, 2024
Yuncong Yang, Han Yang, Jiachen Zhou, Peihao Chen, Hongxin Zhang, Yilun Du, and Chuang Gan · 2024
Later among the works it cites.
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
Naoki Yokoyama, Ram Ramrakhya, Abhishek Das, Dhruv Batra, and Sehoon Ha · 2024
Later among the works it cites.
BEV perception for autonomous driving: State of the art and future perspectives
Junhui Zhao, Jingyue Shi, and Li Zhuo · 2024
Later among the works it cites.
One Map to Find Them All: Real-time Open-Vocabulary Mapping for Zero-shot Multi-Object Navigation
Finn L Busch, Timon Homberger, Jesús Ortega-Peimbert, Quantao Yang, and Olov Andersson · 2025
Closest in time.
ASHiTA: Automatic Scene-grounded HIerarchical Task Analysis
Yun Chang, Leonor Fermoselle, Duy Ta, Bernadette Bucher, Luca Carlone, and Jiuguang Wang · 2025
Closest in time.
Semantic visual simultaneous localization and mapping: A survey
Kaiqi Chen, Junhao Xiao, Jialing Liu, Qiyi Tong, Heng Zhang, Ruyu Liu, Jianhua Zhang, Arash Ajoudani, and Shengyong Chen · 2025
Closest in time.
OpenLex3D: A New Evaluation Benchmark for Open-Vocabulary 3D Scene Representations
Christina Kassab, Sacha Morin, Martin Büchner, Matías Mattamala, Kumaraditya Gupta, Abhinav Valada, Liam Paull, and Maurice Fallon · 2025
Closest in time.
GeomGS: LiDAR-Guided Geometry-Aware Gaussian Splatting for Robot Localization
Jaewon Lee, Mangyu Kong, Minseong Park, and Euntai Kim · 2025
Closest in time.
Gaussnav: Gaussian splatting for visual navigation
Xiaohan Lei, Min Wang, Wengang Zhou, and Houqiang Li · 2025
Closest in time.
Sonia Raychaudhuri, Duy Ta, Katrina Ashton, Angel X. Chang, Jiuguang Wang, and Bernadette Bucher · 2025
Closest in time.
Semantic mapping techniques for indoor mobile robots: Review and prospect
Xu Song, Xuan Liang, and Zhou Huaidong · 2025
Closest in time.
S Talha Bukhari, Daniel Lawson, and Ahmed H Qureshi · 2025
Closest in time.
OpenIN: Open-Vocabulary Instance-Oriented Navigation in Dynamic Domestic Environments
Yujie Tang, Meiling Wang, Yinan Deng, Zibo Zheng, Jingchuan Deng, and Yufeng Yue · 2025
Closest in time.
Instruction-guided path planning with 3D semantic maps for vision-language navigation
Zehao Wang, Mingxiao Li, Minye Wu, Marie-Francine Moens, and Tinne Tuytelaars · 2025
Closest in time.
OVL-MAP: An Online Visual Language Map Approach for Vision-and-Language Navigation in Continuous Environments
Shuhuan Wen, Ziyuan Zhang, Yuxiang Sun, and Zhiwen Wang · 2025
Closest in time.
Bo Yang, Tri Minh Triet Pham, and Jinqiu Yang · 2025
Closest in time.
Lingfeng Zhang, Xiaoshuai Hao, Qinwen Xu, Qiang Zhang, Xinyao Zhang, Pengwei Wang, Jing Zhang, Zhongyuan Wang, Shanghang Zhang, and Renjing Xu · 2025
Closest in time.
Topological Mapping and Navigation in Real-World Environments
Collin Eugene Johnson · 2027
Closest in time.
Collaborative mobile robotics for semantic mapping: A survey
Abdessalem Achour, Hiba Al-Assaad, Yohan Dupuis, and Madeleine El Zaher · 2076
Closest in time.
An improved initialization method for monocular visual-inertial SLAM
Jun Cheng, Liyan Zhang, and Qihong Chen · 2079
Closest in time.
Constructing maps for autonomous robotics: An introductory conceptual overview
Peteris Racinskis, Janis Arents, and Modris Greitans · 2079
Closest in time.