Fetching the paper…
Reading the bibliography…
We consider learning based methods for visual localization that do not require the construction of explicit maps in the form of point clouds or voxels.
Bundle adjustment—a modern synthesis
B. Triggs, P. F. McLauchlan, R. I. Hartley, and A. W. Fitzgibbon · 1999
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
D. G. Lowe · 2004
Earlier work this paper cites.
Probabilistic robotics
S. Thrun, W. Burgard, and D. Fox · 2005
Earlier work this paper cites.
A tutorial on graph-based slam
G. Grisetti, R. Kummerle, C. Stachniss, and W. Burgard · 2010
Earlier work this paper cites.
Worldwide pose estimation using 3d point clouds
Y. Li, N. Snavely, D. Huttenlocher, and P. Fua · 2012
Earlier work this paper cites.
Slam++: Simultaneous localisation and mapping at the level of objects
R. F. Salas-Moreno, R. A. Newcombe, H. Strasdat, P. H. Kelly, and A. J. Davison · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Posenet: A convolutional network for real-time 6-dof camera relocalization
A. Kendall, M. Grimes, and R. Cipolla · 2015
Earlier work this paper cites.
Orb-slam: a versatile and accurate monocular slam system
R. Mur-Artal, J. M. M. Montiel, and J. D. Tardos · 2015
Earlier work this paper cites.
The return of the gating network: combining generative models and discriminative training in natural image priors
D. Rosenbaum and Y. Weiss · 2015
Earlier work this paper cites.
Camera pose voting for large-scale image-based localization
B. Zeisl, T. Sattler, and M. Pollefeys · 2015
Earlier work this paper cites.
3d semantic parsing of large-scale indoor spaces
I. Armeni, O. Sener, A. R. Zamir, H. Jiang, I. Brilakis, M. Fischer, and S. Savarese · 2016
Earlier work this paper cites.
Past, present, and future of simultaneous localization and mapping: Toward the robust-perception age
C. Cadena, L. Carlone, H. Carrillo, Y. Latif, D. Scaramuzza, J. Neira, I. Reid, and J. J. Leonard · 2016
Earlier work this paper cites.
The malmo platform for artificial intelligence experimentation
M. Johnson, K. Hofmann, T. Hutton, and D. Bignell · 2016
Earlier work this paper cites.
Learning to navigate in complex environments
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. J. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, et al · 2016
Earlier work this paper cites.
Matching networks for one shot learning
O. Vinyals, C. Blundell, T. Lillicrap, D. Wierstra, et al · 2016
Cited alongside, same era.
Generic 3d representation via pose estimation and matching
A. R. Zamir, T. Wekel, P. Agrawal, C. Wei, J. Malik, and S. Savarese · 2016
Cited alongside, same era.
Matterport3d: Learning from rgb-d data in indoor environments
A. Chang, A. Dai, T. Funkhouser, M. Halber, M. Nießner, M. Savva, S. Song, A. Zeng, and Y. Zhang · 2017
Cited alongside, same era.
Vidloc: 6-dof video-clip relocalization. arxiv preprint
R. Clark, S. Wang, A. Markham, N. Trigoni, and H. Wen · 2017
Cited alongside, same era.
Cognitive mapping and planning for visual navigation
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik · 2017
Cited alongside, same era.
Demon: Depth and motion network for learning monocular stereo
B. Ummenhofer, H. Zhou, J. Uhrig, N. Mayer, E. Ilg, A. Dosovitskiy, and T. Brox · 2017
Later among the works it cites.
Image-based localization using lstms for structured feature correlation
F. Walch, C. Hazirbas, L. Leal-Taixe, T. Sattler, S. Hilsenbeck, and D. Cremers · 2017
Later among the works it cites.
J. Zhang, L. Tai, J. Boedecker, W. Burgard, and M. Liu · 2017
Later among the works it cites.
Unsupervised learning of depth and ego-motion from video
T. Zhou, M. Brown, N. Snavely, and D. G. Lowe · 2017
Later among the works it cites.
Vector-based navigation using grid-like representations in artificial agents
A. Banino, C. Barry, B. Uria, C. Blundell, T. Lillicrap, P. Mirowski, A. Pritzel, M. J. Chadwick, T. Degris, et al · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unifying map and landmark based representations for visual navigation
S. Gupta, D. Fouhey, S. Levine, and J. Malik · 2017
Cited alongside, same era.
Semanticfusion: Dense 3d semantic mapping with convolutional neural networks
J. McCormac, A. Handa, A. Davison, and S. Leutenegger · 2017
Cited alongside, same era.
Neural map: Structured memory for deep reinforcement learning
E. Parisotto and R. Salakhutdinov · 2017
Cited alongside, same era.
Few-shot autoregressive density estimation: Towards learning to learn distributions
S. Reed, Y. Chen, T. Paine, A. v. d. Oord, S. Eslami, D. J. Rezende, O. Vinyals, and N. de Freitas · 2017
Cited alongside, same era.
Efficient & effective prioritized matching for large-scale image-based localization
T. Sattler, B. Leibe, and L. Kobbelt · 2017
Cited alongside, same era.
Are large-scale 3d models really necessary for accurate visual localization?
T. Sattler, A. Torii, J. Sivic, M. Pollefeys, H. Taira, M. Okutomi, and T. Pajdla · 2017
Cited alongside, same era.
Minos: Multimodal indoor simulator for navigation in complex environments
M. Savva, A. X. Chang, A. Dosovitskiy, T. Funkhouser, and V. Koltun · 2017
Cited alongside, same era.
M. Bloesch, J. Czarnowski, R. Clark, S. Leutenegger, and A. J. Davison · 2018
Closest in time.
C. J. Cueva and X.-X. Wei · 2018
Closest in time.
Neural scene representation and rendering
S. A. Eslami, D. J. Rezende, F. Besse, F. Viola, A. S. Morcos, M. Garnelo, A. Ruderman, A. A. Rusu, I. Danihelka, et al · 2018
Closest in time.
Mapnet: An allocentric spatial memory for mapping environments
J. F. Henriques and A. Vedaldi · 2018
Closest in time.
Global pose estimation with an attention-based recurrent network
E. Parisotto, D. S. Chaplot, J. Zhang, and R. Salakhutdinov · 2018
Closest in time.
Semi-parametric topological memory for navigation
N. Savinov, A. Dosovitskiy, and V. Koltun · 2018
Closest in time.
Ba-net: Dense bundle adjustment network
C. Tang and P. Tan · 2018
Closest in time.
Multi-view consistency as supervisory signal for learning shape and pose prediction
S. Tulsiani, A. A. Efros, and J. Malik · 2018
Closest in time.
Unsupervised predictive memory in a goal-directed agent
G. Wayne, C.-C. Hung, D. Amos, M. Mirza, A. Ahuja, A. Grabska-Barwinska, J. Rae, P. Mirowski, J. Z. Leibo, A. Santoro, et al · 2018
Closest in time.