Fetching the paper…
Reading the bibliography…
The ability to navigate from visual observations in unfamiliar environments is a core component of intelligent agents and an ongoing challenge for Deep Reinforcement Learning (RL).
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, Patrick Haffner, et al · 1998
Earlier work this paper cites.
Transfer learning for reinforcement learning domains: A survey
Matthew E. Taylor and Peter Stone · 2009
Earlier work this paper cites.
A survey on transfer learning
Sinno Jialin Pan and Qiang Yang · 2010
Earlier work this paper cites.
Geo-localization of street views with aerial image databases
Mayank Bansal, Harpreet S. Sawhney, Hui Cheng, and Kostas Daniilidis · 2011
Earlier work this paper cites.
Do deep nets really need to be deep?
Jimmy Ba and Rich Caruana · 2014
Earlier work this paper cites.
Planar structure matching under projective uncertainty for geolocation
Ang Li, Vlad I. Morariu, and Larry S. Davis · 2014
Earlier work this paper cites.
Fitnets: Hints for thin deep nets
Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou, Antoine Chassang, Carlo Gatta, and Yoshua Bengio · 2014
Earlier work this paper cites.
Visual localization within lidar maps for automated urban driving
Ryan W Wolcott and Ryan M Eustice · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Earlier work this paper cites.
Learning deep representations for ground-to-aerial geolocalization
Tsung-Yi Lin, Yin Cui, Serge Belongie, and James Hays · 2015
Earlier work this paper cites.
Geo-localization using volumetric representations of overhead imagery
Özge Can Özcanli, Yi Dong, and Joseph L. Mundy · 2015
Earlier work this paper cites.
Transfer deep reinforcement learning in 3d environments: An empirical study
Devendra Singh Chaplot, Guillaume Lample, Kanthashree Mysore Sathyendra, and Ruslan Salakhutdinov · 2016
Earlier work this paper cites.
Cross modal distillation for supervision transfer
Saurabh Gupta, Judy Hoffman, and Jitendra Malik · 2016
Earlier work this paper cites.
Learning with side information through modality hallucination
Judy Hoffman, Saurabh Gupta, and Trevor Darrell · 2016
Earlier work this paper cites.
Control of memory, active perception, and action in Minecraft
Junhyuk Oh, Valliappa Chockalingam, Satinder Singh, and Honglak Lee · 2016
Earlier work this paper cites.
One-shot reinforcement learning for robot navigation with interactive replay
Jake Bruce, Niko Sünderhauf, Piotr Mirowski, Raia Hadsell, and Michael Milford · 2017
Cited alongside, same era.
Gated-attention architectures for task-oriented language grounding
Devendra Singh Chaplot, Kanthashree Mysore Sathyendra, Rama Kumar Pasumarthi, Dheeraj Rajagopal, and Ruslan Salakhutdinov · 2017
Cited alongside, same era.
Unifying map and landmark based representations for visual navigation
Saurabh Gupta, David Fouhey, Sergey Levine, and Jitendra Malik · 2017
Cited alongside, same era.
Schema networks: Zero-shot transfer with a generative causal model of intuitive physics
Ken Kansky, Tom Silver, David A Mély, Mohamed Eldawy, Miguel Lázaro-Gredilla, Xinghua Lou, Nimrod Dorfman, Szymon Sidor, Scott Phoenix, and Dileep George · 2017
Cited alongside, same era.
Learned contextual feature reweighting for image geo-localization
Hyo Jin Kim, Enrique Dunn, and Jan-Michael Frahm · 2017
Emergence of grid-like representations by training recurrent neural networks to perform spatial localization
Christopher J Cueva and Xue-Xin Wei · 2018
Later among the works it cites.
Human Spatial Navigation
Arne D Ekstrom, Hugo J Spiers, Véronique D Bohbot, and R Shayna Rosenbaum · 2018
Later among the works it cites.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
Lasse Espeholt, Hubert Soyer, Remi Munos, Karen Simonyan, Volodymir Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, et al · 2018
Later among the works it cites.
Modality distillation with multiple stream networks for action recognition
Nuno C Garcia, Pietro Morerio, and Vittorio Murino · 2018
Later among the works it cites.
Graph distillation for action detection with privileged modalities
Zelun Luo, Jun-Ting Hsieh, Lu Jiang, Juan Carlos Niebles, and Li Fei-Fei · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Ai2-thor: An interactive 3d environment for visual ai
Eric Kolve, Roozbeh Mottaghi, Daniel Gordon, Yuke Zhu, Abhinav Gupta, and Ali Farhadi · 2017
Cited alongside, same era.
Learning from noisy labels with distillation
Yuncheng Li, Jianchao Yang, Yale Song, Liangliang Cao, Jiebo Luo, and Li-Jia Li · 2017
Cited alongside, same era.
Learning to navigate in complex environments
Piotr Mirowski, Razvan Pascanu, Fabio Viola, Hubert Soyer, Andrew J Ballard, Andrea Banino, Misha Denil, Ross Goroshin, Laurent Sifre, Koray Kavukcuoglu, et al · 2017
Cited alongside, same era.
Virtual-to-real deep reinforcement learning: Continuous control of mobile robots for mapless navigation
Lei Tai, Giuseppe Paolo, and Ming Liu · 2017
Cited alongside, same era.
Cross-view image matching for geo-localization in urban environments
Yicong Tian, Chen Chen, and Mubarak Shah · 2017
Cited alongside, same era.
Deep reinforcement learning with successor features for navigation across similar environments
Jingwei Zhang, Jost Tobias Springenberg, Joschka Boedecker, and Wolfram Burgard · 2017
Cited alongside, same era.
Neural slam: Learning to explore with external memory
Jingwei Zhang, Lei Tai, Joschka Boedecker, Wolfram Burgard, and Ming Liu · 2017
Cited alongside, same era.
Piotr W. Mirowski, Matthew Koichi Grimes, Mateusz Malinowski, Karl Moritz Hermann, Keith Anderson, Denis Teplyashin, Karen Simonyan, Koray Kavukcuoglu, Andrew Zisserman, and Raia Hadsell · 2018
Later among the works it cites.
Kaichun Mo, Haoxiang Li, Zhe Lin, and Joon-Young Lee · 2018
Later among the works it cites.
Visual representations for semantic target driven navigation
Arsalan Mousavian, Alexander Toshev, Marek Fiser, Jana Kosecka, and James Davidson · 2018
Later among the works it cites.
Global pose estimation with an attention-based recurrent network
Emilio Parisotto, Devendra Singh Chaplot, Jian Zhang, and Ruslan Salakhutdinov · 2018
Later among the works it cites.
Airsim: High-fidelity visual and physical simulation for autonomous vehicles
Shital Shah, Debadeepta Dey, Chris Lovett, and Ashish Kapoor · 2018
Later among the works it cites.
Xin Wang, Qiuyuan Huang, Asli Celikyilmaz, Jianfeng Gao, Dinghan Shen, Yuan-Fang Wang, William Yang Wang, and Lei Zhang · 2018
Later among the works it cites.
Unsupervised predictive memory in a goal-directed agent
Greg Wayne, Chia-Chun Hung, David Amos, Mehdi Mirza, Arun Ahuja, Agnieszka Grabska-Barwinska, Jack W. Rae, Piotr Mirowski, Joel Z. Leibo, Adam Santoro, Mevlana Gemici, Malcolm Reynolds, Tim Harley, Josh Abramson, Shakir Mohamed, Danilo Jimenez Rezende, David Saxton, Adam Cain, Chloe Hillier, David Silver, Koray Kavukcuoglu, Matthew Botvinick, Demis Hassabis, and Timothy P. Lillicrap · 2018
Later among the works it cites.
Deep mutual learning
Ying Zhang, Tao Xiang, Timothy M Hospedales, and Huchuan Lu · 2018
Later among the works it cites.
Learning to follow directions in street view
Karl Moritz Hermann, Mateusz Malinowski, P. Mirowski, Andras Banki-Horvath, Keith Anderson, and Raia Hadsell · 2019
Closest in time.
The streetlearn environment and dataset
Piotr Mirowski, Andras Banki-Horvath, Keith Anderson, Denis Teplyashin, Karl Moritz Hermann, Mateusz Malinowski, Matthew Koichi Grimes, Karen Simonyan, Koray Kavukcuoglu, Andrew Zisserman, et al · 2019
Closest in time.
Improved knowledge distillation via teacher assistant: Bridging the gap between student and teacher
Seyed-Iman Mirzadeh, Mehrdad Farajtabar, Ang Li, and Hassan Ghasemzadeh · 2019
Closest in time.