Fetching the paper…
Reading the bibliography…
We describe a robotic learning system for autonomous exploration and navigation in diverse, open-world environments.
B. Kuipers and Y.-T. Byun, “A robot exploration and mapping strategy based on a semantic hierarchy of spatial representations,” Robotics and Autonomous Systems , 1991, special Issue Toward Learning Robots
1991
Earlier work this paper cites.
L. P. Kaelbling, “Learning to achieve goals,” in IJCAI . Citeseer, 1993, pp. 1094–1099
1993
Earlier work this paper cites.
B. Yamauchi, “A frontier-based approach for autonomous exploration,” in IEEE International Symposium on Computational Intelligence in Robotics and Automation (CIRA) , 1997
1997
Earlier work this paper cites.
N. Tishby, F. C. Pereira, and W. Bialek, “The information bottleneck method,” arXiv preprint physics/0004057 , 2000
2000
Earlier work this paper cites.
F. Bourgault, A. A. Makarenko, S. B. Williams, B. Grocholsky, and H. F. Durrant-Whyte, “Information based adaptive robotic exploration,” in IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , vol. 1, 2002, pp. 540–545 vol.1
2002
Earlier work this paper cites.
M. E. Taylor and P. Stone, “Cross-domain transfer for reinforcement learning,” in Proceedings of the 24th international conference on Machine learning , 2007, pp. 879–886
2007
Earlier work this paper cites.
T. Kollar and N. Roy, “Efficient optimization of information-theoretic exploration in slam,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , 2008
2008
Earlier work this paper cites.
D. Holz, N. Basilico, F. Amigoni, and S. Behnke, “A comparative evaluation of exploration strategies and heuristics to improve them,” 2011
2011
Earlier work this paper cites.
A. Lazaric, “Transfer in reinforcement learning: a framework and a survey,” in Reinforcement Learning . Springer, 2012, pp. 143–173
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
B. Charrow, S. Liu, V. Kumar, and N. Michael, “Information-theoretic mapping using cauchy-schwarz quadratic mutual information,” in IEEE International Conference on Robotics and Automation (ICRA) , 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
T. Schaul, D. Horgan, K. Gregor, and D. Silver, “Universal Value Function Approximators,” in International Conference on Machine Learning (ICML) , 2015
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell, “Curiosity-driven exploration by self-supervised prediction,” in International Conference on Machine Learning . PMLR, 2017, pp. 2778–2787
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov, “Proximal Policy Optimization Algorithms,” 2017
2017
Earlier work this paper cites.
A. G. Howard, M. Zhu, B. Chen, D. Kalenichenko, W. Wang, T. Weyand, M. Andreetto, and H. Adam, “MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications,” 2017
2017
Cited alongside, same era.
N. Savinov, A. Dosovitskiy, and V. Koltun, “Semi-Parametric Topological Memory for Navigation,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
A. Faust, K. Oslund, O. Ramirez, A. Francis, L. Tapia, M. Fiser, and J. Davidson, “Prm-rl: Long-range robotic navigation tasks by combining reinforcement learning and sampling-based planning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 5113–5120
2018
Cited alongside, same era.
R. Bajcsy, Y. Aloimonos, and J. K. Tsotsos, “Revisiting active perception,” Autonomous Robots , vol. 42, no. 2, pp. 177–196, 2018
2018
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
C. Colas, P. Fournier, M. Chetouani, O. Sigaud, and P.-Y. Oudeyer, “Curious: intrinsically motivated modular multi-goal reinforcement learning,” in International conference on machine learning . PMLR, 2019, pp. 1331–1340
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
A. Achille and S. Soatto, “Emergence of invariance and disentanglement in deep representations,” The Journal of Machine Learning Research , vol. 19, no. 1, pp. 1947–1980, 2018
2018
Cited alongside, same era.
W. Tabib, K. Goel, J. Yao, M. Dabhi, C. Boirum, and N. Michael, “Real-time information-theoretic exploration with gaussian mixture model maps.” in Robotics: Science and Systems , 2019
2019
Cited alongside, same era.
B. Eysenbach, R. R. Salakhutdinov, and S. Levine, “Search on the Replay Buffer: Bridging Planning and RL,” in Advances in Neural Information Processing Systems (NeurIPS) , 2019
2019
Cited alongside, same era.
X. Meng, N. Ratliff, Y. Xiang, and D. Fox, “Scaling Local Control to Large-Scale Topological Navigation,” in IEEE International Conference on Robotics and Automation (ICRA) , 2020
2020
Later among the works it cites.
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, and P. Sermanet, “Learning latent plans from play,” in Conference on Robot Learning . PMLR, 2020, pp. 1113–1132
2020
Later among the works it cites.
2020
Later among the works it cites.
D. S. Chaplot, D. Gandhi, S. Gupta, A. Gupta, and R. Salakhutdinov, “Learning to Explore using Active Neural SLAM,” in International Conference on Learning Representations (ICLR) , 2020
2020
Later among the works it cites.
D. Singh Chaplot, R. Salakhutdinov, A. Gupta, and S. Gupta, “Neural topological slam for visual navigation,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Later among the works it cites.
D. S. Chaplot, H. Jiang, S. Gupta, and A. Gupta, “Semantic curiosity for active visual learning,” in ECCV , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
N. Rhinehart, R. McAllister, and S. Levine, “Deep imitative models for flexible inference, planning, and control,” in International Conference on Learning Representations , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
S. Pitis, H. Chan, S. Zhao, B. Stadie, and J. Ba, “Maximum entropy gain exploration for long horizon multi-goal reinforcement learning,” in International Conference on Machine Learning . PMLR, 2020, pp. 7750–7761
2020
Later among the works it cites.
K. Liu, T. Kurutach, C. Tung, P. Abbeel, and A. Tamar, “Hallucinative topological memory for zero-shot visual planning,” in International Conference on Machine Learning . PMLR, 2020, pp. 6259–6270
2020
Later among the works it cites.
G. Kahn, P. Abbeel, and S. Levine, “BADGR: An Autonomous Self-Supervised Learning-Based Navigation System,” 2020
2020
Later among the works it cites.
E. Wijmans, A. Kadian, A. Morcos, S. Lee, I. Essa, D. Parikh, M. Savva, and D. Batra, “DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames,” in International Conference on Learning Representations (ICLR) , 2020
2020
Later among the works it cites.
D. Shah, B. Eysenbach, G. Kahn, N. Rhinehart, and S. Levine, “ViNG: Learning Open-World Navigation with Visual Goals,” in IEEE International Conference on Robotics and Automation (ICRA) , 2021
2021
Closest in time.
N. Yokoyama, S. Ha, and D. Batra, “Success weighted by completion time: A dynamics-aware evaluation criteria for embodied navigation,” 2021
2021
Closest in time.
S. Gamrian and Y. Goldberg, “Transfer learning for related reinforcement learning tasks via image-to-image translation,” in International Conference on Machine Learning . PMLR, 2019, pp. 2063–2072
2072
Closest in time.