Fetching the paper…
Reading the bibliography…
An important goal of research in Deep Reinforcement Learning in mobile robotics is to train agents capable of solving complex tasks, which require a high level of scene understanding and reasoning from an egocentric perspective.
Williams, R.J.: Simple statistical gradient following algorithms for connectionist reinforcement learning. Machine Learning pp. 229–256 (1992)
1992
Earlier work this paper cites.
Hochreiter, S., Schmidhuber, J.: Long Short-Term Memory. Neural Computation 9
1997
Earlier work this paper cites.
Lecun, Y., Eon Bottou, L., Bengio, Y., Haaner, P.: Gradient-Based Learning Applied to Document Recognition. IEEE 86
1998
Earlier work this paper cites.
Van Der Maaten, L., Hinton, G.: Visualizing Data using t-SNE. Journal of Machine Learning Research 9
2008
Earlier work this paper cites.
Civera, J., Galvez-Lopez, D., Riazuelo, L., Tardós, J.D., Montiel, J.M.M.: Towards semantic slam using a monocular camera. In: IROS (2011)
2011
Earlier work this paper cites.
Chung, J., Gulcehre, C., Cho, K., Bengio, Y.: Gated Feedback Recurrent Neural Networks. In: ICML (2015)
2015
Earlier work this paper cites.
Lillicrap, T.P., Hunt, J.J., Pritzel, A., Heess, N., Erez, T., Tassa, Y., Silver, D., Wierstra, D.: Continuous Control with Deep Reinforcement Learning. arxiv pre-print 1509.02971 (2015)
2015
Earlier work this paper cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., Hassabis, D.: Human-level control through deep reinforcement learning. Nature (2015)
2015
Earlier work this paper cites.
Beattie, C., Leibo, J.Z., Teplyashin, D., Ward, T., Wainwright, M., Küttler, H., Lefrancq, A., Green, S., Valdés, V., Sadik, A., Schrittwieser, J., Anderson, K., York, S., Cant, M., Cain, A., Bolton, A., Gaffney, S., King, H., Hassabis, D., Legg, S., Petersen, S.: DeepMind Lab. arxiv pre-print 1612.03801 (2016)
2016
Earlier work this paper cites.
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., Zaremba, W.: OpenAI Gym. arxiv pre-print 1606.01540 (2016)
2016
Earlier work this paper cites.
Levine, S., Finn, C., Darrell, T., Abbeel, P.: End-to-End Training of Deep Visuomotor Policies. Journal of Machine Learning Research 17
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Silver, D., Huang, A., Maddison, C.J., Guez, A., Sifre, L., van den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., Dieleman, S., Grewe, D., Nham, J., Kalchbrenner, N., Sutskever, I., Lillicrap, T., Leach, M., Kavukcuoglu, K., Graepel, T., Hassabis, D.: Mastering the game of Go with deep neural networks and tree search. Nature 529
2016
Earlier work this paper cites.
Wang, J.X., Kurth-Nelson, Z., Tirumala, D., Soyer, H., Leibo, J.Z., Munos, R., Blundell, C., Kumaran, D., Botvinick, M.: Learning to reinforcement learn. arxiv pre-print 1611.05763 (2016)
2016
Earlier work this paper cites.
Dhariwal, P., Hesse, C., Klimov, O., Nichol, A., Plappert, M., Radford, A., Schulman, J., Sidor, S., Wu, Y., Zhokhov, P.: Openai baselines. https://github.com/openai/baselines (2017)
2017
Cited alongside, same era.
Dosovitskiy, A., Ros, G., Codevilla, F., Lopez, A., Koltun, V.: CARLA: An open urban driving simulator. In: CoRL. pp. 1–16 (2017)
2017
Cited alongside, same era.
Jaderberg, M., Mnih, V., Czarnecki, W.M., Schaul, T., Leibo, J.Z., Silver, D., Kavukcuoglu, K.: Reinforcement learning with unsupervised auxiliary tasks. In: ICLR (2017)
2017
Cited alongside, same era.
Kempka, M., Wydmuch, M., Runc, G., Toczek, J., Jaskowski, W.: ViZDoom: A Doom-based AI research platform for visual reinforcement learning. IEEE Conference on Computatonal Intelligence and Games, CIG (2017)
2017
Cited alongside, same era.
Kolve, E., Mottaghi, R., Gordon, D., Zhu, Y., Gupta, A., Farhadi, A.: AI2-THOR: An Interactive 3D Environment for Visual AI. arXiv (2017)
Wu, Y., Mansimov, E., Liao, S., Grosse, R., Ba, J.: Scalable trust-region method for deep reinforcement learning using Kronecker-factored approximation. In: NIPS (2017)
2017
Later among the works it cites.
Anderson, P., Wu, Q., Teney, D., Bruce, J., Johnson, M., Sünderhauf, N., Reid, I., Gould, S., den Hengel, A.: Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments. In: CVPR (2018)
2018
Later among the works it cites.
Brodeur, S., Perez, E., Anand, A., Golemo, F., Celotti, L., Strub, F., Rouat, J., Larochelle, H., Courville, A.: HoME: a Household Multimodal Environment. In: ICLR (2018)
2018
Later among the works it cites.
Das, A., Gkioxari, G., Lee, S., Parikh, D., Batra, D.: Neural Modular Control for Embodied Question Answering. In: CVPR (2018)
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
Lample, G., Chaplot, D.S.: Playing FPS games with deep reinforcement learning. In: AAAI (2017)
2017
Cited alongside, same era.
Mirowski, P., Pascanu, R., Viola, F., Soyer, H., Ballard, A.J., Banino, A., Denil, M., Goroshin, R., Sifre, L., Kavukcuoglu, K., Kumaran, D., Hadsell, R.: Learning to Navigate in Complex Environments. In: ICLR (2017)
2017
Cited alongside, same era.
Moritz, P., Nishihara, R., Wang, S., Tumanov, A., Liaw, R., Liang, E., Elibol, M., Yang, Z., Paul, W., Jordan, M.I.: Ray: A Distributed Framework for Emerging AI Applications. In: USENIX Symposium on Operating Systems Design and Implementation (2017)
2017
Cited alongside, same era.
Müller, M., Casser, V., Lahoud, J., Smith, N., Ghanem, B.: Sim4CV: A Photo-Realistic Simulator for Computer Vision Applications. arXiv (Aug 2017)
2017
Cited alongside, same era.
Parisotto, E., Salakhutdinov, R.: Neural map: Structured memory for deep reinforcement learning. arxiv pre-print 1702.08360 (2017)
2017
Cited alongside, same era.
Savva, M., Chang, A.X., Dosovitskiy, A., Funkhouser, T., Koltun, V.: MINOS: Multimodal Indoor Simulator for Navigation in Complex Environments. arxiv pre-print 1712.03931 (2017)
2017
Cited alongside, same era.
Schaarschmidt, M., Kuhnle, A., Fricke, K.: Tensorforce: A tensorflow library for applied reinforcement learning (2017)
2017
Cited alongside, same era.
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., Legg, S., Kavukcuoglu, K.: IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures. In: ICML (2018)
2018
Later among the works it cites.
Henderson, P., Islam, R., Bachman, P., Pineau, J., Precup, D., Meger, D.: Deep reinforcement learning that matters. In: AAAI (2018)
2018
Later among the works it cites.
Kayalibay, B., Mirchev, A., Soelch, M., Van Der Smagt, P., Bayer, J.: Navigation and planning in latent maps. In: FAIM workshop “Prediction and Generative Modeling in Reinforcement Learning” (2018)
2018
Later among the works it cites.
Kostrikov, I.: Pytorch implementations of reinforcement learning algorithms. https://github.com/ikostrikov/pytorch-a2c-ppo-acktr (2018)
2018
Later among the works it cites.
OpenAI: Openai five. https://blog.openai.com/openai-five/ (2018)
2018
Later among the works it cites.
Stooke, A., Abbeel, P.: Accelerated methods for deep reinforcement learning. arxiv pre-print 1803.02811 (2018)
2018
Later among the works it cites.
Xia, F., R. Zamir, A., He, Z.Y., Sax, A., Malik, J., Savarese, S.: Gibson env: real-world perception for embodied agents. In: CVPR (2018)
2018
Later among the works it cites.
2018
Later among the works it cites.
Manolis Savva*, Abhishek Kadian*, Oleksandr Maksymets*, Batra, D.: Habitat: A platform for embodied ai research. arXiv (2019)
2019
Closest in time.