Collaborative multi-robot exploration
Wolfram Burgard, Mark Moors, Dieter Fox, Reid Simmons, and Sebastian Thrun · 2000
Earlier work this paper cites.
A mathematician looks at wolfram’s new kind of science
Lawrence Gray, A New, et al · 2003
Earlier work this paper cites.
Exploration strategies based on multi-criteria decision making for searching environments in rescue operations
Nicola Basilico and Francesco Amigoni · 2011
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling · 2013
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Earlier work this paper cites.
Fractal trajectories for online non-uniform aerial coverage
Seyed Abbas Sadat, Jens Wawerla, and Richard Vaughan · 2015
Earlier work this paper cites.
Real-time adaptive multi-robot exploration with application to underwater map construction
Athanasios Ch Kapoutsis, Savvas A Chatzichristofis, Lefteris Doitsidis, Joao Borges de Sousa, Jose Pinto, Jose Braga, and Elias B Kosmatopoulos · 2016
Earlier work this paper cites.
Distributed coverage estimation and control for multirobot persistent tasks
José Manuel Palacios-Gasós, Eduardo Montijano, Carlos Sagüés, and Sergio Llorente · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Earlier work this paper cites.
Openai gym
Original
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Extending the openai gym for robotics: a toolkit for reinforcement learning using ros and gazebo
Original
Iker Zamora, Nestor Gonzalez Lopez, Victor Mayoral Vilches, and Alejandro Hernandez Cordero · 2016
Earlier work this paper cites.