Fetching the paper…
Reading the bibliography…
The high probability of hardware failures prevents many advanced robots (e.g., legged robots) from being confidently deployed in real-world situations (e.g., post-disaster rescue).
J. Quiñonero-Candela, C. E. Rasmussen, A unifying view of sparse approximate Gaussian process regression, Journal of Machine Learning Research 6 (2005) 1939–1959
1959
Earlier work this paper cites.
L. E. Kavraki, P. Svestka, J.-C. Latombe, M. H. Overmars, Probabilistic roadmaps for path planning in high-dimensional configuration spaces, IEEE Trans. on Robotics and Automation 12 (4) (1996) 566–580
1996
Earlier work this paper cites.
R. S. Sutton, A. G. Barto, Reinforcement learning: An introduction, MIT press, 1998
1998
Earlier work this paper cites.
S. M. LaValle, Rapidly-exploring random trees: A new tool for path planning, Tech. Rep. TR 98-11, Computer Science Dept., Iowa State University (1998)
1998
Earlier work this paper cites.
A. J. Ijspeert, J. Nakanishi, S. Schaal, Learning attractor landscapes for learning motor primitives, in: Proc. of NIPS, 2002, pp. 1547–1554
2002
Earlier work this paper cites.
M. Blanke, J. Schröder, Diagnosis and fault-tolerant control, Vol. 115, Springer, 2003
2003
Earlier work this paper cites.
V. Verma, G. Gordon, R. Simmons, S. Thrun, Real-time fault diagnosis, IEEE Robotics & Automation Magazine 11 (2) (2004) 56–66
2004
Earlier work this paper cites.
R. Tedrake, T. W. Zhang, H. S. Seung, Stochastic policy gradient reinforcement learning on a simple 3D biped, in: Proc. of IROS, 2004, pp. 2849–2854
2004
Earlier work this paper cites.
J. Carlson, R. R. Murphy, How UGVs physically fail in the field, IEEE Trans. on Robotics 21 (3) (2005) 423–437
2005
Earlier work this paper cites.
R. Isermann, Fault-diagnosis systems: an introduction from fault detection to fault tolerance, Springer Science & Business Media, 2006
2006
Earlier work this paper cites.
J. C. Bongard, V. Zykov, H. Lipson, Resilient machines through continuous self-modeling, Science 314 (5802) (2006) 1118–1121
2006
Earlier work this paper cites.
S. M. LaValle, Planning algorithms, Cambridge University Press, 2006
2006
Earlier work this paper cites.
C. E. Rasmussen, C. K. I. Williams, Gaussian processes for machine learning, MIT Press, 2006
2006
Earlier work this paper cites.
2006
Earlier work this paper cites.
D. J. Lizotte, T. Wang, M. H. Bowling, D. Schuurmans, Automatic gait optimization with gaussian process regression, in: Proc. of IJCAI, 2007, pp. 944–949
2007
Earlier work this paper cites.
T. Cazenave, N. Jouandeau, On the parallelization of UCT, in: Proc. of the Computer Games Workshop, 2007, pp. 93–101
2007
Earlier work this paper cites.
F. Corbato, On Building Systems That Will Fail, ACM Turing award lectures 34 (9) (2007) 72–81
2007
Earlier work this paper cites.
G. Chaslot, S. Bakkes, I. Szita, P. Spronck, Monte-carlo tree search: A new framework for game AI, in: Proc. of AIIDE, 2008, pp. 216–217
2008
Earlier work this paper cites.
J. Peters, S. Schaal, Reinforcement learning of motor skills with policy gradients, Neural Networks 21 (4) (2008) 682–697
2008
Earlier work this paper cites.
P. Rolet, M. Sebag, O. Teytaud, Boosting active learning to optimality: A tractable monte-carlo, billiard-based algorithm, in: Proc. of ECML, 2009, pp. 302–317
2009
Earlier work this paper cites.
J. Peters, K. Mülling, Y. Altun, Relative entropy policy search, in: Proc. of AAAI, 2010, pp. 1607–1612
2010
Cited alongside, same era.
K. Mostafa, C. Tsai, I. Her, Alternative gaits for multiped robots with leg failures to retain maneuverability, International Journal of Advanced Robotic Systems 7 (4) (2010) 31
2010
Cited alongside, same era.
D. Silver, J. Veness, Monte-carlo planning in large POMDPs, in: Proc. of NIPS, 2010, pp. 2164–2172
2010
Cited alongside, same era.
J.-B. Mouret, S. Doncieux, Sferes v2
2010
Cited alongside, same era.
M. P. Deisenroth, C. E. Rasmussen, D. Fox, Learning to control a low-cost manipulator using data-efficient reinforcement learning, in: Robotics: Science & Systems (RSS), 2011, pp. 57–64
2011
Cited alongside, same era.
A. Baranes, P.-Y. Oudeyer, Active learning of inverse models with intrinsically motivated goal exploration in robots, Robotics and Autonomous Systems 61 (1) (2013) 49–73
2013
Later among the works it cites.
C. Atkeson, et al., No falls, no resets: Reliable humanoid behavior in the DARPA robotics challenge, in: Proc. of Humanoids, 2015, pp. 623–630
2015
Later among the works it cites.
A. Cully, J. Clune, D. Tarapore, J.-B. Mouret, Robots that can adapt like animals, Nature 521 (7553) (2015) 503–507
2015
Later among the works it cites.
G. Ren, W. Chen, S. Dasgupta, C. Kolodziejski, F. Wörgötter, P. Manoonpong, Multiple chaotic central pattern generators with learning for legged locomotion and malfunction compensation, Information Sciences 294 (2015) 666–682
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Nguyen-Tuong, J. Peters, Model learning for robot control: a survey, Cognitive Processing 12 (4) (2011) 319–340
2011
Cited alongside, same era.
A. Couëtoux, J.-B. Hoock, N. Sokolovska, O. Teytaud, N. Bonnard, Continuous upper confidence trees, in: Proc. of LION, 2011, pp. 433–445
2011
Cited alongside, same era.
A. Couetoux, M. Milone, M. Brendel, H. Doghmen, M. Sebag, O. Teytaud, Continuous rapid action value estimates, in: Proc. of ACML, 2011, p. 19–31
2011
Cited alongside, same era.
E. Guizzo, Fukushima robot operator writes tell-all blog, in: IEEE Spectrum, 2011, URL: http://spectrum.ieee.org/automaton/robotics/industrial-robots/fukushima-robot-operator-diaries
2011
Cited alongside, same era.
J.-B. Mouret, S. Doncieux, Encouraging behavioral diversity in evolutionary robotics: an empirical study, Evolutionary Computation 20 (1) (2012) 91–133
2012
Cited alongside, same era.
T. Hester, M. Quinlan, P. Stone, RTMBA: A real-time model-based reinforcement learning architecture for robot control, in: Proc. of ICRA, IEEE, 2012, pp. 85–90
2012
Cited alongside, same era.
C. B. Browne, et al., A survey of monte carlo tree search methods, IEEE Trans. on Computational Intelligence and AI in Games 4 (1) (2012) 1–43
2012
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, D. Hassabis, Human-level control through deep reinforcement learning, Nature 518 (7540) (2015) 529–533
2015
Later among the works it cites.
M. P. Deisenroth, D. Fox, C. E. Rasmussen, Gaussian processes for data-efficient learning in robotics and control, IEEE Trans. Pattern Anal. Mach. Intell. 37 (2) (2015) 408–423
2015
Later among the works it cites.
F. Nori, S. Traversaro, J. Eljaik, F. Romano, A. Del Prete, D. Pucci, iCub whole-body control through force regulation on rigid non-coplanar contacts, Frontiers in Robotics and AI 2 (2015) 6
2015
Later among the works it cites.
R. Calandra, A. Seyfarth, J. Peters, M. Deisenroth, Bayesian optimization for learning gaits under uncertainty, Annals of Mathematics and Artificial Intelligence 76 (2015) 5–23
2015
Later among the works it cites.
J. Schulman, S. Levine, P. Moritz, M. I. Jordan, P. Abbeel, Trust region policy optimization, in: Proc. of ICML, 2015, pp. 1889–1897
2015
Later among the works it cites.
A. Nguyen, J. Yosinski, J. Clune, Deep neural networks are easily fooled: High confidence predictions for unrecognizable images, in: Proc. of CVPR, 2015, pp. 427–436
2015
Later among the works it cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel, D. Hassabis, Mastering the game of Go with deep neural networks and tree search, Nature 529 (7587) (2016) 484–489
2016
Closest in time.
B. Shahriari, K. Swersky, Z. Wang, R. P. Adams, N. de Freitas, Taking the human out of the loop: A review of bayesian optimization, Proceedings of the IEEE 104 (1) (2016) 148–175
2016
Closest in time.
M. Duarte, J. Gomes, S. M. Oliveira, A. L. Christensen, EvoRBC: evolutionary repertoire-based control for robots with arbitrary locomotion complexity, in: Proc. of GECCO, 2016, pp. 93–100
2016
Closest in time.
J. K. Pugh, L. B. Soros, K. O. Stanley, Quality diversity: A new frontier for evolutionary computation, Frontiers in Robotics and AI 3 (2016) 40, doi: 10.3389/frobt.2016.00040
2016
Closest in time.
A. Nguyen, J. Yosinski, J. Clune, Understanding Innovation Engines: Automated Creativity and Improved Stochastic Optimization via Deep Learning, Evolutionary Computation 24 (2016) 545–572
2016
Closest in time.
J. Lehman, S. Risi, J. Clune, Creative generation of 3D objects with deep learning and innovation engines, in: Proc. of the 7th Intern. Conf. on Comput. Creativity, 2016, pp. 180–187
2016
Closest in time.
Y. Gal, Z. Ghahramani, Dropout as a Bayesian approximation: Representing model uncertainty in deep learning, in: Proc. of ICML, 2016, pp. 1050–1059
2016
Closest in time.
M. DeDonato, F. Polido, K. Knoedler, B. P. Babu, N. Banerjee, C. P. Bove, X. Cui, R. Du, P. Franklin, J. P. Graff, et al., Team WPI-CMU: Achieving Reliable Humanoid Behavior in the DARPA Robotics Challenge, Journal of Field Robotics 34 (2) (2017) 381–399
2017
Closest in time.
K. Chatzilygeroudis, R. Rama, R. Kaushik, D. Goepp, V. Vassiliades, J.-B. Mouret, Black-Box Data-efficient Policy Search for Robotics, in: Proc. of IROS, 2017
2017
Closest in time.
A. Gaier, A. Asteroth, J.-B. Mouret, Feature space modeling through surrogate illumination, in: Proc. of GECCO, 2017
2017
Closest in time.