Fetching the paper…
Reading the bibliography…
Robotic algorithms typically depend on various parameters, the choice of which significantly affects the robot's performance.
In: Proc. of the IEEE International Conference on Robotics and Automation (ICRA) . pp. 1951–1958
Calandra R, Seyfarth A, Peters J and Deisenroth MP (2014b) An experimental comparison of bayesian optimization for bipedal locomotion · 1958
Earlier work this paper cites.
World Scientific
Davidor Y (1991) Genetic algorithms and robotics: a heuristic strategy for optimization · 1991
Earlier work this paper cites.
Control Engineering Practice 1(4): 699–714
Åström KJ, Hägglund T, Hang CC and Ho WK (1993) Automatic tuning and adaptation for pid controllers - a survey · 1993
Earlier work this paper cites.
MIT press
Sutton RS and Barto AG (1998) Reinforcement learning: an introduction · 1998
Earlier work this paper cites.
Prentice Hall
Zhou K and Doyle JC (1998) Essentials of robust control , volume 104 · 1998
Earlier work this paper cites.
Journal of Global Optimization 21(4): 345–383
Jones DR (2001) A taxonomy of global optimization methods based on response surfaces · 2001
Earlier work this paper cites.
In: Becker S, Thrun S and Obermayer K (eds.) Proc. of Neural Information Processing Systems (NIPS) . MIT Press, pp. 1057–1064
Solak E, Murray-smith R, Leithead WE, Leith DJ and Rasmussen CE (2003) Derivative observations in gaussian process models of dynamic systems · 2003
Earlier work this paper cites.
The Annals of Statistics 34(5): 2413–2429
Ghosal S and Roy A (2006) Posterior consistency of gaussian process prior for nonparametric binary regression · 2006
Earlier work this paper cites.
IEEE Control Systems 26(1): 70–79
Killingsworth NJ and Krstić M (2006) Pid tuning using extremum seeking: online, model-free performance optimization · 2006
Earlier work this paper cites.
In: Proc. of the IEEE/RSJ International Conference on Intelligent Robots and Systems . pp. 2219–2225
Peters J and Schaal S (2006) Policy gradient methods for robotics · 2006
Earlier work this paper cites.
Cambridge MA: MIT Press
Rasmussen CE and Williams CK (2006) Gaussian processes for machine learning · 2006
Earlier work this paper cites.
In: Proc. of the International Joint Conference on Artificial Intelligence (IJCAI) , volume 7. pp. 944–949
Lizotte DJ, Wang T, Bowling MH and Schuurmans D (2007) Automatic gait optimization with gaussian process regression · 2007
Earlier work this paper cites.
Information Science and Statistics. New York, NY: Springer
Christmann A and Steinwart I (2008) Support Vector Machines · 2008
Earlier work this paper cites.
Neural Networks 21(4): 682–697
Peters J and Schaal S (2008) Reinforcement learning of motor skills with policy gradients · 2008
Earlier work this paper cites.
IEEE Robotics & Automation Magazine 17(2): 20–29
Schaal S and Atkeson CG (2010) Learning control in robotics · 2010
Earlier work this paper cites.
Journal of Machine Learning Research 12(Oct): 2879–2904
Bull AD (2011) Convergence rates of efficient global optimization algorithms · 2011
Cited alongside, same era.
In: NIPS . pp. 226–234
Duvenaud DK, Nickisch H and Rasmussen CE (2011) Additive gaussian processes · 2011
Cited alongside, same era.
In: Proc. of Neural Information Processing Systems (NIPS) . pp. 2447–2455
Krause A and Ong CS (2011) Contextual gaussian process bandit optimization · 2011
Cited alongside, same era.
In: Proc. of the American Control Conference (ACC) . pp. 3843–3849
Schoellig AP, Hehn M, Lupashin S and D’Andrea R (2011) Feasiblity of motion primitives for choreographed quadrocopter flight · 2011
Cited alongside, same era.
In: Proc. of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . pp. 1069–1074
Tesch M, Schneider J and Choset H (2011) Using response surfaces and expected improvement to optimize snake robot gait parameters · 2011
Cited alongside, same era.
In: Proc. of the Conference on Uncertainty in Artificial Intelligence (UAI) . pp. 250–259
Gelbart MA, Snoek J and Adams RP (2014) Bayesian optimization with unknown constraints · 2014
Later among the works it cites.
Kober J and Peters J (2014) Reinforcement learning in robotics: a survey
2014
Later among the works it cites.
Mechatronics 24(1): 41–54
Lupashin S, Hehn M, Mueller MW, Schoellig AP, Sherback M and D’Andrea R (2014) A platform for aerial robotics research and demonstration: The flying machine arena · 2014
Later among the works it cites.
In: Proc. of the European Control Conference (ECC) . pp. 2501–2506
Berkenkamp F and Schoellig AP (2015) Safe and robust learning control with gaussian processes · 2015
Later among the works it cites.
Lillicrap TP, Hunt JJ, Pritzel A, Heess N, Erez T, Tassa Y, Silver D and Wierstra D (2015) Continuous control with deep reinforcement learning · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Álvarez MA, Rosasco L and Lawrence ND (2012) Kernels for vector-valued functions: A review · 2012
Cited alongside, same era.
Springer Science & Business Media
Mockus J (2012) Bayesian approach to global optimization: theory and applications · 2012
Cited alongside, same era.
In: Proc. of the International Conference on Machine Learning (ICML) . pp. 1711–1718
Moldovan TM and Abbeel P (2012) Safe exploration in markov decision processes · 2012
Cited alongside, same era.
In: Proc. of the American Control Conference (ACC) . pp. 4313–4318
Schoellig A, Wiltsche C and D’Andrea R (2012) Feed-forward parameter identification for precise periodic quadrocopter motions · 2012
Cited alongside, same era.
IEEE Transactions on Information Theory 58(5): 3250–3265
Srinivas N, Krause A, Kakade SM and Seeger M (2012) Gaussian process optimization in the bandit setting: No regret and experimental design · 2012
Cited alongside, same era.
https://github.com/SheffieldML/GPy
The GPy authors (2012) Gpy: A gaussian process framework in python · 2012
Cited alongside, same era.
Automatica 49(5): 1216–1226
Aswani A, Gonzalez H, Sastry SS and Tomlin C (2013) Provably safe and robust learning-based model predictive control · 2013
Cited alongside, same era.
In: Machine Learning and Knowledge Discovery in Databases , 9286. Springer International Publishing, pp. 133–149
Schreiter J, Nguyen-Tuong D, Eberts M, Bischoff B, Markert H and Toussaint M (2015) Safe exploration for active learning with gaussian processes · 2015
Later among the works it cites.
In: Proc. of the International Conference on Machine Learning (ICML) . pp. 997–1005
Sui Y, Gotovos A, Burdick JW and Krause A (2015) Safe exploration for optimization with gaussian processes · 2015
Later among the works it cites.
In: Proc. of the IEEE International Conference on Robotics and Automation (ICRA) . pp. 493–496
Berkenkamp F, Schoellig AP and Krause A (2016) Safe controller optimization for quadrotors with gaussian processes · 2016
Closest in time.
The International Journal of Robotics Research (IJRR) 35(13): 1547–1536
Ostafew CJ, Schoellig AP and Barfoot TD (2016) Robust constrained learning-based nmpc enabling reliable mobile robot path tracking · 2016
Closest in time.
pp. 4305–4313
Turchetta M, Berkenkamp F and Krause A (2016) Safe exploration in finite markov decision processes with gaussian processes · 2016
Closest in time.
In: Proc. of the International Conference on Machine Learning (ICML)
Achiam J, Held D, Tamar A and Abbeel P (2017) Constrained policy optimization · 2017
Closest in time.
In: Proc. of Neural Information Processing Systems (NIPS)
Berkenkamp F, Turchetta M, Schoellig AP and Krause A (2017) Safe model-based reinforcement learning with stability guarantees · 2017
Closest in time.
In: Proc. of the International Conference on Machine Learning (ICML) . pp. 844–853
Chowdhury SR and Gopalan A (2017) On kernelized multi-armed bandits · 2017
Closest in time.
In: Proc. of the IFAC (International Federation of Automatic Control) World Congress . pp. 12306–12313
Duivenvoorden RR, Berkenkamp F, Carion N, Krause A and Schoellig AP (2017) Constrained bayesian optimization with particle swarms for adaptive controller tuning · 2017
Closest in time.
In: Proc. of the IEEE International Conference on Robotics and Automation (ICRA) . pp. 1557–1563
Marco A, Berkenkamp F, Hennig P, Schoellig AP, Krause A, Schaal S and Trimpe S (2017) Virtual vs. real: Trading off simulations and physical experiments in reinforcement learning with bayesian optimization · 2017
Closest in time.