Fetching the paper…
Reading the bibliography…
Trajectory optimization methods have achieved an exceptional level of performance on real-world robots in recent years.
K. Åström and P. Eykhoff, “System identification - a survey,” Automatica , vol. 7, no. 2, pp. 123–162, 1971
1971
Earlier work this paper cites.
K. Hornik, M. Stinchcombe, and H. White, “Multilayer feedforward networks are universal approximators,” Neural Networks , vol. 2, no. 5, pp. 359–366, 1989
1989
Earlier work this paper cites.
C. E. GarcÃa, D. M. Prett, and M. Morari, “Model predictive control: Theory and practice - a survey,” Automatica , pp. 335–348, 1989
1989
Earlier work this paper cites.
K. Narendra and K. Parthasarathy, “Identification and control of dynamical systems using neural networks,” IEEE Transactions on Neural Networks , vol. 1, no. 1, pp. 4–27, 1990
1990
Earlier work this paper cites.
J. Sjöberg, H. Hjalmarsson, and L. Ljung, “Neural networks in system identification,” IFAC Proceedings Volumes , vol. 27, no. 8, pp. 359–382, 1994
1994
Earlier work this paper cites.
K. Kozlowski, Modelling and Identification in Robotics . John Wiley & Sons, Ltd, 1998
1998
Earlier work this paper cites.
L. Ljung, System Identification . John Wiley & Sons, Ltd, 1999
1999
Earlier work this paper cites.
C. E. Rasmussen and C. K. I. Williams, Gaussian Processes for Machine Learning (Adaptive Computation and Machine Learning) . The MIT Press, 2005
2005
Earlier work this paper cites.
E. Todorov and W. Li, “A generalized iterative lqg method for locally-optimal feedback control of constrained nonlinear stochastic systems,” in Proceedings of the 2005, American Control Conference, 2005. , 2005, pp. 300–306 vol. 1
2005
Earlier work this paper cites.
J.-S. Wang and Y.-P. Chen, “A fully automated recurrent neural network for unknown dynamic system identification and control,” IEEE Transactions on Circuits and Systems I: Regular Papers , vol. 53, no. 6, pp. 1363–1372, 2006
2006
Earlier work this paper cites.
S. Thrun, M. Montemerlo, H. Dahlkamp, D. Stavens, A. Aron, J. Diebel, P. Fong, J. Gale, M. Halpenny, G. Hoffmann, K. Lau, C. Oakley, M. Palatucci, V. Pratt, P. Stang, S. Strohband, C. Dupont, L.-E. Jendrossek, C. Koelen, C. Markey, C. Rummel, J. van Niekerk, E. Jensen, P. Alessandrini, G. Bradski, B. Davies, S. Ettinger, A. Kaehler, A. Nefian, and P. Mahoney, Stanley: The Robot That Won the DARPA Grand Challenge . Springer Berlin Heidelberg, 2007
2007
Earlier work this paper cites.
L. Biagiotti and C. Melchiorri, Trajectory Planning for Automatic Machines and Robots , 1st ed. Springer Publishing Company, Incorporated, 2008
2008
Earlier work this paper cites.
M. Blösch, S. Weiss, D. Scaramuzza, and R. Siegwart, “Vision based mav navigation in unknown and unstructured environments,” in 2010 IEEE International Conference on Robotics and Automation , 2010, pp. 21–28
2010
Earlier work this paper cites.
J. Z. Kolter, C. Plagemann, D. T. Jackson, A. Y. Ng, and S. Thrun, “A probabilistic approach to mixed open-loop and closed-loop control, with application to extreme autonomous driving,” in 2010 IEEE International Conference on Robotics and Automation , 2010, pp. 839–845
2010
Earlier work this paper cites.
D. Nguyen-Tuong and J. Peters, “Model learning for robot control: a survey,” Cognitive Processing , vol. 12, no. 4, pp. 319–340, 2011
2011
Earlier work this paper cites.
M. P. Deisenroth and C. E. Rasmussen, “Pilco: A model-based and data-efficient approach to policy search,” in Proceedings of the 28th International Conference on machine learning (ICML-11) , ser. ICML’11. Omnipress, 2011
2011
Earlier work this paper cites.
Z. I. Botev, D. P. Kroese, R. Y. Rubinstein, and P. L’Ecuyer, “Chapter 3 - the cross-entropy method for optimization,” in Handbook of Statistics . Elsevier, 2013, pp. 35–59
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
A. Liniger, A. Domahidi, and M. Morari, “Optimization-based autonomous racing of 1:43 scale rc cars,” Optimal Control Applications and Methods , vol. 36, no. 5, p. 628–647, Jul 2014
2014
Earlier work this paper cites.
S. Levine and P. Abbeel, “Learning neural network policies with guided policy search under unknown dynamics,” in Advances in Neural Information Processing Systems , Z. Ghahramani, M. Welling, C. Cortes, N. Lawrence, and K. Q. Weinberger, Eds., vol. 27. Curran Associates, Inc., 2014
2014
Cited alongside, same era.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,” International Conference on Learning Representations , 2014
2014
Cited alongside, same era.
M. L. Puterman, Markov decision processes: discrete stochastic dynamic programming . John Wiley & Sons, 2014
2014
Cited alongside, same era.
J. Chung, C. Gulcehre, K. Cho, and Y. Bengio, “Empirical evaluation of gated recurrent neural networks on sequence modeling,” in NeurIPS 2014 Workshop on Deep Learning, December 2014 , 2014
2014
Cited alongside, same era.
2018
Later among the works it cites.
I. Clavera, J. Rothfuss, J. Schulman, Y. Fujita, T. Asfour, and P. Abbeel, “Model-based reinforcement learning via meta-policy optimization,” in Proceedings of The 2nd Conference on Robot Learning , ser. Proceedings of Machine Learning Research, vol. 87, 29–31 Oct 2018, pp. 617–629
2018
Later among the works it cites.
F. Rubio, F. Valero, and C. Llopis-Albert, “A review of mobile robots: Concepts, methods, theoretical framework, and applications,” International Journal of Advanced Robotic Systems , vol. 16, no. 2, 2019
2019
Later among the works it cites.
J. M. Bern, P. Banzet, R. Poranne, and S. Coros, “Trajectory optimization for cable-driven soft robot locomotion,” Robotics: Science and Systems XV , 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. P. Deisenroth, D. Fox, and C. E. Rasmussen, “Gaussian processes for data-efficient learning in robotics and control,” IEEE transactions on pattern analysis and machine intelligence , pp. 408–423, 2015
2015
Cited alongside, same era.
C. Gehring, S. Coros, M. Hutter, C. Dario Bellicoso, H. Heijnen, R. Diethelm, M. Bloesch, P. Fankhauser, J. Hwangbo, M. Hoepflinger, and R. Siegwart, “Practice makes perfect: An optimization-based approach to controlling agile motions for a quadruped robot,” IEEE Robotics Automation Magazine , vol. 23, no. 1, pp. 34–43, 2016
2016
Cited alongside, same era.
I. Goodfellow, Y. Bengio, and A. Courville, Deep learning . MIT press, 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
S. Gu, T. Lillicrap, I. Sutskever, and S. Levine, “Continuous deep q-learning with model-based acceleration,” in International conference on machine learning . PMLR, 2016, pp. 2829–2838
2016
Cited alongside, same era.
2016
Cited alongside, same era.
M. Geilinger, R. Poranne, R. Desai, B. Thomaszewski, and S. Coros, “Skaterbots: Optimization-based design and motion synthesis for robotic creatures with legs and wheels,” in Proceedings of ACM SIGGRAPH , A. T. on Graphics (TOG), Ed., vol. 37. ACM, August 2018
2018
Cited alongside, same era.
S. Zimmermann, R. Poranne, J. M. Bern, and S. Coros, “PuppetMaster,” ACM Transactions on Graphics , vol. 38, no. 4, pp. 1–11, 2019
2019
Later among the works it cites.
J. Kabzan, L. Hewing, A. Liniger, and M. N. Zeilinger, “Learning-based model predictive control for autonomous racing,” IEEE Robotics and Automation Letters , vol. 4, pp. 3363–3370, 2019
2019
Later among the works it cites.
A. Chiuso and G. Pillonetto, “System identification: A machine learning perspective,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 2, no. 1, pp. 281–304, 2019
2019
Later among the works it cites.
R. Boney, N. Di Palo, M. Berglund, A. Ilin, J. Kannala, A. Rasmus, and H. Valpola, “Regularizing trajectory optimization with denoising autoencoders,” in Advances in Neural Information Processing Systems , 2019
2019
Later among the works it cites.
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson, “Learning latent dynamics for planning from pixels,” in International conference on machine learning . PMLR, 2019, pp. 2555–2565
2019
Later among the works it cites.
M. Janner, J. Fu, M. Zhang, and S. Levine, “When to trust your model: Model-based policy optimization,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Later among the works it cites.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science Robotics , no. 47, 2020
2020
Later among the works it cites.
A. Nagabandi, K. Konolige, S. Levine, and V. Kumar, “Deep dynamics models for learning dexterous manipulation,” in Proceedings of the Conference on Robot Learning , ser. Proceedings of Machine Learning Research, L. P. Kaelbling, D. Kragic, and K. Sugiura, Eds., vol. 100. PMLR, 30 Oct–01 Nov 2020, pp. 1101–1112
2020
Later among the works it cites.
S. Curi, F. Berkenkamp, and A. Krause, “Efficient model-based reinforcement learning through optimistic policy search and planning,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. F. Balcan, and H. Lin, Eds., 2020, pp. 14 156–14 170
2020
Later among the works it cites.
H. Bharadhwaj, K. Xie, and F. Shkurti, “Model-predictive control via cross-entropy and gradient-based optimization,” in Proceedings of the 2nd Conference on Learning for Dynamics and Control , ser. Proceedings of Machine Learning Research, vol. 120. PMLR, 2020, pp. 277–286
2020
Later among the works it cites.
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi, “Dream to control: Learning behaviors by latent imagination,” 2020
2020
Later among the works it cites.
T. M. Moerland, J. Broekens, and C. M. Jonker, “Model-based reinforcement learning: A survey,” 2021
2021
Later among the works it cites.
S. Zimmermann, R. Poranne, and S. Coros, “Go fetch! - dynamic grasps using boston dynamics spot with external robotic arm,” in 2021 IEEE International Conference on Robotics and Automation (ICRA) , 2021, pp. 4488–4494
2021
Later among the works it cites.
A. R. Geist and S. Trimpe, “Structured learning of rigid-body dynamics: A survey and unified view from a robotics perspective,” GAMM-Mitteilungen , vol. 44, no. 2, 2021
2021
Later among the works it cites.