Fetching the paper…
Reading the bibliography…
Recent advancement in combining trajectory optimization with function approximation (especially neural networks) shows promise in learning complex control policies for diverse tasks in robot systems.
“Differential dynamic programming”
David Jacobson and David Mayne · 1970
Earlier work this paper cites.
“Virtual adversarial training: a regularization method for supervised and semi-supervised learning”
Takeru Miyato, Shin-ichi Maeda, Masanori Koyama and Shin Ishii · 1993
Earlier work this paper cites.
“A class of smoothing functions for nonlinear and mixed complementarity problems”
Chunhui Chen and Olvi Mangasarian · 1996
Earlier work this paper cites.
“Learning from demonstration”
Stefan Schaal · 1997
Earlier work this paper cites.
“Survey of numerical methods for trajectory optimization”
John Betts · 1998
Earlier work this paper cites.
“Robust reinforcement learning”
Jun Morimoto and Kenji Doya · 2000
Earlier work this paper cites.
“Estimating contact dynamics”
Marcus Brubaker, Leonid Sigal and David Fleet · 2009
Earlier work this paper cites.
“Market structure and equilibrium”
Heinrich Von · 2010
Earlier work this paper cites.
“Distributed optimization and statistical learning via the alternating direction method of multipliers”
Stephen Boyd, Neal Parikh and Eric Chu · 2011
Earlier work this paper cites.
“A reduction of imitation learning and structured prediction to no-regret online learning”
Stéphane Ross, Geoffrey Gordon and Drew Bagnell · 2011
Earlier work this paper cites.
“A convex, smooth and invertible contact model for trajectory optimization”
Emanuel Todorov · 2011
Earlier work this paper cites.
“Discovery of complex behaviors through contact-invariant optimization”
Igor Mordatch, Emanuel Todorov and Zoran Popović · 2012
Earlier work this paper cites.
“Mujoco: A physics engine for model-based control”
Emanuel Todorov, Tom Erez and Yuval Tassa · 2012
Earlier work this paper cites.
“A survey on policy search for robotics”
Marc Deisenroth, Gerhard Neumann and Jan Peters · 2013
Earlier work this paper cites.
“Guided policy search”
Sergey Levine and Vladlen Koltun · 2013
Earlier work this paper cites.
“Variational policy search via trajectory optimization”
Sergey Levine and Vladlen Koltun · 2013
Earlier work this paper cites.
“Adam: A method for stochastic optimization”
Diederik Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
“Learning complex neural network policies with trajectory optimization”
Sergey Levine and Vladlen Koltun · 2014
Cited alongside, same era.
“Combining the benefits of function approximation and trajectory optimization.”
Igor Mordatch and Emo Todorov · 2014
Cited alongside, same era.
“A direct method for trajectory optimization of rigid bodies through contact”
Michael Posa, Cecilia Cantu and Russ Tedrake · 2014
Cited alongside, same era.
“Control-limited differential dynamic programming”
Yuval Tassa, Nicolas Mansard and Emo Todorov · 2014
Cited alongside, same era.
“Generative adversarial imitation learning”
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
“Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot”
Scott Kuindersma et al · 2016
“On Value Discrepancy of Imitation Learning”
Tian Xu, Ziniu Li and Yang Yu · 2019
Later among the works it cites.
“Theoretically principled trade-off between robustness and accuracy”
Hongyang Zhang et al · 2019
Later among the works it cites.
“Task-relevant adversarial imitation learning”
Konrad Zolna et al · 2019
Later among the works it cites.
“Adversarial feature training for generalizable robotic visuomotor control”
Xi Chen, Ali Ghadirzadeh, Mårten Björkman and Patric Jensfelt · 2020
Later among the works it cites.
“Online trajectory planning through combined trajectory optimization and function approximation: Application to the exoskeleton Atalante”
Alexis Duburcq, Yann Chevaleyre, Nicolas Bredeche and Guilhem Boéris · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Decoupled weight decay regularization”
Ilya Loshchilov and Frank Hutter · 2017
Cited alongside, same era.
“Robust adversarial reinforcement learning”
Lerrel Pinto, James Davidson, Rahul Sukthankar and Abhinav Gupta · 2017
Cited alongside, same era.
“Lipschitz continuity in model-based reinforcement learning”
Kavosh Asadi, Dipendra Misra and Michael Littman · 2018
Cited alongside, same era.
“Plan online, learn offline: Efficient learning and exploration via model-based control”
Kendall Lowrey et al · 2018
Cited alongside, same era.
“Generalized Inner Loop Meta-Learning”
Edward Grefenstette et al · 2019
Cited alongside, same era.
“Using Self-Supervised Learning Can Improve Model Robustness and Uncertainty”
Dan Hendrycks, Mantas Mazeika, Saurav Kadavath and Dawn Song · 2019
Cited alongside, same era.
Later among the works it cites.
“Implicit bias of gradient descent based adversarial training on separable data”, 2020
Yan Li, Ethan Fang, Huan Xu and Tuo Zhao · 2020
Later among the works it cites.
“Crocoddyl: An Efficient and Versatile Framework for Multi-Contact Optimal Control”
Carlos Mastalli et al · 2020
Later among the works it cites.
“Deep Reinforcement Learning with Robust and Smooth Policy”
Qianli Shen et al · 2020
Later among the works it cites.
“Adversarial robustness through local lipschitzness”
Yao-Yuan Yang et al · 2020
Later among the works it cites.
Zhigen Zhao, Ziyi Zhou, Michael Park and Ye Zhao · 2020
Later among the works it cites.
“Accelerated ADMM based Trajectory Optimization for Legged Locomotion with Coupled Rigid Body Dynamics”
Ziyi Zhou and Ye Zhao · 2020
Later among the works it cites.
“PyBullet, a Python module for physics simulation for games, robotics and machine learning”, http://pybullet.org , 2016–2021
Erwin Coumans and Yunfei Bai · 2021
Closest in time.
“Robust trajectory optimization over uncertain terrain with stochastic complementarity”
Luke Drnach and Ye Zhao · 2021
Closest in time.
“SEAGuL: Sample Efficient Adversarially Guided Learning of Value Functions”
Benoit Landry, Hongkai Dai and Marco Pavone · 2021
Closest in time.
“Adversarial Training is Not Ready for Robot Learning”
Mathias Lechner et al · 2021
Closest in time.
“Adversarial Training as Stackelberg Game: An Unrolled Optimization Approach”
Simiao Zuo et al · 2021
Closest in time.
“ARCH: Efficient Adversarial Regularized Training with Caching”
Simiao Zuo et al · 2021
Closest in time.