Fetching the paper…
Reading the bibliography…
Model predictive control (MPC) is an effective method for controlling robotic systems, particularly autonomous aerial vehicles such as quadcopters.
D. Jacobson and D. Mayne, Differential Dynamic Programming . Elsevier, 1970
1970
Earlier work this paper cites.
D. Pomerleau, “ALVINN: an autonomous land vehicle in a neural network,” in Advances in Neural Information Processing Systems (NIPS) , 1989
1989
Earlier work this paper cites.
K. J. Hunt, D. Sbarbaro, R. Żbikowski, and P. J. Gawthrop, “Neural networks for control systems: A survey,” Automatica , vol. 28, no. 6, pp. 1083–1112, Nov. 1992
1992
Earlier work this paper cites.
D. H. Shim, H. J. Kim, and S. Sastry, “Nonlinear model predictive tracking control for rotorcraft-based unmanned aerial vehicles,” in American Control Conference (ACC) , 2002
2002
Earlier work this paper cites.
N. Kohl and P. Stone, “Policy gradient reinforcement learning for fast quadrupedal locomotion,” in International Conference on Robotics and Automation (IROS) , 2004
2004
Earlier work this paper cites.
W. Li and E. Todorov, “Iterative linear quadratic regulator design for nonlinear biological movement systems,” in ICINCO (1) , 2004, pp. 222–229
2004
Earlier work this paper cites.
A. G. Richards, “Robust constrained model predictive control,” Ph.D. dissertation, Massachusetts Institute of Technology, 2004
2004
Earlier work this paper cites.
R. Tedrake, T. Zhang, and H. Seung, “Stochastic policy gradient reinforcement learning on a simple 3d biped,” in International Conference on Intelligent Robots and Systems (IROS) , 2004
2004
Earlier work this paper cites.
D. Q. Mayne, M. M. Seron, and S. V. Raković, “Robust model predictive control of constrained linear systems with bounded disturbances,” Automatica , vol. 41, no. 2, Feb. 2005
2005
Earlier work this paper cites.
P. Abbeel, A. Coates, M. Quigley, and A. Ng, “An application of reinforcement learning to aerobatic helicopter flight,” in Advances in Neural Information Processing Systems (NIPS) , 2006
2006
Earlier work this paper cites.
T. Geng, B. Porr, and F. Wörgötter, “Fast biped walking with a reflexive controller and realtime policy searching,” in Advances in Neural Information Processing Systems (NIPS) , 2006
2006
Earlier work this paper cites.
G. Endo, J. Morimoto, T. Matsubara, J. Nakanishi, and G. Cheng, “Learning CPG-based biped locomotion with a policy gradient method: Application to a humanoid robot,” International Journal of Robotic Research , vol. 27, no. 2, pp. 213–228, 2008
2008
Earlier work this paper cites.
J. Peters and S. Schaal, “Reinforcement learning of motor skills with policy gradients,” Neural Networks , vol. 21, no. 4, pp. 682–697, 2008
2008
Earlier work this paper cites.
P. Pastor, H. Hoffmann, T. Asfour, and S. Schaal, “Learning and generalization of motor skills by learning from demonstration,” in International Conference on Robotics and Automation (ICRA) , 2009
2009
Earlier work this paper cites.
J. Kober, E. Oztop, and J. Peters, “Reinforcement learning to adjust robot movements to new situations,” in Robotics: Science and Systems (RSS) , 2010
2010
Cited alongside, same era.
P. Martin and E. Salaun, “The true role of accelerometer feedback in quadrotor control,” in International Conference on Robotics and Automation (ICRA) , 2010
2010
Cited alongside, same era.
V. Nair and G. Hinton, “Rectified linear units improve restricted boltzmann machines,” in International Conference on Machine Learning (ICML) , 2010
2010
Cited alongside, same era.
G. V. Raffo, M. G. Ortega, and F. R. Rubio, “An integral predictive/nonlinear control structure for a quadrotor helicopter,” Automatica , vol. 46, no. 1, pp. 29 – 39, 2010
2010
Cited alongside, same era.
M. Deisenroth, C. Rasmussen, and D. Fox, “Learning to control a low-cost manipulator using data-efficient reinforcement learning,” in Robotics: Science and Systems (RSS) , 2011
M. Deisenroth, G. Neumann, and J. Peters, “A survey on policy search for robotics,” Foundations and Trends in Robotics , vol. 2, no. 1-2, pp. 1–142, 2013
2013
Later among the works it cites.
J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement learning in robotics: A survey,” International Journal of Robotic Research , vol. 32, no. 11, pp. 1238–1274, 2013
2013
Later among the works it cites.
S. Levine and V. Koltun, “Guided policy search,” in International Conference on Machine Learning (ICML) , 2013
2013
Later among the works it cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing Atari with deep reinforcement learning,” NIPS ’13 Workshop on Deep Learning , 2013
2013
Later among the works it cites.
M. Mueller and R. D’Andrea, “A model predictive controller for quadrocopter state interception,” in European Control Conference (ECC) , 2013
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2011
Cited alongside, same era.
L. Heng, L. Meier, P. Tanskanen, F. Fraundorfer, and M. Pollefeys, “Autonomous obstacle avoidance and maneuvering on a vision-guided mav using on-board processing,” in International Conference on Robotics and Automation (ICRA) , 2011
2011
Cited alongside, same era.
S. Ross, G. Gordon, and A. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” Journal of Machine Learning Research , vol. 15, pp. 627–635, 2011
2011
Cited alongside, same era.
K. Alexis, G. Nikolakopoulos, and A. Tzes, “Model predictive quadrotor control: attitude, altitude and position experimental studies,” Control Theory Applications, IET , vol. 6, no. 12, pp. 1812–1827, Aug 2012
2012
Cited alongside, same era.
F. Augugliaro, A. P. Schoellig, and R. D’Andrea, “Generation of collision-free trajectories for a quadrocopter fleet: A sequential convex programming approach,” in International Conference on Intelligent Robots and Systems (IROS) , 2012
2012
Cited alongside, same era.
P. Bouffard, A. Aswani, and C. Tomlin, “Learning-based model predictive control on a quadrotor: Onboard implementation and experimental results,” in International Conference on Robotics and Automation (ICRA) , 2012
2012
Cited alongside, same era.
F. Fraundorfer, L. Heng, D. Honegger, G. Lee, L. Meier, P. Tanskanen, and M. Pollefeys, “Vision-based autonomous mapping and exploration using a quadrotor mav,” in International Conference on Intelligent Robots and Systems (IROS) , 2012
2012
Cited alongside, same era.
F. L. Mueller, A. P. Schoellig, and R. D’Andrea, “Iterative learning of feed-forward corrections for high-performance tracking,” in International Conference on Intelligent Robots and Systems (IROS) , 2012
2012
Cited alongside, same era.
2013
Later among the works it cites.
S. Ross, N. Melik-Barkhudarov, K. S. Shankar, A. Wendel, D. Dey, J. A. Bagnell, and M. Hebert, “Learning monocular reactive UAV control in cluttered natural environments,” in International Conference on Robotics and Automation (ICRA) , 2013
2013
Later among the works it cites.
S. Levine and P. Abbeel, “Learning neural network policies with guided policy search under unknown dynamics,” in Advances in Neural Information Processing Systems (NIPS) , 2014
2014
Later among the works it cites.
——, “Learning complex neural network policies with trajectory optimization,” in International Conference on Machine Learning (ICML) , 2014
2014
Later among the works it cites.
R. Deits and R. Tedrake, “Efficient mixed-integer planning for UAVs in cluttered environments,” in International Conference on Robotics and Automation (ICRA) , 2015
2015
Closest in time.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in International Conference on Learning Representations (ICLR) , 2015
2015
Closest in time.
I. Lenz, R. Knepper, and A. Saxena, “Deepmpc: Learning deep latent features for model predictive control,” in Robotics: Science and Systems (RSS) , 2015
2015
Closest in time.
S. Levine, N. Wagener, and P. Abbeel, “Learning contact-rich manipulation skills with guided policy search,” in International Conference on Robotics and Automation (ICRA) , 2015
2015
Closest in time.