Fetching the paper…
Reading the bibliography…
A key open challenge in agile quadrotor flight is how to combine the flexibility and task-level generality of model-free reinforcement learning (RL) with the structure and online replanning capabilities of model predictive control (MPC), aiming to leverage their complementary strengths in dynamic and uncertain environments.
E. Arthur Jr and J.-C. Ho, Applied optimal control: optimization, estimation, and control . Hemisphere, 1975
1975
Earlier work this paper cites.
C. E. Garcia, D. M. Prett, and M. Morari, “Model predictive control: Theory and practice—a survey,” Automatica , vol. 25, no. 3, pp. 335–348, 1989
1989
Earlier work this paper cites.
R. P. N. Rao and D. H. Ballard, “Predictive coding in the visual cortex: a functional interpretation of some extra-classical receptive-field effects,” Nature Neuroscience , vol. 2, no. 1, pp. 79–87, 1999
1999
Earlier work this paper cites.
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” Advances in neural information processing systems , vol. 12, 1999
1999
Earlier work this paper cites.
M. Diehl, H. G. Bock, and J. P. Schlöder, “A real-time iteration scheme for nonlinear optimization in optimal feedback control,” SIAM Journal on control and optimization , vol. 43, no. 5, pp. 1714–1736, 2005
2005
Earlier work this paper cites.
M. Diehl, H. G. Bock, H. Diedam, and P. B. Wieber, “Fast direct multiple shooting algorithms for optimal robot control,” in Fast motions in biomechanics and robotics . Springer, 2006
2006
Earlier work this paper cites.
Y. Wang and S. Boyd, “Fast model predictive control using online optimization,” IEEE Transactions on control systems technology , vol. 18, no. 2, pp. 267–278, 2009
2009
Earlier work this paper cites.
K. Friston, “The free-energy principle: a unified brain theory?” Nature Reviews Neuroscience , vol. 11, no. 2, pp. 127–138, 2010
2010
Earlier work this paper cites.
L. Jaillet, J. Cortés, and T. Siméon, “Sampling-based path planning on configuration-space costmaps,” IEEE Transactions on Robotics , vol. 26, no. 4, pp. 635–646, 2010
2010
Earlier work this paper cites.
R. González, M. Fiacchini, J. L. Guzmán, T. Álamo, and F. Rodríguez, “Robust tube-based predictive control for mobile robots in off-road conditions,” Robotics and Autonomous Systems , vol. 59, no. 10, pp. 711–726, 2011
2011
Earlier work this paper cites.
J. B. Rawlings, D. Angeli, and C. N. Bates, “Fundamentals of economic model predictive control,” in 2012 IEEE 51st IEEE conference on decision and control (CDC) . IEEE, 2012, pp. 3851–3861
2012
Earlier work this paper cites.
D. V. Lu, D. Hershberger, and W. D. Smart, “Layered costmaps for context-sensitive navigation,” in 2014 IEEE/RSJ International Conference on Intelligent Robots and Systems , 2014, pp. 709–715
2014
Earlier work this paper cites.
M. Bangura and R. Mahony, “Real-time model predictive control for quadrotors,” IFAC World Congress , 2014
2014
Earlier work this paper cites.
M. Ellis, H. Durand, and P. D. Christofides, “A tutorial review of economic model predictive control methods,” Journal of Process Control , vol. 24, no. 8, pp. 1156–1178, 2014
2014
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
A. Liniger, A. Domahidi, and M. Morari, “Optimization-based autonomous racing of 1: 43 scale rc cars,” Optimal Control Applications and Methods , vol. 36, no. 5, pp. 628–647, 2015
2015
Earlier work this paper cites.
I. Lenz, R. A. Knepper, and A. Saxena, “Deepmpc: Learning deep latent features for model predictive control.” in Robotics: Science and Systems , vol. 10. Rome, Italy, 2015
2015
Earlier work this paper cites.
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller, “Embed to control: A locally linear latent dynamics model for control from raw images,” Advances in neural information processing systems , vol. 28, 2015
2015
Earlier work this paper cites.
J. Schulman, P. Moritz, S. Levine, M. Jordan, and P. Abbeel, “High-dimensional continuous control using generalized advantage estimation,” 2016
2015
Earlier work this paper cites.
M. Ellis, J. Liu, and P. D. Christofides, Economic Model Predictive Control: Theory, Formulations and Chemical Process Applications . Springer, 2016
2016
Earlier work this paper cites.
P.-B. Wieber, R. Tedrake, and S. Kuindersma, “Modeling and control of legged robots,” in Springer handbook of robotics . Springer, 2016, pp. 1203–1234
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of go with deep neural networks and tree search,” Nature , vol. 529, no. 7587, pp. 484–489, 2016
2016
Earlier work this paper cites.
M. Neunert, C. De Crousaz, F. Furrer, M. Kamel, F. Farshidian, R. Siegwart, and J. Buchli, “Fast nonlinear model predictive control for unified trajectory optimization and tracking,” in 2016 IEEE international conference on robotics and automation (ICRA) . IEEE, 2016, pp. 1398–1404
2016
Earlier work this paper cites.
S. Kuindersma, R. Deits, M. Fallon, A. Valenzuela, H. Dai, F. Permenter, T. Koolen, P. Marion, and R. Tedrake, “Optimization-based locomotion planning, estimation, and control design for the atlas humanoid robot,” Autonomous robots , vol. 40, no. 3, pp. 429–455, 2016
2016
Earlier work this paper cites.
R. Verschueren, M. Zanon, R. Quirynen, and M. Diehl, “Time-optimal race car driving using an online exact hessian based nonlinear mpc algorithm,” in 2016 European control conference (ECC) . IEEE, 2016, pp. 141–147
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
F. Farshidian, E. Jelavic, A. Satapathy, M. Giftthaler, and J. Buchli, “Real-time motion planning of legged robots: A model predictive control approach,” in 2017 IEEE-RAS 17th International Conference on Humanoid Robotics (Humanoids) . IEEE, 2017, pp. 577–584
2017
Earlier work this paper cites.
C. Liu, S. Lee, S. Varnhagen, and H. E. Tseng, “Path planning for autonomous vehicles using model predictive control,” in 2017 IEEE Intelligent Vehicles Symposium (IV) . IEEE, 2017, pp. 174–179
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
B. Amos, I. Jimenez, J. Sacks, B. Boots, and J. Z. Kolter, “Differentiable mpc for end-to-end planning and control,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
G. Bledt, M. J. Powell, B. Katz, J. Di Carlo, P. M. Wensing, and S. Kim, “Mit cheetah 3: Design and control of a robust, dynamic quadruped robot,” in 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2018, pp. 2245–2252
2018
Earlier work this paper cites.
M. Neunert, M. Stäuble, M. Giftthaler, C. D. Bellicoso, J. Carius, C. Gehring, M. Hutter, and J. Buchli, “Whole-body nonlinear model predictive control through contacts for quadrupeds,” IEEE Robotics and Automation Letters , vol. 3, no. 3, pp. 1458–1465, 2018
2018
Earlier work this paper cites.
G. Williams, P. Drews, B. Goldfain, J. M. Rehg, and E. A. Theodorou, “Information-theoretic model predictive control: Theory and applications to autonomous driving,” IEEE Transactions on Robotics , vol. 34, no. 6, pp. 1603–1622, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Nagabandi, G. Kahn, R. S. Fearing, and S. Levine, “Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 7559–7566
2018
Earlier work this paper cites.
K. Chua, R. Calandra, R. McAllister, and S. Levine, “Deep reinforcement learning in a handful of trials using probabilistic dynamics models,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
O. Vinyals, I. Babuschkin, W. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. Choi, R. Powell, T. Ewalds, P. Georgiev et al. , “Grandmaster level in starcraft ii using multi-agent reinforcement learning,” Nature , vol. 575, no. 7782, pp. 350–354, 2019
2019
Earlier work this paper cites.
A. Romero, P. N. Beuchat, Y. R. Stürz, R. S. Smith, and J. Lygeros, “Nonlinear control of quadcopters via approximate dynamic programming,” in 2019 18th European Control Conference (ECC) . IEEE, 2019, pp. 3752–3759
2019
Earlier work this paper cites.
G. Shi, X. Shi, M. O’Connell, R. Yu, K. Azizzadenesheli, A. Anandkumar, Y. Yue, and S.-J. Chung, “Neural lander: Stable drone landing control using learned dynamics,” in 2019 international conference on robotics and automation (icra) . IEEE, 2019, pp. 9784–9790
2019
Earlier work this paper cites.
S. East, M. Gallieri, J. Masci, J. Koutnik, and M. Cannon, “Infinite-horizon differentiable model predictive control,” in International Conference on Learning Representations , 2019
2019
Earlier work this paper cites.
S. Gros and M. Zanon, “Data-driven economic nmpc using reinforcement learning,” IEEE Transactions on Automatic Control , vol. 65, no. 2, pp. 636–648, 2019
2019
Earlier work this paper cites.
P. Foehn, D. Brescianini, E. Kaufmann, T. Cieslewski, M. Gehrig, M. Muglikar, and D. Scaramuzza, “Alphapilot: Autonomous drone racing,” Robotics: Science and Systems (RSS) , 2020
2020
Cited alongside, same era.
L. Hewing, K. P. Wabersich, M. Menner, and M. N. Zeilinger, “Learning-based model predictive control: Toward safe learning in control,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 3, pp. 269–296, 2020
2020
Cited alongside, same era.
O. M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray et al. , “Learning dexterous in-hand manipulation,” The International Journal of Robotics Research , vol. 39, no. 1, pp. 3–20, 2020
2020
Cited alongside, same era.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science robotics , vol. 5, no. 47, p. eabc5986, 2020
2020
Cited alongside, same era.
C. De Wagter, F. Paredes-Vallés, N. Sheth, and G. C. H. E. de Croon, “The sensing, state-estimation, and control behind the winning entry to the 2019 artificial intelligence robotic racing competition,” Field Robotics , vol. 2, pp. 1263–1290, 2022
2022
Later among the works it cites.
R. Penicka and D. Scaramuzza, “Minimum-time quadrotor waypoint flight in cluttered environments,” IEEE Robotics and Automation Letters , vol. 7, no. 2, pp. 5719–5726, 2022
2022
Later among the works it cites.
P. Foehn, E. Kaufmann, A. Romero, R. Penicka, S. Sun, L. Bauersfeld, T. Laengle, G. Cioffi, Y. Song, A. Loquercio, and D. Scaramuzza, “Agilicious: Open-source and open-hardware agile quadrotor for vision-based flight,” Science Robotics , vol. 7, no. 67, p. eabl6259, 2022
2022
Later among the works it cites.
E. Kaufmann, L. Bauersfeld, and D. Scaramuzza, “A benchmark comparison of learned control policies for agile quadrotor flight,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 10 504–10 510
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Hoeller, F. Farshidian, and M. Hutter, “Deep value model predictive control,” in Conference on Robot Learning . PMLR, 2020, pp. 990–1004
2020
Cited alongside, same era.
A. Nagabandi, K. Konolige, S. Levine, and V. Kumar, “Deep dynamics models for learning dexterous manipulation,” in Conference on Robot Learning . PMLR, 2020, pp. 1101–1112
2020
Cited alongside, same era.
M. Zanon, “A gauss–newton-like hessian approximation for economic nmpc,” IEEE Transactions on Automatic Control , vol. 66, no. 9, pp. 4206–4213, 2020
2020
Cited alongside, same era.
S. Gros, M. Zanon, R. Quirynen, A. Bemporad, and M. Diehl, “From linear to nonlinear mpc: bridging the gap via the real-time iteration,” International Journal of Control , vol. 93, no. 1, pp. 62–80, 2020
2020
Cited alongside, same era.
Y. Song, S. Naji, E. Kaufmann, A. Loquercio, and D. Scaramuzza, “Flightmare: A flexible quadrotor simulator,” in Conference on Robot Learning , 2020
2020
Cited alongside, same era.
P. Foehn, A. Romero, and D. Scaramuzza, “Time-optimal planning for quadrotor waypoint flight,” Science Robotics , vol. 6, no. 56, 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
K. P. Wabersich, L. Hewing, A. Carron, and M. N. Zeilinger, “Probabilistic model predictive safety certification for learning-based control,” IEEE Transactions on Automatic Control , vol. 67, no. 1, pp. 176–188, 2021
2021
Cited alongside, same era.
2022
Later among the works it cites.
Y. Song, A. Romero, M. Mueller, V. Koltun, and D. Scaramuzza, “Reaching the limit in autonomous racing: Optimal control versus reinforcement learning,” Science Robotics , p. adg1462, 2023
2023
Closest in time.
C. Chi, Z. Xu, S. Feng, E. Cousineau, Y. Du, B. Burchfiel, R. Tedrake, and S. Song, “Diffusion policy: Visuomotor policy learning via action diffusion,” The International Journal of Robotics Research , p. 02783649241273668, 2023
2023
Closest in time.
A. Ajay, Y. Du, A. Gupta, J. Tenenbaum, T. Jaakkola, and P. Agrawal, “Is conditional generative modeling all you need for decision making?” in The Eleventh International Conference on Learning Representations , 2023
2023
Closest in time.
E. Kaufmann, L. Bauersfeld, A. Loquercio, M. Müller, V. Koltun, and D. Scaramuzza, “Champion-level drone racing using deep reinforcement learning,” Nature , vol. 620, no. 7976, pp. 982–987, Aug 2023
2023
Closest in time.
E. Arcari, M. V. Minniti, A. Scampicchio, A. Carron, F. Farshidian, M. Hutter, and M. N. Zeilinger, “Bayesian multi-task learning mpc for robotic mobile manipulation,” IEEE Robotics and Automation Letters , vol. 8, no. 6, pp. 3222–3229, 2023
2023
Closest in time.
A. Saviolo, J. Frey, A. Rathod, M. Diehl, and G. Loianno, “Active learning of discrete-time dynamics for uncertainty-aware model predictive control,” IEEE Transactions on Robotics , 2023
2023
Closest in time.
G. Li and G. Loianno, “Nonlinear model predictive control for cooperative transportation and manipulation of cable suspended payloads with multiple quadrotors,” in 2023 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2023, pp. 5034–5041
2023
Closest in time.
D. Mankowitz, A. Michi, A. Zhernov et al. , “Faster sorting algorithms discovered using deep reinforcement learning,” Nature , vol. 618, pp. 257–263, 2023
2023
Closest in time.
A. Romero, S. Govil, G. Yilmaz, Y. Song, and D. Scaramuzza, “Weighted maximum likelihood for controller tuning,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 1334–1341
2023
Closest in time.
G. Grandesso, E. Alboni, G. P. R. Papini, P. M. Wensing, and A. Del Prete, “Cacto: Continuous actor-critic with trajectory optimization—towards global optimality,” IEEE Robotics and Automation Letters , vol. 8, no. 6, pp. 3318–3325, 2023
2023
Closest in time.
R. Reiter, J. Hoffmann, J. Boedecker, and M. Diehl, “A hierarchical approach for strategic motion planning in autonomous racing,” in 2023 European Control Conference (ECC) . IEEE, 2023, pp. 1–8
2023
Closest in time.
E. Bøhn, S. Gros, S. Moe, and T. A. Johansen, “Optimization of the model predictive control meta-parameters through reinforcement learning,” Engineering Applications of Artificial Intelligence , vol. 123, p. 106211, 2023
2023
Closest in time.
H. N. Esfahani, A. B. Kordabad, W. Cai, and S. Gros, “Learning-based state estimation and control using mhe and mpc schemes with imperfect models,” European Journal of Control , vol. 73, p. 100880, 2023
2023
Closest in time.
A. S. Anand, D. Reinhardt, S. Sawant, J. T. Gravdahl, and S. Gros, “A painless deterministic policy gradient method for learning-based mpc,” in 2023 European Control Conference (ECC) . IEEE, 2023, pp. 1–7
2023
Closest in time.
W. Cai, S. Sawant, D. Reinhardt, S. Rastegarpour, and S. Gros, “A learning-based model predictive control strategy for home energy management systems,” IEEE Access , 2023
2023
Closest in time.
S. Cheng, L. Song, M. Kim, S. Wang, and N. Hovakimyan, “Difftune + : Hyperparameter-free auto-tuning using auto-differentiation,” in Proceedings of The 5th Annual Learning for Dynamics and Control Conference , ser. Proceedings of Machine Learning Research, N. Matni, M. Morari, and G. J. Pappas, Eds., vol. 211. PMLR, 15–16 Jun 2023, pp. 170–183
2023
Closest in time.
F. Yang, C. Wang, C. Cadena, and M. Hutter, “iplanner: Imperative path planning,” Robotics: Science and Systems Conference (RSS) , 2023
2023
Closest in time.
P. Karkus, B. Ivanovic, S. Mannor, and M. Pavone, “Diffstack: A differentiable and modular control stack for autonomous vehicles,” in Conference on Robot Learning . PMLR, 2023, pp. 2170–2180
2023
Closest in time.
B. Wang, Z. Ma, S. Lai, and L. Zhao, “Neural moving horizon estimation for robust flight control,” IEEE Transactions on Robotics , 2023
2023
Closest in time.
C. Wang, D. Gao, K. Xu, J. Geng, Y. Hu, Y. Qiu, B. Li, F. Yang, B. Moon, A. Pandey, Aryan, J. Xu, T. Wu, H. He, D. Huang, Z. Ren, S. Zhao, T. Fu, P. Reddy, X. Lin, W. Wang, J. Shi, R. Talak, K. Cao, Y. Du, H. Wang, H. Yu, S. Wang, S. Chen, A. Kashyap, R. Bandaru, K. Dantu, J. Wu, L. Xie, L. Carlone, M. Hutter, and S. Scherer, “PyPose: A library for robot learning with physics-based optimization,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023
2023
Closest in time.
A. Romero, Y. Song, and D. Scaramuzza, “Actor-critic model predictive control,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 14 777–14 784
2024
Closest in time.
T. Han, A. Liu, A. Li, A. Spitzer, G. Shi, and B. Boots, “Model predictive control for aggressive driving over uneven terrain,” 2024
2024
Closest in time.
M. Krinner, A. Romero, L. Bauersfeld, M. Zeilinger, A. Carron, and D. Scaramuzza, “Mpcc++: Model predictive contouring control for time-optimal flight with safety constraints,” Proceedings of Robotics: Science and Systems, Delft, Netherlands , 2024
2024
Closest in time.
A. Tagliabue and J. P. How, “Efficient deep learning of robust policies from mpc using imitation and tube-guided data augmentation,” IEEE Transactions on Robotics , 2024
2024
Closest in time.
——, “Tube-nerf: Efficient imitation learning of visuomotor policies from mpc via tube-guided data augmentation and nerfs,” IEEE Robotics and Automation Letters , 2024
2024
Closest in time.
M. Minarik, R. Penicka, V. Vonasek, and M. Saska, “Model predictive path integral control for agile unmanned aerial vehicles,” pp. 13 144–13 151, 2024
2024
Closest in time.
J. Sacks, R. Rana, K. Huang, A. Spitzer, G. Shi, and B. Boots, “Deep model predictive optimization,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 16 945–16 953
2024
Closest in time.
B. Zarrouki, M. Spanakakis, and J. Betz, “A safe reinforcement learning driven weights-varying model predictive control for autonomous vehicle motion control,” pp. 1401–1408, 2024
2024
Closest in time.
F. Jenelten, J. He, F. Farshidian, and M. Hutter, “Dtc: Deep tracking control,” Science Robotics , vol. 9, no. 86, p. eadh5401, 2024
2024
Closest in time.
W. Cao, A. Capone, R. Yadav, S. Hirche, and W. Pan, “Computation-aware learning for stable control with gaussian process,” 2024
2024
Closest in time.
S. Adhau, S. Gros, and S. Skogestad, “Reinforcement learning based mpc with neural dynamical models,” European Journal of Control , p. 101048, 2024
2024
Closest in time.
N. Hansen, H. Su, and X. Wang, “Td-mpc2: Scalable, robust world models for continuous control,” 2024
2024
Closest in time.
E. Aljalbout, F. Frank, M. Karl, and P. van der Smagt, “On the role of the action space in robot manipulation learning and sim-to-real transfer,” IEEE Robotics and Automation Letters , 2024
2024
Closest in time.
A. Das, R. D. Yadav, S. Sun, M. Sun, S. Kaski, and W. Pan, “Dronediffusion: Robust quadrotor dynamics learning with diffusion models,” in 2025 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2025, pp. 1604–1610
2025
Closest in time.
H. Xue, C. Pan, Z. Yi, G. Qu, and G. Shi, “Full-order sampling-based mpc for torque-level locomotion control via diffusion-style annealing,” pp. 4974–4981, 2025
2025
Closest in time.
R. Reiter, A. Ghezzi, K. Baumgärtner, J. Hoffmann, R. D. McAllister, and M. Diehl, “Ac4mpc: Actor-critic reinforcement learning for nonlinear model predictive control,” IEEE Transactions on Control Systems Technology , 2025
2025
Closest in time.
X. Zhang, W. Pan, C. Li, X. Xu, X. Wang, R. Zhang, and D. Hu, “Toward scalable multirobot control: Fast policy learning in distributed mpc,” IEEE Transactions on Robotics , 2025
2025
Closest in time.
E. Aljalbout, N. Sotirakis, P. van der Smagt, M. Karl, and N. Chen, “Limt: Language-informed multi-task visual world models,” pp. 8226–8233, 2025
2025
Closest in time.