Fetching the paper…
Reading the bibliography…
Model predictive control (MPC) is a powerful technique for solving dynamic control tasks.
A measure of asymptotic efficiency for tests of a hypothesis based on the sum of observations
Herman Chernoff · 1952
Earlier work this paper cites.
Differential quadrature and long-term integration
Richard Bellman and John Casti · 1971
Earlier work this paper cites.
Likelihood ratio gradient estimation for stochastic systems
Peter W. Glynn · 1990
Earlier work this paper cites.
Natural gradient descent for on-line learning
Magnus Rattray, David Saad, and Shun-ichi Amari · 1998
Earlier work this paper cites.
Using the Cross-Entropy Method to Guide/Govern Mobile Agent’s Path Finding in Networks
Bjarne E. Helvik and Otto Wittner · 2002
Earlier work this paper cites.
Mirror descent and nonlinear projected subgradient methods for convex optimization
Amir Beck and Marc Teboulle · 2003
Earlier work this paper cites.
Reducing the Time Complexity of the Derandomized Evolution Strategy with Covariance Matrix Adaptation (CMA-ES)
Nikolaus Hansen, Sibylle D. Müller, and Petros Koumoutsakos · 2003
Earlier work this paper cites.
The cross entropy method for fast policy search
Shie Mannor, Reuven Y. Rubinstein, and Yohai Gat · 2003
Earlier work this paper cites.
Clustering with Bregman divergences
Arindam Banerjee, Srujana Merugu, Inderjit S. Dhillon, and Joydeep Ghosh · 2005
Earlier work this paper cites.
Basis Function Adaptation in Temporal Difference Reinforcement Learning
Ishai Menache, Shie Mannor, and Nahum Shimkin · 2005
Earlier work this paper cites.
Learning Tetris using the noisy cross-entropy method
István Szita and András Lörincz · 2006
Earlier work this paper cites.
Statistical exponential families: A digest with flash cards
Frank Nielsen and Vincent Garcia · 2009
Cited alongside, same era.
Autonomous helicopter aerobatics through apprenticeship learning
Pieter Abbeel, Adam Coates, and Andrew Y. Ng · 2010
Cited alongside, same era.
Risk sensitive path integral control
LJ van den Broek, WAJJ Wiegerinck, and HJ Kappen · 2010
Cited alongside, same era.
Theoretical foundation for CMA-ES from information geometry perspective
Youhei Akimoto, Yuichi Nagata, Isao Ono, and Shigenobu Kobayashi · 2012
Cited alongside, same era.
Cross-entropy randomized motion planning
Marin Kobilarov · 2012
Cited alongside, same era.
The cross-entropy method for optimization
Zdravko I. Botev, Dirk P. Kroese, Reuven Y. Rubinstein, and Pierre L’Ecuyer · 2013
Natural evolution strategies
Daan Wierstra, Tom Schaul, Tobias Glasmachers, Yi Sun, Jan Peters, and Jürgen Schmidhuber · 2014
Later among the works it cites.
Introduction to online convex optimization
Elad Hazan · 2016
Later among the works it cites.
Aggressive driving with model predictive path integral control
Grady Williams, Paul Drews, Brian Goldfain, James M. Rehg, and Evangelos A. Theodorou · 2016
Later among the works it cites.
Information Theoretic MPC for Model-Based Reinforcement Learning
Grady Williams, Nolan Wagener, Brian Goldfain, Paul Drews, James M. Rehg, Byron Boots, and Evangelos A. Theodorou · 2017
Later among the works it cites.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine · 2018
Later among the works it cites.
Mirror descent search and its acceleration
Megumi Miyashita, Shiro Yano, and Toshiyuki Kondo · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Model predictive control
Eduardo F. Camacho and Carlos Bordons Alba · 2013
Cited alongside, same era.
Linear-exponential-quadratic Gaussian control
Tyrone E. Duncan · 2013
Cited alongside, same era.
Dynamical models and tracking regret in online convex programming
Eric Hall and Rebecca Willett · 2013
Cited alongside, same era.
Model predictive control: Recent developments and future promise
David Q. Mayne · 2014
Cited alongside, same era.
Control-limited differential dynamic programming
Yuval Tassa, Nicolas Mansard, and Emo Todorov · 2014
Cited alongside, same era.
Information-Theoretic Model Predictive Control: Theory and Applications to Autonomous Driving
Grady Williams, Paul Drews, Brian Goldfain, James M. Rehg, and Evangelos A. Theodorou
Cited in the paper.
Later among the works it cites.
Acceleration of Gradient-based Path Integral Method for Efficient Optimal and Inverse Optimal Control
Masashi Okada and Tadahiro Taniguchi · 2018
Later among the works it cites.
Adaptive Online Learning in Dynamic Environments
Lijun Zhang, Shiyin Lu, and Zhi-Hua Zhou · 2018
Later among the works it cites.
Vision-Based High-Speed Driving With a Deep Dynamic Observer
Paul Drews, Grady Williams, Brian Goldfain, Evangelos A. Theodorou, and James M. Rehg · 2019
Closest in time.
AutoRally: An Open Platform for Aggressive Autonomous Driving
Brian Goldfain, Paul Drews, Changxi You, Matthew Barulic, Orlin Velev, Panagiotis Tsiotras, and James M. Rehg · 2019
Closest in time.