Fetching the paper…
Reading the bibliography…
Trajectory optimizers for model-based reinforcement learning, such as the Cross-Entropy Method (CEM), can yield compelling results even in high-dimensional control tasks and sparse-reward environments.
What is the fast fourier transform?
W. T. Cochran, J. W. Cooley, D. L. Favin, H. D. Helms, R. A. Kaenel, W. W. Lang, G. C. Maling, D. E. Nelson, C. M. Rader, and P. D. Welch · 1967
Earlier work this paper cites.
Evolutionsstrategie: Optimierung technischer Systeme nach Prinzipien der biologischen Evolution
I. Rechenberg · 1971
Earlier work this paper cites.
On generating power law noise
J. Timmer and M. Koenig · 1995
Earlier work this paper cites.
Adapting arbitrary normal mutation distributions in evolution strategies: the covariance matrix adaptation
N. Hansen and A. Ostermeier · 1996
Earlier work this paper cites.
The cross-entropy method for combinatorial and continuous optimization
R. Rubinstein and W. Davidson · 1999
Earlier work this paper cites.
The Cross Entropy Method: A Unified Approach To Combinatorial Optimization, Monte-Carlo Simulation (Information Science and Statistics)
R. Y. Rubinstein and D. P. Kroese · 2004
Earlier work this paper cites.
On the convergence of the cross-entropy method
L. Margolin · 2005
Earlier work this paper cites.
The cross-entropy method for network reliability estimation
K.-P. Hui, N. Bean, M. Kraetzl, and D. P. Kroese · 2005
Earlier work this paper cites.
A tutorial on the cross-entropy method
P.-T. De Boer, D. P. Kroese, S. Mannor, and R. Y. Rubinstein · 2005
Earlier work this paper cites.
Robust variable horizon model predictive control for vehicle maneuvering
A. Richards and J. P. How · 2006
Earlier work this paper cites.
Integration of ranked lists via cross entropy monte carlo with applications to mRNA and microRNA studies
S. Lin and J. Ding · 2008
Earlier work this paper cites.
Environmental context explains Lévy and Brownian movement patterns of marine predators
N. Humphries, N. Queiroz, J. Dyer, N. Pade, M. Musyl, K. Schaefer, D. Fuller, J. Brunnschweiler, T. Doyle, J. Houghton, G. Hays, C. Jones, L. Noble, V. Wearmouth, E. Southall, and D. Sims · 2010
Cited alongside, same era.
An adaptive coupled-layer visual model for robust visual tracking
L. Čehovin, M. Kristan, and A. Leonardis · 2011
Cited alongside, same era.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Cited alongside, same era.
Cross-entropy motion planning
M. Kobilarov · 2012
Cited alongside, same era.
Chapter 3. The Cross-Entropy Method for Optimization , volume 31, pages 35–59
Z. Botev, D. Kroese, R. Rubinstein, and P. L’Ecuyer · 2013
Cited alongside, same era.
Model predictive path integral control using covariance variable importance sampling
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
A. Nagabandi, G. Kahn, R. S. Fearing, and S. Levine · 2018
Later among the works it cites.
Structured evolution with compact architectures for scalable policy optimization
K. Choromanski, M. Rowland, V. Sindhwani, R. Turner, and A. Weller · 2018
Later among the works it cites.
Simple random search of static linear policies is competitive for reinforcement learning
H. Mania, A. Guy, and B. Recht · 2018
Later among the works it cites.
Evolution-guided policy gradient in reinforcement learning
S. Khadka and K. Tumer · 2018
Later among the works it cites.
CEM-RL: Combining evolutionary and gradient-based methods for policy search
A. Pourchot and O. Sigaud · 2018
Later among the works it cites.
Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
G. Williams, A. Aldrich, and E. Theodorou · 2015
Cited alongside, same era.
Aggressive driving with model predictive path integral control
G. Williams, P. Drews, B. Goldfain, J. M. Rehg, and E. Theodorou · 2016
Cited alongside, same era.
Benchmarking deep reinforcement learning for continuous control
Y. Duan, X. Chen, R. Houthooft, J. Schulman, and P. Abbeel · 2016
Cited alongside, same era.
OpenAI gym
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Cited alongside, same era.
Evolution strategies as a scalable alternative to reinforcement learning
T. Salimans, J. Ho, X. Chen, S. Sidor, and I. Sutskever · 2017
Cited alongside, same era.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
K. Chua, R. Calandra, R. McAllister, and S. Levine · 2018
Cited alongside, same era.
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine · 2018
Later among the works it cites.
Learning latent dynamics for planning from pixels
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2019
Later among the works it cites.
Exploring model-based planning with policy networks
T. Wang and J. Ba · 2020
Closest in time.
Model-predictive control via cross-entropy and gradient-based optimization
H. Bharadhwaj, K. Xie, and F. Shkurti · 2020
Closest in time.
The differentiable cross-entropy method
B. Amos and D. Yarats · 2020
Closest in time.