Fetching the paper…
Reading the bibliography…
Gradient-based methods enable efficient search capabilities in high dimensions.
Elementary Classical Analysis
J. Marsden and M. Hoffman · 1974
Earlier work this paper cites.
Model predictive heuristic control: Applications to industrial processes
J. Richalet, A. Rault, J. Testud, and J. Papon · 1978
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
D. Pomerleau · 1989
Earlier work this paper cites.
Exploiting model uncertainty estimates for safe dynamic control learning
J. Schneider · 1996
Earlier work this paper cites.
A comparison of direct and model-based reinforcement learning
C. Atkeson and J. Santamaria · 1997
Earlier work this paper cites.
An introduction to statistical modeling of extreme values
S. Coles · 2001
Earlier work this paper cites.
Gamma-Convergence for Beginners
A. Braides · 2002
Earlier work this paper cites.
Gaussian process model based predictive control
J. Kocijan, R. Murray-Smith, C. Rasmussen, and A. Girard · 2004
Earlier work this paper cites.
A tutorial on the cross-entropy method
P.-T. de Boer, D. P. Kroese, S. Mannor, and R. Y. Rubinstein · 2005
Earlier work this paper cites.
Using inaccurate models in reinforcement learning
P. Abbeel, M. Quigley, and A. Y. Ng · 2006
Earlier work this paper cites.
Relative entropy policy search
J. Peters, K. Muelling, and Y. Altun · 2010
Earlier work this paper cites.
Pilco: A model-based and data-efficient approach to policy search
M. Deisenroth and C. Rasmussen · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning, 2011
S. Ross, G. J. Gordon, and J. A. Bagnell · 2011
Earlier work this paper cites.
A connection between score matching and denoising autoencoders
P. Vincent · 2011
Earlier work this paper cites.
Randomized smoothing for stochastic optimization, 2012
J. C. Duchi, P. L. Bartlett, and M. J. Wainwright · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Stochastic first- and zeroth-order methods for nonconvex stochastic programming, 2013
S. Ghadimi and G. Lan · 2013
Earlier work this paper cites.
Generative adversarial networks, 2014
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Embed to control: A locally linear latent dynamics model for control from raw images, 2015
M. Watter, J. T. Springenberg, J. Boedecker, and M. Riedmiller · 2015
Earlier work this paper cites.
Variational inference with normalizing flows
D. Rezende and S. Mohamed · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics, 2015
J. Sohl-Dickstein, E. A. Weiss, N. Maheswaranathan, and S. Ganguli · 2015
Earlier work this paper cites.
Model predictive path integral control using covariance variable importance sampling, 2015
G. Williams, A. Aldrich, and E. Theodorou · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
A theoretically grounded application of dropout in recurrent neural networks, 2016
Y. Gal and Z. Ghahramani · 2016
Earlier work this paper cites.
Generative adversarial imitation learning, 2016
J. Ho and S. Ermon · 2016
Earlier work this paper cites.
Gradient descent only converges to minimizers
J. D. Lee, M. Simchowitz, M. I. Jordan, and B. Recht · 2016
Earlier work this paper cites.
Adam: A method for stochastic optimization, 2017
D. P. Kingma and J. Ba · 2017
Cited alongside, same era.
Deep visual foresight for planning robot motion, 2017
C. Finn and S. Levine · 2017
Cited alongside, same era.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models, 2018
K. Chua, R. Calandra, R. McAllister, and S. Levine · 2018
Cited alongside, same era.
Accurate uncertainties for deep learning using calibrated regression, 2018
V. Kuleshov, N. Fenner, and S. Ermon · 2018
Cited alongside, same era.
Model-ensemble trust-region policy optimization, 2018
T. Kurutach, I. Clavera, Y. Duan, A. Tamar, and P. Abbeel · 2018
Cited alongside, same era.
Recurrent world models facilitate policy evolution, 2018
D. Ha and J. Schmidhuber · 2018
Cited alongside, same era.
Model error propagation via learned contraction metrics for safe feedback motion planning of unknown systems
G. Chou, N. Ozay, and D. Berenson · 2021
Later among the works it cites.
D4rl: Datasets for deep data-driven reinforcement learning, 2021
J. Fu, A. Kumar, O. Nachum, G. Tucker, and S. Levine · 2021
Later among the works it cites.
Conservative objective models for effective offline model-based optimization, 2021
B. Trabucco, A. Kumar, X. Geng, and S. Levine · 2021
Later among the works it cites.
Model-based reinforcement learning via latent-space collocation, 2021
O. Rybkin, C. Zhu, A. Nagabandi, K. Daniilidis, I. Mordatch, and S. Levine · 2021
Later among the works it cites.
Model-based offline planning, 2021
A. Argenson and G. Dulac-Arnold · 2021
Later among the works it cites.
Awac: Accelerating online reinforcement learning with offline datasets, 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Self-consistent trajectory autoencoder: Hierarchical reinforcement learning with trajectory embeddings, 2018
J. D. Co-Reyes, Y. Liu, A. Gupta, B. Eysenbach, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Learning robust rewards with adversarial inverse reinforcement learning, 2018
J. Fu, K. Luo, and S. Levine · 2018
Cited alongside, same era.
Deep dynamics models for learning dexterous manipulation
A. Nagabandi, K. Konolige, S. Levine, and V. Kumar · 2019
Cited alongside, same era.
Off-policy deep reinforcement learning without exploration, 2019
S. Fujimoto, D. Meger, and D. Precup · 2019
Cited alongside, same era.
A divergence minimization perspective on imitation learning methods, 2019
S. K. S. Ghasemipour, R. Zemel, and S. Gu · 2019
Cited alongside, same era.
Calibrated model-based deep reinforcement learning
A. Malik, V. Kuleshov, J. Song, D. Nemer, H. Seymour, and S. Ermon · 2019
Cited alongside, same era.
A. Nair, A. Gupta, M. Dalal, and S. Levine · 2021
Later among the works it cites.
Offline reinforcement learning with implicit q-learning, 2021
I. Kostrikov, A. Nair, and S. Levine · 2021
Later among the works it cites.
Lyapunov density models: Constraining distribution shift in learning-based control
K. Kang, P. Gradu, J. J. Choi, M. Janner, C. Tomlin, and S. Levine · 2022
Later among the works it cites.
Planning with diffusion for flexible behavior synthesis
M. Janner, Y. Du, J. Tenenbaum, and S. Levine · 2022
Later among the works it cites.
Safe output feedback motion planning from images via learned perception modules and contraction theory
G. Chou, N. Ozay, and D. Berenson · 2022
Later among the works it cites.
Auto-encoding variational bayes, 2022
D. P. Kingma and M. Welling · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models, 2022
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Later among the works it cites.
Design-bench: Benchmarks for data-driven offline model-based optimization, 2022
B. Trabucco, X. Geng, A. Kumar, and S. Levine · 2022
Later among the works it cites.
Risk-averse zero-order trajectory optimization
M. Vlastelica, S. Blaes, C. Pinneri, and G. Martius · 2022
Later among the works it cites.
Do differentiable simulators give better policy gradients?, 2022
H. J. T. Suh, M. Simchowitz, K. Zhang, and R. Tedrake · 2022
Later among the works it cites.
Bundled gradients through contact via randomized smoothing, 2022
H. J. T. Suh, T. Pang, and R. Tedrake · 2022
Later among the works it cites.
Is conditional generative modeling all you need for decision-making?, 2022
A. Ajay, Y. Du, A. Gupta, J. Tenenbaum, T. Jaakkola, and P. Agrawal · 2022
Later among the works it cites.
Transporter networks: Rearranging the visual world for robotic manipulation, 2022
A. Zeng, P. Florence, J. Tompson, S. Welker, J. Chien, M. Attarian, T. Armstrong, I. Krasin, D. Duong, A. Wahid, V. Sindhwani, and J. Lee · 2022
Later among the works it cites.
Statistical safety and robustness guarantees for feedback motion planning of unknown underactuated stochastic systems
C. Knuth, G. Chou, J. Reese, and J. Moore · 2023
Closest in time.
Diffusion policy: Visuomotor policy learning via action diffusion, 2023
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Closest in time.
Interpreting and improving diffusion models using the euclidean distance function, 2023
F. Permenter and C. Yuan · 2023
Closest in time.
When data geometry meets deep function: Generalizing offline reinforcement learning
J. Li, X. Zhan, H. Xu, X. Zhu, J. Liu, and Y.-Q. Zhang · 2023
Closest in time.
Underactuated Robotics
R. Tedrake · 2023
Closest in time.
G. Chou and R. Tedrake · 2023
Closest in time.
Mastering diverse domains through world models, 2023
D. Hafner, J. Pasukonis, J. Ba, and T. Lillicrap · 2023
Closest in time.
Global planning for contact-rich manipulation via local smoothing of quasi-dynamic contact models, 2023
T. Pang, H. J. T. Suh, L. Yang, and R. Tedrake · 2023
Closest in time.