Fetching the paper…
Reading the bibliography…
We introduce MuJoCo MPC (MJPC), an open-source, interactive application and software framework for real-time predictive control, based on MuJoCo physics.
On the general theory of control systems
Rudolf E. Kalman · 1960
Earlier work this paper cites.
When Is a Linear Control System Optimal?
Rudolf E. Kalman · 1964
Earlier work this paper cites.
Random optimization
J. Matyas · 1965
Earlier work this paper cites.
Sufficient conditions for the optimal control of nonlinear systems
O. L. Mangasarian · 1966
Earlier work this paper cites.
Trajectory optimization for an Apollo-type vehicle under entry conditions encountered during lunar return
R. E. Smith and J. W. Yound · 1967
Earlier work this paper cites.
Differential Dynamic Programming
David H. Jacobson and David Q. Mayne · 1970
Earlier work this paper cites.
Apollo 11 Mission Report , volume 238
NASA · 1971
Earlier work this paper cites.
Optimal stochastic linear systems with exponential performance criteria and their relation to deterministic differential games
David Jacobson · 1973
Earlier work this paper cites.
Evolutionsstrategie
Ingo Rechenberg · 1973
Earlier work this paper cites.
Model predictive heuristic control: Applications to industrial processes
Jacques Richalet, André Rault, J. L. Testud, and J. Papon · 1978
Earlier work this paper cites.
Risk-sensitive linear/quadratic/Gaussian control
Peter Whittle · 1981
Earlier work this paper cites.
Genetic algorithms
John H. Holland · 1992
Earlier work this paper cites.
Direct and indirect methods for trajectory optimization
Oskar Von Stryk and Roland Bulirsch · 1992
Earlier work this paper cites.
Numerical solution of optimal control problems by direct collocation
Oskar Von Stryk · 1993
Earlier work this paper cites.
Survey of numerical methods for trajectory optimization
John T. Betts · 1998
Earlier work this paper cites.
Completely derandomized self-adaptation in evolution strategies
Nikolaus Hansen and Andreas Ostermeier · 2001
Earlier work this paper cites.
A model predictive controller for nuclear reactor power
Man Gyun Na, Sun Ho Shin, and Whee Cheol Kim · 2003
Earlier work this paper cites.
The Shadow robot mimics human actions
Paul Tuffield and Hugo Elias · 2003
Cited alongside, same era.
Iterative linear quadratic regulator design for nonlinear biological movement systems
Weiwei Li and Emanuel Todorov · 2004
Cited alongside, same era.
SNOPT: An SQP algorithm for large-scale constrained optimization
Philip E. Gill, Walter Murray, and Michael A. Saunders · 2005
Cited alongside, same era.
On the implementation of an interior-point filter line-search algorithm for large-scale nonlinear programming
Andreas Wächter and Lorenz T. Biegler · 2006
Cited alongside, same era.
Predictive active steering control for autonomous vehicle systems
Paolo Falcone, Francesco Borrelli, Jahan Asgari, Hongtei Eric Tseng, and Davor Hrovat · 2007
Cited alongside, same era.
Synthesis and stabilization of complex behaviors through online trajectory optimization
PhysicsForests: Real-time fluid simulation using machine learning
Lubor Ladicky, SoHyeon Jeong, Nemanja Bartolovic, Marc Pollefeys, and Markus Gross · 2017
Later among the works it cites.
Evolution strategies as a scalable alternative to reinforcement learning
Tim Salimans, Jonathan Ho, Xi Chen, Szymon Sidor, and Ilya Sutskever · 2017
Later among the works it cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Model predictive path integral control: From theory to parallel computation
Grady Williams, Andrew Aldrich, and Evangelos A. Theodorou · 2017
Later among the works it cites.
Simple random search of static linear policies is competitive for reinforcement learning
Horia Mania, Aurelia Guy, and Benjamin Recht · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yuval Tassa, Tom Erez, and Emanuel Todorov · 2012
Cited alongside, same era.
MuJoCo: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Cited alongside, same era.
Fast nonlinear model predictive control: Formulation and industrial process applications
Rodrigo Lopez-Negrete, Fernando J. D’Amato, Lorenz T. Biegler, and Aditya Kumar · 2013
Cited alongside, same era.
Online motion synthesis using sequential Monte Carlo
Perttu Hämäläinen, Sebastian Eriksson, Esa Tanskanen, Ville Kyrki, and Jaakko Lehtinen · 2014
Cited alongside, same era.
Physically-consistent sensor fusion in contact-rich behaviors
Kendall Lowrey, Svetoslav Kolev, Yuval Tassa, Tom Erez, and Emanuel Todorov · 2014
Cited alongside, same era.
Control-limited differential dynamic programming
Yuval Tassa, Nicolas Mansard, and Emo Todorov · 2014
Cited alongside, same era.
Online control of simulated humanoids using particle belief propagation
Perttu Hämäläinen, Joose Rajamäki, and C. Karen Liu · 2015
Cited alongside, same era.
Deep dynamics models for learning dexterous manipulation
Anusha Nagabandi, Kurt Konolige, Sergey Levine, and Vikash Kumar · 2020
Later among the works it cites.
Mastering Atari, Go, chess and shogi by planning with a learned model
Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert, Karen Simonyan, Laurent Sifre, Simon Schmitt, Arthur Guez, Edward Lockhart, Demis Hassabis, Thore Graepel, Timothy Lillicrap, and David Silver · 2020
Later among the works it cites.
Local search for policy iteration in continuous control
Jost Tobias Springenberg, Nicolas Heess, Daniel Mankowitz, Josh Merel, Arunkumar Byravan, Abbas Abdolmaleki, Jackie Kay, Jonas Degrave, Julian Schrittwieser, Yuval Tassa, et al · 2020
Later among the works it cites.
dm_control: Software and tasks for continuous control
Saran Tunyasuvunakool, Alistair Muldal, Yotam Doron, Siqi Liu, Steven Bohez, Josh Merel, Tom Erez, Timothy Lillicrap, Nicolas Heess, and Yuval Tassa · 2020
Later among the works it cites.
Evaluating model-based planning and planner amortization for continuous control
Arunkumar Byravan, Leonard Hasenclever, Piotr Trochim, Mehdi Mirza, Alessandro Davide Ialongo, Yuval Tassa, Jost Tobias Springenberg, Abbas Abdolmaleki, Nicolas Heess, Josh Merel, et al · 2021
Later among the works it cites.
A system for general in-hand object re-orientation
Tao Chen, Jie Xu, and Pulkit Agrawal · 2022
Closest in time.
Current directions in visual perceptual learning
Zhong-Lin Lu and Barbara Anne Dosher · 2022
Closest in time.
MuJoCo Menagerie: A collection of high-quality simulation models for MuJoCo, 2022
MuJoCo Menagerie Contributors · 2022
Closest in time.
Learning to walk in minutes using massively parallel deep reinforcement learning
Nikita Rudin, David Hoeller, Philipp Reist, and Marco Hutter · 2022
Closest in time.
A walk in the park: Learning to walk in 20 minutes with model-free reinforcement learning
Laura Smith, Ilya Kostrikov, and Sergey Levine · 2022
Closest in time.
Unitree A1 Quadruped
Unitree · 2022
Closest in time.
DayDreamer: World models for physical robot learning
Philipp Wu, Alejandro Escontrela, Danijar Hafner, Ken Goldberg, and Pieter Abbeel · 2022
Closest in time.