Fetching the paper…
Reading the bibliography…
To solve tasks in complex environments, robots need to learn from experience.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2013
Earlier work this paper cites.
A survey on policy search for robotics
M. P. Deisenroth, G. Neumann, J. Peters, et al · 2013
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
D. J. Rezende, S. Mohamed, and D. Wierstra · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2015
Earlier work this paper cites.
Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours, 2015
L. Pinto and A. Gupta · 2015
Earlier work this paper cites.
Adapting deep visuomotor representations with weak pairwise constraints, 2015
E. Tzeng, C. Devin, J. Hoffman, C. Finn, P. Abbeel, S. Levine, K. Saenko, and T. Darrell · 2015
Earlier work this paper cites.
Improving pilco with bayesian neural network dynamics models
Y. Gal, R. McAllister, and C. E. Rasmussen · 2016
Earlier work this paper cites.
Sim-to-real robot learning from pixels with progressive nets, 2016
A. A. Rusu, M. Vecerik, T. Rothörl, N. Heess, R. Pascanu, and R. Hadsell · 2016
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
C. Finn, I. Goodfellow, and S. Levine · 2016
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Earlier work this paper cites.
Deep visual foresight for planning robot motion
C. Finn and S. Levine · 2017
Earlier work this paper cites.
Learning image-conditioned dynamics models for control of under-actuated legged millirobots, 2017
A. Nagabandi, G. Yang, T. Asmar, R. Pandya, G. Kahn, S. Levine, and R. S. Fearing · 2017
Earlier work this paper cites.
Visual foresight: Model-based deep reinforcement learning for vision-based robotic control
F. Ebert, C. Finn, S. Dasari, A. Xie, A. Lee, and S. Levine · 2018
Earlier work this paper cites.
Learning latent dynamics for planning from pixels
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2018
Earlier work this paper cites.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Earlier work this paper cites.
Rainbow: Combining improvements in deep reinforcement learning
M. Hessel, J. Modayil, H. Van Hasselt, T. Schaul, G. Ostrovski, W. Dabney, D. Horgan, B. Piot, M. Azar, and D. Silver · 2018
Earlier work this paper cites.
Sim-to-real transfer of robotic control with dynamics randomization
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel · 2018
Earlier work this paper cites.
Learning dexterous in-hand manipulation, 2018
OpenAI, M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba · 2018
Earlier work this paper cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation, 2018
D. Kalashnikov, A. Irpan, P. Pastor, J. Ibarz, A. Herzog, E. Jang, D. Quillen, E. Holly, M. Kalakrishnan, V. Vanhoucke, and S. Levine · 2018
Cited alongside, same era.
Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen · 2018
Cited alongside, same era.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
K. Chua, R. Calandra, R. McAllister, and S. Levine · 2018
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi · 2019
Cited alongside, same era.
Model-predictive policy learning with uncertainty regularization for driving in dense traffic
M. Henaff, A. Canziani, and Y. LeCun · 2019
Mastering visual continuous control: Improved data-augmented reinforcement learning
D. Yarats, R. Fergus, A. Lazaric, and L. Pinto · 2021
Later among the works it cites.
Learning to walk in minutes using massively parallel deep reinforcement learning, 2021
N. Rudin, D. Hoeller, P. Reist, and M. Hutter · 2021
Later among the works it cites.
Rma: Rapid motor adaptation for legged robots, 2021
A. Kumar, Z. Fu, D. Pathak, and J. Malik · 2021
Later among the works it cites.
Blind bipedal stair traversal via sim-to-real reinforcement learning, 2021
J. Siekmann, K. Green, J. Warila, A. Fern, and J. Hurst · 2021
Later among the works it cites.
Legged robots that keep on learning: Fine-tuning locomotion policies in the real world, 2021
L. Smith, J. C. Kew, X. B. Peng, S. Ha, J. Tan, and S. Levine · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Mastering atari, go, chess and shogi by planning with a learned model
J. Schrittwieser, I. Antonoglou, T. Hubert, K. Simonyan, L. Sifre, S. Schmitt, A. Guez, E. Lockhart, D. Hassabis, T. Graepel, et al · 2019
Cited alongside, same era.
Robonet: Large-scale multi-robot learning, 2019
S. Dasari, F. Ebert, S. Tian, S. Nair, B. Bucher, K. Schmeckpeper, S. Singh, S. Levine, and C. Finn · 2019
Cited alongside, same era.
Improvisation through physical understanding: Using novel objects as tools with visual foresight
A. Xie, F. Ebert, S. Levine, and C. Finn · 2019
Cited alongside, same era.
Deep reinforcement learning for industrial insertion tasks with visual inputs and natural rewards, 2019
G. Schoettler, A. Nair, J. Luo, S. Bahl, J. A. Ojea, E. Solowjow, and S. Levine · 2019
Cited alongside, same era.
Data efficient reinforcement learning for legged robots, 2019
Y. Yang, K. Caluwaerts, A. Iscen, T. Zhang, J. Tan, and V. Sindhwani · 2019
Cited alongside, same era.
Solar: deep structured representations for model-based reinforcement learning
M. Zhang, S. Vikram, L. Smith, P. Abbeel, M. Johnson, and S. Levine · 2019
Cited alongside, same era.
Deep dynamics models for learning dexterous manipulation, 2019
A. Nagabandi, K. Konoglie, S. Levine, and V. Kumar · 2019
Cited alongside, same era.
Mt-opt: Continuous multi-task robotic reinforcement learning at scale, 2021
D. Kalashnikov, J. Varley, Y. Chebotar, B. Swanson, R. Jonschkowski, C. Finn, S. Levine, and K. Hausman · 2021
Later among the works it cites.
Bridge data: Boosting generalization of robotic skills with cross-domain datasets, 2021
F. Ebert, Y. Yang, K. Schmeckpeper, B. Bucher, G. Georgakis, K. Daniilidis, C. Finn, and S. Levine · 2021
Later among the works it cites.
Coarse-to-fine q-attention: Efficient learning for visual robotic manipulation via discretisation, 2021
S. James, K. Wada, T. Laidlow, and A. J. Davison · 2021
Later among the works it cites.
Flingbot: The unreasonable effectiveness of dynamic manipulation for cloth unfolding
H. Ha and S. Song · 2021
Later among the works it cites.
Q-attention: Enabling efficient learning for vision-based robotic manipulation, 2021
S. James and A. J. Davison · 2021
Later among the works it cites.
Dreamerpro: Reconstruction-free model-based reinforcement learning with prototypical representations
F. Deng, I. Jang, and S. Ahn · 2021
Later among the works it cites.
Dreaming: Model-based reinforcement learning by latent imagination without reconstruction
M. Okada and T. Taniguchi · 2021
Later among the works it cites.
Adversarial motion priors make good substitutes for complex reward functions, 2022
A. Escontrela, X. B. Peng, W. Yu, T. Zhang, A. Iscen, K. Goldberg, and P. Abbeel · 2022
Closest in time.
Learning robust perceptive locomotion for quadrupedal robots in the wild
T. Miki, J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter · 2022
Closest in time.
Viking: Vision-based kilometer-scale navigation with geographic hints, 2022
D. Shah and S. Levine · 2022
Closest in time.
Imitate and repurpose: Learning reusable robot movement skills from human and animal behaviors, 2022
S. Bohez, S. Tunyasuvunakool, P. Brakel, F. Sadeghi, L. Hasenclever, Y. Tassa, E. Parisotto, J. Humplik, T. Haarnoja, R. Hafner, M. Wulfmeier, M. Neunert, B. Moran, N. Siegel, A. Huber, F. Romano, N. Batchelor, F. Casarini, J. Merel, R. Hadsell, and N. Heess · 2022
Closest in time.
Robotic telekinesis: Learning a robotic hand imitator by watching humans on youtube, 2022
A. Sivakumar, K. Shaw, and D. Pathak · 2022
Closest in time.
Fast and efficient locomotion via learned gait transitions
Y. Yang, T. Zhang, E. Coumans, J. Tan, and B. Boots · 2022
Closest in time.
Safe reinforcement learning for legged locomotion, 2022
T.-Y. Yang, T. Zhang, L. Luu, S. Ha, J. Tan, and W. Yu · 2022
Closest in time.
Information prioritization through empowerment in visual model-based rl
H. Bharadhwaj, M. Babaeizadeh, D. Erhan, and S. Levine · 2022
Closest in time.