Fetching the paper…
Reading the bibliography…
When faced with a novel scenario, it can be hard to succeed on the first attempt.
Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables
K. Rakelly, A. Zhou, D. Quillen, C. Finn, and S. Levine · 1903
Earlier work this paper cites.
Search on the Replay Buffer: Bridging Planning and Reinforcement Learning
B. Eysenbach, R. Salakhutdinov, and S. Levine · 1906
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
D. Pomerleau · 1989
Earlier work this paper cites.
Computational approaches to motor learning by imitation
S. Schaal, A. Ijspeert, and A. Billard · 2002
Earlier work this paper cites.
Self-Supervised Policy Adaptation during Deployment
N. Hansen, R. Jangir, Y. Sun, G. Alenyà, P. Abbeel, A. A. Efros, L. Pinto, and X. Wang · 2007
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2008
Earlier work this paper cites.
One Solution is Not All You Need: Few-Shot Extrapolation via Structured MaxEnt RL
S. Kumar, A. Kumar, S. Levine, and C. Finn · 2010
Earlier work this paper cites.
Recovery RL: Safe Reinforcement Learning with Learned Recovery Zones
B. Thananjeyan, A. Balakrishna, S. Nair, M. Luo, K. Srinivasan, M. Hwang, J. E. Gonzalez, J. Ibarz, C. Finn, and K. Goldberg · 2010
Earlier work this paper cites.
Dual arm manipulation—A survey
C. Smith, Y. Karayiannidis, L. Nalpantidis, X. Gratal, P. Qi, D. V. Dimarogonas, and D. Kragic · 2012
Cited alongside, same era.
Error detection and surprise in stochastic robot actions
L. Y. Ku, D. Ruiken, E. Learned-Miller, and R. Grupen · 2015
Cited alongside, same era.
ShapeNet: An Information-Rich 3D Model Repository
A. X. Chang, T. Funkhouser, L. Guibas, P. Hanrahan, Q. Huang, Z. Li, S. Savarese, M. Savva, S. Song, H. Su, J. Xiao, L. Yi, and F. Yu · 2015
Cited alongside, same era.
Robust and efficient transfer learning with hidden parameter markov decision processes
T. W. Killian, S. Daulton, G. Konidaris, and F. Doshi-Velez · 2017
Cited alongside, same era.
Deep Imitation Learning for Bimanual Robotic Manipulation
F. Xie, A. Chowdhury, M. C. De Paolis Kaluza, L. Zhao, L. Wong, and R. Yu · 2020
Cited alongside, same era.
What matters in learning from offline human demonstrations for robot manipulation
A. Mandlekar, D. Xu, J. Wong, S. Nasiriany, C. Wang, R. Kulkarni, L. Fei-Fei, S. Savarese, Y. Zhu, and R. Martín-Martín · 2021
Later among the works it cites.
Behavior Transformers: Cloning $k$ modes with one stone
N. M. Shafiullah, Z. Cui, A. A. Altanzaya, and L. Pinto · 2022
Later among the works it cites.
Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets
F. Ebert, Y. Yang, K. Schmeckpeper, B. Bucher, G. Georgakis, K. Daniilidis, C. Finn, and S. Levine · 2022
Later among the works it cites.
Learning fine-grained bimanual manipulation with low-cost hardware
T. Zhao, V. Kumar, S. Levine, and C. Finn · 2023
Later among the works it cites.
Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Ho, A. Jain, and P. Abbeel · 2020
Cited alongside, same era.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, R. Martín-Martín, A. Joshi, S. Nasiriany, and Y. Zhu · 2020
Cited alongside, same era.
Adaptable Agent Populations via a Generative Model of Policies
K. Derek and P. Isola · 2021
Cited alongside, same era.
Stabilize to Act: Learning to Coordinate for Bimanual Manipulation
J. Grannen, Y. Wu, B. Vu, and D. Sadigh
Cited in the paper.
Is imitation learning the route to humanoid robots?
S. Schaal
Cited in the paper.
A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning
S. Ross, G. Gordon, and D. Bagnell
Cited in the paper.
On the Sample Complexity of Stability Constrained Imitation Learning
S. Tu, A. Robey, T. Zhang, and N. Matni
Cited in the paper.
Model-based runtime monitoring with interactive imitation learning, 2023
H. Liu, S. Dass, R. Martín-Martín, and Y. Zhu · 2023
Later among the works it cites.
RT-1: Robotics Transformer for Real-World Control at Scale
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, J. Ibarz, B. Ichter, A. Irpan, T. Jackson, S. Jesmonth, N. Joshi, R. Julian, D. Kalashnikov, Y. Kuang, I. Leal, K.-H. Lee, S. Levine, Y. Lu, U. Malla, D. Manjunath, I. Mordatch, O. Nachum, C. Parada, J. Peralta, E. Perez, K. Pertsch, J. Quiambao, K. Rao, M. Ryoo, G. Salazar, P. Sanketi, K. Sayed, J. Singh, S. Sontakke, A. Stone, C. Tan, H. Tran, V. Vanhoucke, S. Vega, Q. Vuong, F. Xia, T. Xiao, P. Xu, S. Xu, T. Yu, and B. Zitkovich · 2023
Later among the works it cites.