Fetching the paper…
Reading the bibliography…
We introduce BridgeData V2, a large and diverse dataset of robotic manipulation behaviors designed to facilitate research on scalable robot learning.
Learning to Achieve Goals
L. P. Kaelbling · 1993
Earlier work this paper cites.
Universal Value Function Approximators
T. Schaul, D. Horgan, K. Gregor, and D. Silver · 2015
Earlier work this paper cites.
Learning Structured Output Representation using Deep Conditional Generative Models
K. Sohn, X. Yan, and H. Lee · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2015
Earlier work this paper cites.
End-to-End Training of Deep Visuomotor Policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Earlier work this paper cites.
Generative Adversarial Imitation Learning
J. Ho and S. Ermon · 2016
Earlier work this paper cites.
Supersizing Self-supervision: Learning to Grasp from 50K Tries and 700 Robot Hours
L. Pinto and A. Gupta · 2016
Earlier work this paper cites.
Learning to Poke by Poking: Experiential Learning of Intuitive Physics
P. Agrawal, A. Nair, P. Abbeel, J. Malik, and S. Levine · 2016
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
C. Finn, I. Goodfellow, and S. Levine · 2016
Earlier work this paper cites.
Learning Hand-Eye Coordination for Robotic Grasping with Deep Learning and Large-Scale Data Collection
S. Levine, P. Pastor, A. Krizhevsky, and D. Quillen · 2017
Earlier work this paper cites.
Learning to push by grasping: Using multiple tasks for effective learning
L. Pinto and A. Gupta · 2017
Earlier work this paper cites.
Combining self-supervised learning and imitation for vision-based rope manipulation
A. Nair, D. Chen, P. Agrawal, P. Isola, P. Abbeel, J. Malik, and S. Levine · 2017
Earlier work this paper cites.
Ai2-thor: An interactive 3d environment for visual ai
E. Kolve, R. Mottaghi, W. Han, E. VanderBilt, L. Weihs, A. Herrasti, M. Deitke, K. Ehsani, D. Gordon, Y. Zhu, et al · 2017
Earlier work this paper cites.
FiLM: Visual Reasoning with a General Conditioning Layer, Dec. 2017
E. Perez, F. Strub, H. de Vries, V. Dumoulin, and A. Courville · 2017
Earlier work this paper cites.
Multiple interactions made easy (mime): Large scale demonstrations data for imitation
P. Sharma, L. Mohan, L. Pinto, and A. Gupta · 2018
Earlier work this paper cites.
Roboturk: A crowdsourcing platform for robotic skill learning through imitation
A. Mandlekar, Y. Zhu, A. Garg, J. Booher, M. Spero, A. Tung, J. Gao, J. Emmons, A. Gupta, E. Orbay, et al · 2018
Earlier work this paper cites.
Visual foresight: Model-based deep reinforcement learning for vision-based robotic control
F. Ebert, C. Finn, S. Dasari, A. Xie, A. Lee, and S. Levine · 2018
Earlier work this paper cites.
Deep imitation learning for complex manipulation tasks from virtual reality teleoperation
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel · 2018
Earlier work this paper cites.
Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation
D. Kalashnikov, A. Irpan, P. Pastor, J. Ibarz, A. Herzog, E. Jang, D. Quillen, E. Holly, M. Kalakrishnan, V. Vanhoucke, et al · 2018
Earlier work this paper cites.
Robot learning in homes: Improving generalization and reducing dataset bias
A. Gupta, A. Murali, D. Gandhi, and L. Pinto · 2018
Earlier work this paper cites.
One-shot imitation from observing humans via domain-adaptive meta-learning
T. Yu, C. Finn, A. Xie, S. Dasari, T. Zhang, P. Abbeel, and S. Levine · 2018
Cited alongside, same era.
Gonet: A semi-supervised deep learning approach for traversability estimation
N. Hirose, A. Sadeghian, M. Vázquez, P. Goebel, and S. Savarese · 2018
Cited alongside, same era.
Universal sentence encoder for english
D. Cer, Y. Yang, S.-y. Kong, N. Hua, N. Limtiaco, R. S. John, N. Constant, M. Guajardo-Cespedes, S. Yuan, C. Tar, et al · 2018
Cited alongside, same era.
Robonet: Large-scale multi-robot learning
S. Dasari, F. Ebert, S. Tian, S. Nair, B. Bucher, K. Schmeckpeper, S. Singh, S. Levine, and C. Finn · 2019
Cited alongside, same era.
Scaling Data-Driven Robotics With Reward Sketching and Batch Reinforcement Learning
S. Cabi, S. G. Colmenarejo, A. Novikov, K. Konyushkova, S. Reed, R. Jeong, K. Zolna, Y. Aytar, D. Budden, M. Vecerik, et al · 2019
Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills
Y. Chebotar, K. Hausman, Y. Lu, T. Xiao, D. Kalashnikov, J. Varley, A. Irpan, B. Eysenbach, R. Julian, C. Finn, and S. Levine · 2021
Later among the works it cites.
Visual imitation made easy
S. Young, D. Gandhi, S. Tulsiani, A. Gupta, P. Abbeel, and L. Pinto · 2021
Later among the works it cites.
Maniskill: Generalizable manipulation skill benchmark with large-scale demonstrations
T. Mu, Z. Ling, F. Xiang, D. Yang, X. Li, S. Tao, Z. Huang, Z. Jia, and H. Su · 2021
Later among the works it cites.
Offline Reinforcement Learning With Implicit Q-Learning
I. Kostrikov, A. Nair, and S. Levine · 2021
Later among the works it cites.
C-learning: Learning to achieve goals via recursive classification
B. Eysenbach, R. Salakhutdinov, and S. Levine · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Habitat: A platform for embodied ai research
M. Savva, A. Kadian, O. Maksymets, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, et al · 2019
Cited alongside, same era.
Multilingual Universal Sentence Encoder for Semantic Retrieval, July 2019
Y. Yang, D. Cer, A. Ahmad, M. Guo, J. Law, N. Constant, G. H. Abrego, S. Yuan, C. Tar, Y.-H. Sung, et al · 2019
Cited alongside, same era.
Efficientnet: Rethinking model scaling for convolutional neural networks
M. Tan and Q. Le · 2019
Cited alongside, same era.
Language models are few-shot learners
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Cited alongside, same era.
COG: Connecting New Skills to Past Experience with Offline Reinforcement Learning
A. Singh, A. Yu, J. Yang, J. Zhang, A. Kumar, and S. Levine · 2020
Cited alongside, same era.
Grasping in the wild: Learning 6dof closed-loop grasping from low-cost demonstrations
S. Song, A. Zeng, J. Lee, and T. Funkhouser · 2020
Cited alongside, same era.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine · 2020
Cited alongside, same era.
Rt-1: Robotics transformer for real-world control at scale
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, et al · 2022
Later among the works it cites.
How to spend your robot time: Bridging kickstarting and offline reinforcement learning for vision-based robotic manipulation
A. X. Lee, C. Devin, J. T. Springenberg, Y. Zhou, T. Lampe, A. Abdolmaleki, and K. Bousmalis · 2022
Later among the works it cites.
Socially compliant navigation dataset (scand): A large-scale dataset of demonstrations for social navigation
H. Karnan, A. Nair, X. Xiao, G. Warnell, S. Pirk, A. Toshev, J. Hart, J. Biswas, and P. Stone · 2022
Later among the works it cites.
Tartandrive: A large-scale dataset for learning off-road dynamics models
S. Triest, M. Sivaprakasam, S. J. Wang, W. Wang, A. M. Johnson, and S. Scherer · 2022
Later among the works it cites.
Gnm: A general navigation model to drive any robot
D. Shah, A. Sridhar, A. Bhorkar, N. Hirose, and S. Levine · 2022
Later among the works it cites.
Pre-training for robots: Offline rl enables learning new tasks from a handful of trials
A. Kumar, A. Singh, F. Ebert, Y. Yang, C. Finn, and S. Levine · 2022
Later among the works it cites.
Contrastive learning as goal-conditioned reinforcement learning
B. Eysenbach, T. Zhang, S. Levine, and R. R. Salakhutdinov · 2022
Later among the works it cites.
Behavior transformers: Cloning k k modes with one stone
N. M. M. Shafiullah, Z. J. Cui, A. Altanzaya, and L. Pinto · 2022
Later among the works it cites.
Learning fine-grained bimanual manipulation with low-cost hardware
T. Z. Zhao, V. Kumar, S. Levine, and C. Finn · 2023
Closest in time.
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Closest in time.
Roboagent: Towards sample efficient robot manipulation with semantic augmentations and action chunking
H. Bharadhwaj, J. Vakil, M. Sharma, A. Gupta, S. Tulsiani, and V. Kumar · 2023
Closest in time.
Robocat: A self-improving foundation agent for robotic manipulation
K. Bousmalis, G. Vezzani, D. Rao, C. Devin, A. X. Lee, M. Bauza, T. Davchev, Y. Zhou, A. Gupta, A. Raju, et al · 2023
Closest in time.
Idql: Implicit q-learning as an actor-critic method with diffusion policies
P. Hansen-Estruch, I. Kostrikov, M. Janner, J. G. Kuba, and S. Levine · 2023
Closest in time.
Rh20t: A robotic dataset for learning diverse skills in one-shot
H.-S. Fang, H. Fang, Z. Tang, J. Liu, J. Wang, H. Zhu, and C. Lu · 2023
Closest in time.
Stabilizing contrastive rl: Techniques for offline goal reaching, 2023
C. Zheng, B. Eysenbach, H. Walke, P. Yin, K. Fang, R. Salakhutdinov, and S. Levine · 2023
Closest in time.