Fetching the paper…
Reading the bibliography…
Few-shot imitation learning relies on only a small amount of task-specific demonstrations to efficiently adapt a policy for a given downstream tasks.
Concept2robot: Learning manipulation concepts from instructions and human demonstrations
L. Shao, T. Migimatsu, Q. Zhang, K. Yang, and J. Bohg · 2021
Earlier work this paper cites.
Language conditioned imitation learning over unstructured data
C. Lynch and P. Sermanet · 2021
Earlier work this paper cites.
Ella: Exploration through learned language abstraction
S. Mirchandani, S. Karamcheti, and D. Sadigh · 2021
Earlier work this paper cites.
What matters in learning from offline human demonstrations for robot manipulation
A. Mandlekar, D. Xu, J. Wong, S. Nasiriany, C. Wang, R. Kulkarni, L. Fei-Fei, S. Savarese, Y. Zhu, and R. Martín-Martín · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Earlier work this paper cites.
Learning and retrieval from prior data for skill-based imitation learning
S. Nasiriany, T. Gao, A. Mandlekar, and Y. Zhu · 2022
Earlier work this paper cites.
Rt-1: Robotics transformer for real-world control at scale
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, et al · 2022
Earlier work this paper cites.
Gmflow: Learning optical flow via global matching
H. Xu, J. Zhang, J. Cai, H. Rezatofighi, and D. Tao · 2022
Earlier work this paper cites.
Can foundation models perform zero-shot task specification for robot manipulation?
Y. Cui, S. Niekum, A. Gupta, V. Kumar, and A. Rajeswaran · 2022
Earlier work this paper cites.
Behavior retrieval: Few-shot imitation learning by querying unlabeled datasets
M. Du, S. Nair, D. Sadigh, and C. Finn · 2023
Earlier work this paper cites.
Libero: Benchmarking knowledge transfer for lifelong robot learning
B. Liu, Y. Zhu, C. Gao, Y. Feng, Q. Liu, Y. Zhu, and P. Stone · 2023
Earlier work this paper cites.
Bridgedata v2: A dataset for robot learning at scale
H. Walke, K. Black, A. Lee, M. J. Kim, M. Du, C. Zheng, T. Zhao, P. Hansen-Estruch, Q. Vuong, A. He, V. Myers, K. Fang, C. Finn, and S. Levine · 2023
Earlier work this paper cites.
An unbiased look at datasets for visuo-motor pre-training
S. Dasari, M. K. Srirama, U. Jain, and A. Gupta · 2023
Earlier work this paper cites.
Real-world robot learning with masked visual pre-training
I. Radosavovic, T. Xiao, S. James, P. Abbeel, J. Malik, and T. Darrell · 2023
Cited alongside, same era.
Octo: An open-source generalist robot policy
Octo Model Team, D. Ghosh, H. Walke, K. Pertsch, K. Black, O. Mees, S. Dasari, J. Hejna, C. Xu, J. Luo, T. Kreiman, Y. Tan, D. Sadigh, C. Finn, and S. Levine · 2023
Cited alongside, same era.
Voyager: An open-ended embodied agent with large language models
G. Wang, Y. Xie, Y. Jiang, A. Mandlekar, C. Xiao, Y. Zhu, L. Fan, and A. Anandkumar · 2023
Cited alongside, same era.
Interactive language: Talking to robots in real time
C. Lynch, A. Wahid, J. Tompson, T. Ding, J. Betker, R. Baruch, T. Armstrong, and P. Florence · 2023
Cited alongside, same era.
Language embedded radiance fields for zero-shot task-oriented grasping
A. Rashid, S. Sharma, C. M. Kim, J. Kerr, L. Y. Chen, A. Kanazawa, and K. Goldberg · 2023
Cited alongside, same era.
Liv: Language-image representations and rewards for robotic control
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Later among the works it cites.
Goal-conditioned imitation learning using score-based diffusion policies
M. Reuss, M. Li, X. Jia, and R. Lioutikov · 2023
Later among the works it cites.
R3m: A universal visual representation for robot manipulation
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta · 2023
Later among the works it cites.
Dinov2: Learning robust visual features without supervision
M. Oquab, T. Darcet, T. Moutakanni, H. Vo, M. Szafraniec, V. Khalidov, P. Fernandez, D. Haziza, F. Massa, A. El-Nouby, et al · 2023
Later among the works it cites.
Rt-h: Action hierarchies using language
S. Belkhale, T. Ding, T. Xiao, P. Sermanet, Q. Vuong, J. Tompson, Y. Chebotar, D. Dwibedi, and D. Sadigh · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. J. Ma, W. Liang, V. Som, V. Kumar, A. Zhang, O. Bastani, and D. Jayaraman · 2023
Cited alongside, same era.
Language-driven representation learning for robotics
S. Karamcheti, S. Nair, A. S. Chen, T. Kollar, C. Finn, D. Sadigh, and P. Liang · 2023
Cited alongside, same era.
Roboclip: One demonstration is enough to learn robot policies
S. A. Sontakke, J. Zhang, S. Arnold, K. Pertsch, E. Biyik, D. Sadigh, C. Finn, and L. Itti · 2023
Cited alongside, same era.
Improving long-horizon imitation through instruction prediction
J. Hejna, P. Abbeel, and L. Pinto · 2023
Cited alongside, same era.
Any-point trajectory modeling for policy learning
C. Wen, X. Lin, J. So, K. Chen, Q. Dou, Y. Gao, and P. Abbeel · 2023
Cited alongside, same era.
Robotap: Tracking arbitrary points for few-shot visual imitation
M. Vecerik, C. Doersch, Y. Yang, T. Davchev, Y. Aytar, G. Zhou, R. Hadsell, L. Agapito, and J. Scholz · 2023
Cited alongside, same era.
Learning to Act from Actionless Video through Dense Correspondences
P.-C. Ko, J. Mao, Y. Du, S.-H. Sun, and J. B. Tenenbaum · 2023
Cited alongside, same era.
Closest in time.
Distilling and retrieving generalizable knowledge for robot manipulation via language corrections
L. Zha, Y. Cui, L.-H. Lin, M. Kwon, M. G. Arenas, A. Zeng, F. Xia, and D. Sadigh · 2024
Closest in time.
Mobile aloha: Learning bimanual mobile manipulation with low-cost whole-body teleoperation
Z. Fu, T. Z. Zhao, and C. Finn · 2024
Closest in time.
Data quality in imitation learning
S. Belkhale, Y. Cui, and D. Sadigh · 2024
Closest in time.
Policy learning with a language bottleneck
M. Srivastava, C. Colas, D. Sadigh, and J. Andreas · 2024
Closest in time.
Dinobot: Robot manipulation via retrieval and alignment with vision foundation models
N. D. Palo and E. Johns · 2024
Closest in time.
Track2act: Predicting point tracks from internet videos enables diverse zero-shot robot manipulation, 2024
H. Bharadhwaj, R. Mottaghi, A. Gupta, and S. Tulsiani · 2024
Closest in time.
Vision-based manipulation from single human video with open-world object graphs
Y. Zhu, A. Lim, P. Stone, and Y. Zhu · 2024
Closest in time.