Fetching the paper…
Reading the bibliography…
While natural language offers a convenient shared interface for humans and robots, enabling robots to interpret and follow language commands remains a longstanding challenge in manipulation.
Learning latent plans from play
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, and P. Sermanet · 1903
Earlier work this paper cites.
Feudal reinforcement learning
P. Dayan and G. E. Hinton · 1992
Earlier work this paper cites.
Analysis and observations from the first amazon picking challenge
N. Correll, K. E. Bekris, D. Berenson, O. Brock, A. Causo, K. Hauser, K. Okada, A. Rodriguez, J. M. Romano, and P. R. Wurman · 2016
Earlier work this paper cites.
J. Mahler, J. Liang, S. Niyaz, M. Laskey, R. Doan, X. Liu, J. A. Ojea, and K. Goldberg · 2017
Earlier work this paper cites.
Pointnet++: Deep hierarchical feature learning on point sets in a metric space
C. R. Qi, L. Yi, H. Su, and L. J. Guibas · 2017
Earlier work this paper cites.
End-to-end driving via conditional imitation learning
F. Codevilla, M. Müller, A. López, V. Koltun, and A. Dosovitskiy · 2018
Earlier work this paper cites.
Hierarchical decision making by generating and following natural language instructions
H. Hu, D. Yarats, Q. Gong, Y. Tian, and M. Lewis · 2019
Earlier work this paper cites.
Language-conditioned imitation learning for robot manipulation tasks
S. Stepputtis, J. Campbell, M. Phielipp, S. Lee, C. Baral, and H. Ben Amor · 2020
Earlier work this paper cites.
Language models are few-shot learners
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Earlier work this paper cites.
Dexycb: A benchmark for capturing hand grasping of objects
Y.-W. Chao, W. Yang, Y. Xiang, P. Molchanov, A. Handa, J. Tremblay, Y. S. Narang, K. Van Wyk, U. Iqbal, S. Birchfield, et al · 2021
Earlier work this paper cites.
Contact-graspnet: Efficient 6-dof grasp generation in cluttered scenes
M. Sundermeyer, A. Mousavian, R. Triebel, and D. Fox · 2021
Earlier work this paper cites.
Open-vocabulary object detection via vision and language knowledge distillation
X. Gu, T.-Y. Lin, W. Kuo, and Y. Cui · 2021
Earlier work this paper cites.
Concept2robot: Learning manipulation concepts from instructions and human demonstrations
L. Shao, T. Migimatsu, Q. Zhang, K. Yang, and J. Bohg · 2021
Earlier work this paper cites.
Grounding language to autonomously-acquired skills via goal generation
A. Akakzia, C. Colas, P.-Y. Oudeyer, M. Chetouani, and O. Sigaud · 2021
Earlier work this paper cites.
Zero-shot task adaptation using natural language
P. Goyal, R. J. Mooney, and S. Niekum · 2021
Earlier work this paper cites.
Ella: Exploration through learned language abstraction
S. Mirchandani, S. Karamcheti, and D. Sadigh · 2021
Earlier work this paper cites.
Accelerating robotic reinforcement learning via parameterized action primitives
M. Dalal, D. Pathak, and R. Salakhutdinov · 2021
Earlier work this paper cites.
Clip-nerf: Text-and-image driven manipulation of neural radiance fields
C. Wang, M. Chai, M. He, D. Chen, and J. Liao · 2021
Earlier work this paper cites.
Learning latent graph dynamics for deformable object manipulation
X. Ma, D. Hsu, and W. S. Lee · 2021
Cited alongside, same era.
Flingbot: The unreasonable effectiveness of dynamic manipulation for cloth unfolding
H. Ha and S. Song · 2021
Cited alongside, same era.
Disentangling dense multi-cable knots
V. Viswanath, J. Grannen, P. Sundaresan, B. Thananjeyan, A. Balakrishna, E. Novoseller, J. Ichnowski, M. Laskey, J. E. Gonzalez, and K. Goldberg · 2021
Cited alongside, same era.
Untangling dense non-planar knots by learning manipulation features and recovery policies
P. Sundaresan, J. Grannen, B. Thananjeyan, A. Balakrishna, J. Ichnowski, E. R. Novoseller, M. Hwang, M. Laskey, J. E. Gonzalez, and K. Goldberg · 2021
Cited alongside, same era.
Generalization through hand-eye coordination: An action space for learning spatially-invariant visuomotor control
C. Wang, R. Wang, A. Mandlekar, L. Fei-Fei, S. Savarese, and D. Xu · 2021
Code as policies: Language model programs for embodied control
J. Liang, W. Huang, F. Xia, P. Xu, K. Hausman, B. Ichter, P. Florence, and A. Zeng · 2022
Later among the works it cites.
Skill-based model-based reinforcement learning
L. X. Shi, J. J. Lim, and Y. Lee · 2022
Later among the works it cites.
Stap: Sequencing task-agnostic policies
C. Agia, T. Migimatsu, J. Wu, and J. Bohg · 2022
Later among the works it cites.
Augmenting reinforcement learning with behavior primitives for diverse manipulation tasks
S. Nasiriany, H. Liu, and Y. Zhu · 2022
Later among the works it cites.
Q-attention: Enabling efficient learning for vision-based robotic manipulation
S. James and A. J. Davison · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Cited alongside, same era.
Perceiver io: A general architecture for structured inputs & outputs
A. Jaegle, S. Borgeaud, J.-B. Alayrac, C. Doersch, C. Ionescu, D. Ding, S. Koppula, D. Zoran, A. Brock, E. Shelhamer, et al · 2021
Cited alongside, same era.
Simple open-vocabulary object detection
M. Minderer, A. Gritsenko, A. Stone, M. Neumann, D. Weissenborn, A. Dosovitskiy, A. Mahendran, A. Arnab, M. Dehghani, Z. Shen, et al · 2022
Cited alongside, same era.
Detecting twenty-thousand classes using image-level supervision
X. Zhou, R. Girdhar, A. Joulin, P. Krähenbühl, and I. Misra · 2022
Cited alongside, same era.
Cliport: What and where pathways for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2022
Cited alongside, same era.
Bc-z: Zero-shot task generalization with robotic imitation learning
E. Jang, A. Irpan, M. Khansari, D. Kappler, F. Ebert, C. Lynch, S. Levine, and C. Finn · 2022
Cited alongside, same era.
Rt-1: Robotics transformer for real-world control at scale
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, et al · 2022
Cited alongside, same era.
kpam: Keypoint affordances for category-level robotic manipulation
L. Manuelli, W. Gao, P. Florence, and R. Tedrake · 2022
Later among the works it cites.
Learning visuo-haptic skewering strategies for robot-assisted feeding
P. Sundaresan, S. Belkhale, and D. Sadigh · 2022
Later among the works it cites.
Grounding dino: Marrying dino with grounded pre-training for open-set object detection
S. Liu, Z. Zeng, T. Ren, F. Li, H. Zhang, J. Yang, C. Li, J. Yang, H. Su, J. Zhu, et al · 2023
Closest in time.
Open-world object manipulation using pre-trained vision-language models
A. Stone, T. Xiao, Y. Lu, K. Gopalakrishnan, K.-H. Lee, Q. Vuong, P. Wohlhart, B. Zitkovich, F. Xia, C. Finn, et al · 2023
Closest in time.
Perceiver-actor: A multi-task transformer for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2023
Closest in time.
Do as i can, not as i say: Grounding language in robotic affordances
A. Brohan, Y. Chebotar, C. Finn, K. Hausman, A. Herzog, D. Ho, J. Ibarz, A. Irpan, E. Jang, R. Julian, et al · 2023
Closest in time.
Text2motion: From natural language instructions to feasible plans
K. Lin, C. Agia, T. Migimatsu, M. Pavone, and J. Bohg · 2023
Closest in time.
Open-world object manipulation using pre-trained vision-language model
A. Stone, T. Xiao, Y. Lu, K. Gopalakrishnan, K.-H. Lee, Q. Vuong, P. Wohlhart, B. Zitkovich, F. Xia, C. Finn, and K. Hausman · 2023
Closest in time.
Plato: Predicting latent affordances through object-centric play
S. Belkhale and D. Sadigh · 2023
Closest in time.
Lerf: Language embedded radiance fields, 2023
J. Kerr, C. M. Kim, K. Goldberg, A. Kanazawa, and M. Tancik · 2023
Closest in time.
Tidybot: Personalized robot assistance with large language models
J. Wu, R. Antonova, A. Kan, M. Lepert, A. Zeng, S. Song, J. Bohg, S. Rusinkiewicz, and T. Funkhouser · 2023
Closest in time.
A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W.-Y. Lo, et al · 2023
Closest in time.