Fetching the paper…
Reading the bibliography…
Language-conditioned policies allow robots to interpret and execute human instructions.
The potential field approach and operational space formulation in robot control
O. Khatib · 1986
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
M. McCloskey and N. J. Cohen · 1989
Earlier work this paper cites.
Programming by demonstration: A machine learning approach to support skill acquision for robots
R. Dillmann and H. Friedrich · 1996
Earlier work this paper cites.
Is imitation learning the route to humanoid robots?
S. Schaal · 1999
Earlier work this paper cites.
Neural networks are surprisingly modular, 2020
D. Filan, S. Hod, C. Wild, A. Critch, and S. Russell · 2003
Earlier work this paper cites.
A tutorial on the cross-entropy method
P.-T. de Boer, D. P. Kroese, S. Mannor, and R. Y. Rubinstein · 2004
Earlier work this paper cites.
Dynamic movement primitives-a framework for motor control in humans and humanoid robotics
S. Schaal · 2006
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2009
Earlier work this paper cites.
Apprenticeship learning for helicopter control
A. Coates, P. Abbeel, and A. Y. Ng · 2009
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
E. Todorov, T. Erez, and Y. Tassa · 2012
Earlier work this paper cites.
Learning interaction for collaborative tasks with probabilistic movement primitives
G. Maeda, M. Ewerton, R. Lioutikov, H. B. Amor, J. Peters, and G. Neumann · 2014
Earlier work this paper cites.
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2015
Earlier work this paper cites.
Show, attend and tell: Neural image caption generation with visual attention
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio · 2015
Earlier work this paper cites.
Vqa: Visual question answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Neural module networks
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Earlier work this paper cites.
Neural machine translation with supervised attention
L. Liu, M. Utiyama, A. Finch, and E. Sumita · 2016
Earlier work this paper cites.
One-shot imitation learning
Y. Duan, M. Andrychowicz, B. Stadie, O. Jonathan Ho, J. Schneider, I. Sutskever, P. Abbeel, and W. Zaremba · 2017
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
J. Johnson, B. Hariharan, L. Van Der Maaten, L. Fei-Fei, C. Lawrence Zitnick, and R. Girshick · 2017
Cited alongside, same era.
Visual dialog
A. Das, S. Kottur, K. Gupta, A. Singh, D. Yadav, J. M. Moura, D. Parikh, and D. Batra · 2017
Cited alongside, same era.
One-shot imitation learning
Y. Duan, M. Andrychowicz, B. Stadie, O. Jonathan Ho, J. Schneider, I. Sutskever, P. Abbeel, and W. Zaremba · 2017
Cited alongside, same era.
Exploiting argument information to improve event detection via supervised attention mechanisms
S. Liu, Y. Chen, K. Liu, and J. Zhao · 2017
Cited alongside, same era.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Cited alongside, same era.
Deep imitation learning for complex manipulation tasks from virtual reality teleoperation
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel · 2018
Language-conditioned imitation learning for robot manipulation tasks
S. Stepputtis, J. Campbell, M. Phielipp, S. Lee, C. Baral, and H. B. Amor · 2020
Later among the works it cites.
Deep imitation learning for bimanual robotic manipulation
F. Xie, A. Chowdhury, M. C. De Paolis Kaluza, L. Zhao, L. Wong, and R. Yu · 2020
Later among the works it cites.
Uniter: Universal image-text representation learning
Y.-C. Chen, L. Li, L. Yu, A. El Kholy, F. Ahmed, Z. Gan, Y. Cheng, and J. Liu · 2020
Later among the works it cites.
Deep compositional robotic planners that follow natural language commands
Y.-L. Kuo, B. Katz, and A. Barbu · 2020
Later among the works it cites.
End-to-end object detection with transformers
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko · 2020
Later among the works it cites.
Object-centric learning with slot attention
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Visual coreference resolution in visual dialog using neural module networks
S. Kottur, J. M. Moura, D. Parikh, D. Batra, and M. Rohrbach · 2018
Cited alongside, same era.
Vision-based multi-task manipulation for inexpensive robots using end-to-end learning from demonstration
R. Rahmatizadeh, P. Abolghasemi, L. Bölöni, and S. Levine · 2018
Cited alongside, same era.
Deep imitation learning for complex manipulation tasks from virtual reality teleoperation
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel · 2018
Cited alongside, same era.
Document embedding enhanced event detection with hierarchical and supervised attention
Y. Zhao, X. Jin, Y. Wang, and X. Cheng · 2018
Cited alongside, same era.
A lexicon-based supervised attention model for neural sentiment analysis
Y. Zou, T. Gui, Q. Zhang, and X.-J. Huang · 2018
Cited alongside, same era.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
J. Lu, D. Batra, D. Parikh, and S. Lee · 2019
Cited alongside, same era.
F. Locatello, D. Weissenborn, T. Unterthiner, A. Mahendran, G. Heigold, J. Uszkoreit, A. Dosovitskiy, and T. Kipf · 2020
Later among the works it cites.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, and R. Martín-Martín · 2020
Later among the works it cites.
Language conditioned imitation learning over unstructured data
C. Lynch and P. Sermanet · 2021
Later among the works it cites.
Cliport: What and where pathways for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2021
Later among the works it cites.
Mdetr - modulated detection for end-to-end multi-modal understanding
A. Kamath, M. Singh, Y. LeCun, G. Synnaeve, I. Misra, and N. Carion · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Later among the works it cites.
Are neural nets modular? inspecting functional modularity through differentiable weight masks
R. Csordás, S. van Steenkiste, and J. Schmidhuber · 2021
Later among the works it cites.
Compositional generalization for neural semantic parsing via span-level supervised attention
P. Yin, H. Fang, G. Neubig, A. Pauls, E. A. Platanios, Y. Su, S. Thomson, and J. Andreas · 2021
Later among the works it cites.
Bc-z: Zero-shot task generalization with robotic imitation learning
E. Jang, A. Irpan, M. Khansari, D. Kappler, F. Ebert, C. Lynch, S. Levine, and C. Finn · 2022
Closest in time.
Do as i can, not as i say: Grounding language in robotic affordances
M. Ahn, A. Brohan, N. Brown, Y. Chebotar, O. Cortes, B. David, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, et al · 2022
Closest in time.
Visual object tracking: A survey
F. Chen, X. Wang, Y. Zhao, S. Lv, and X. Niu · 2022
Closest in time.