Fetching the paper…
Reading the bibliography…
Learned language-conditioned robot policies often struggle to effectively adapt to new real-world tasks even when pre-trained across a diverse set of instructions.
Human few-shot learning of compositional instructions
B. M. Lake, T. Linzen, and M. Baroni · 1901
Earlier work this paper cites.
Multilingual Universal Sentence Encoder for Semantic Retrieval
Y. Yang, D. Cer, A. Ahmad, M. Guo, J. Law, N. Constant, G. H. Abrego, S. Yuan, C. Tar, Y.-H. Sung, B. Strope, and R. Kurzweil · 1907
Earlier work this paper cites.
User-friendly Introduction to PAC-Bayes Bounds
P. Alquier · 1935
Earlier work this paper cites.
Probability inequalities for sums of bounded random variables
W. Hoeffding · 1963
Earlier work this paper cites.
Procedures as a representation for data in a computer program for understanding natural language, 1971
T. Winograd · 1971
Earlier work this paper cites.
A theory of the learnable
L. G. Valiant · 1972
Earlier work this paper cites.
Spatial language for human-robot dialogs
M. Skubic, D. Perzanowski, S. Blisard, A. C. Schultz, W. Adams, M. D. Bugajska, and D. P. Brock · 2004
Earlier work this paper cites.
A PAC-Bayesian approach to adaptive classification, 2004
O. Catoni · 2004
Earlier work this paper cites.
Understanding Natural Language Commands for Robotic Navigation and Mobile Manipulation
S. Tellex, T. Kollar, S. Dickerson, M. Walter, A. Banerjee, S. Teller, and N. Roy · 2011
Earlier work this paper cites.
Information Theory: Coding Theorems for Discrete Memoryless Systems
I. Csiszár and J. Körner · 2011
Earlier work this paper cites.
Matching Networks for One Shot Learning
O. Vinyals, C. Blundell, T. Lillicrap, koray kavukcuoglu, and D. Wierstra · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks
C. Finn, P. Abbeel, and S. Levine · 2017
Earlier work this paper cites.
Prototypical Networks for Few-shot Learning
J. Snell, K. Swersky, and R. Zemel · 2017
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization, Jan. 2017
D. P. Kingma and J. Ba · 2017
Earlier work this paper cites.
On First-Order Meta-Learning Algorithms
A. Nichol, J. Achiam, and J. Schulman · 2018
Earlier work this paper cites.
FiLM: Visual Reasoning with a General Conditioning Layer
E. Perez, F. Strub, H. de Vries, V. Dumoulin, and A. Courville · 2018
Earlier work this paper cites.
A Closer Look at Few-shot Classification
W.-Y. Chen, Y.-C. Liu, Z. Kira, Y.-C. F. Wang, and J.-B. Huang · 2019
Earlier work this paper cites.
Language Models are Few-Shot Learners, 2020
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Earlier work this paper cites.
Language-conditioned imitation learning for robot manipulation tasks
S. Stepputtis, J. Campbell, M. Phielipp, S. Lee, C. Baral, and H. B. Amor · 2020
Earlier work this paper cites.
Language conditioned imitation learning over unstructured data
C. Lynch and P. Sermanet · 2021
Earlier work this paper cites.
Composing pick-and-place tasks by grounding language
O. Mees and W. Burgard · 2021
Earlier work this paper cites.
Bayesian Meta-Learning for Few-Shot Policy Adaptation Across Robotic Platforms
A. Ghadirzadeh, X. Chen, P. Poklukar, C. Finn, M. Björkman, and D. Kragic · 2021
Earlier work this paper cites.
Non-Gaussian Gaussian Processes for Few-Shot Regression
M. Sendera, J. Tabor, A. Nowak, A. Bedychaj, M. Patacchiola, T. Trzcinski, P. aw Spurek, and M. Zieba · 2021
Earlier work this paper cites.
Learning to Learn Dense Gaussian Processes for Few-Shot Learning
Z. Wang, Z. Miao, X. Zhen, and Q. Qiu · 2021
Cited alongside, same era.
CLIPort: What and Where Pathways for Robotic Manipulation, Sept. 2021
M. Shridhar, L. Manuelli, and D. Fox · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Cited alongside, same era.
What matters in language conditioned robotic imitation learning over unstructured data
O. Mees, L. Hermann, and W. Burgard · 2022
Cited alongside, same era.
Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language, May 2022
A. Zeng, M. Attarian, B. Ichter, K. Choromanski, A. Wong, S. Welker, F. Tombari, A. Purohit, M. Ryoo, V. Sindhwani, et al · 2022
Cited alongside, same era.
Flamingo: A Visual Language Model for Few-Shot Learning
Goal Representations for Instruction Following: A Semi-Supervised Language Interface to Control
V. Myers, A. W. He, K. Fang, H. R. Walke, P. Hansen-Estruch, C.-A. Cheng, M. Jalobeanu, A. Kolobov, A. Dragan, and S. Levine · 2023
Later among the works it cites.
Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models, Oct. 2023
K. Black, M. Nakamoto, P. Atreya, H. Walke, C. Finn, A. Kumar, and S. Levine · 2023
Later among the works it cites.
VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models, Nov. 2023
W. Huang, C. Wang, R. Zhang, Y. Li, J. Wu, and L. Fei-Fei · 2023
Later among the works it cites.
Grounding language with visual affordances over unstructured data
O. Mees, J. Borja-Diaz, and W. Burgard · 2023
Later among the works it cites.
GPT-4 Technical Report, Mar. 2023
OpenAI · 2023
Later among the works it cites.
VIMA: General Robot Manipulation with Multimodal Prompts, May 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J.-B. Alayrac, J. Donahue, P. Luc, A. Miech, I. Barr, Y. Hasson, K. Lenc, A. Mensch, K. Millican, M. Reynolds, R. Ring, E. Rutherford, S. Cabi, T. Han, Z. Gong, et al · 2022
Cited alongside, same era.
Robust Meta-learning with Sampling Noise and Label Noise via Eigen-Reptile
D. Chen, L. Wu, S. Tang, X. Yun, B. Long, and Y. Zhuang · 2022
Cited alongside, same era.
An Explanation of In-context Learning as Implicit Bayesian Inference, July 2022
S. M. Xie, A. Raghunathan, P. Liang, and T. Ma · 2022
Cited alongside, same era.
Interactive Language: Talking to Robots in Real Time, Oct. 2022
C. Lynch, A. Wahid, J. Tompson, T. Ding, J. Betker, R. Baruch, T. Armstrong, and P. Florence · 2022
Cited alongside, same era.
CALVIN: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks
O. Mees, L. Hermann, E. Rosete-Beas, and W. Burgard · 2022
Cited alongside, same era.
Do As I Can, Not As I Say: Grounding Language in Robotic Affordances, Aug. 2022
M. Ahn, A. Brohan, N. Brown, Y. Chebotar, O. Cortes, B. David, C. Finn, C. Fu, K. Gopalakrishnan, K. Hausman, et al · 2022
Cited alongside, same era.
See, Plan, Predict: Language-guided Cognitive Planning with Video Prediction, Oct. 2022
M. Attarian, A. Gupta, Z. Zhou, W. Yu, I. Gilitschenski, and A. Garg · 2022
Cited alongside, same era.
Y. Jiang, A. Gupta, Z. Zhang, G. Wang, Y. Dou, Y. Chen, L. Fei-Fei, A. Anandkumar, Y. Zhu, and L. Fan · 2023
Later among the works it cites.
MUTEX: Learning Unified Policies from Multimodal Task Specifications
R. Shah, R. Martín-Martín, and Y. Zhu · 2023
Later among the works it cites.
Scaling Robot Learning with Semantically Imagined Experience
T. Yu, T. Xiao, J. Tompson, A. Stone, S. Wang, A. Brohan, J. Singh, C. Tan, D. M, J. Peralta, et al · 2023
Later among the works it cites.
GenAug: Retargeting behaviors to unseen situations via Generative Augmentation
Q. Chen, S. Kiami, A. Gupta, and V. Kumar · 2023
Later among the works it cites.
Audio visual language maps for robot navigation
C. Huang, O. Mees, A. Zeng, and W. Burgard · 2023
Later among the works it cites.
SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling, June 2023
J. Zhang, K. Pertsch, J. Zhang, and J. J. Lim · 2023
Later among the works it cites.
Visual language maps for robot navigation
C. Huang, O. Mees, A. Zeng, and W. Burgard · 2023
Later among the works it cites.
Learning to model the world with language
J. Lin, Y. Du, O. Watkins, D. Hafner, P. Abbeel, D. Klein, and A. Dragan · 2024
Closest in time.
Octo: An open-source generalist robot policy
Octo Model Team, D. Ghosh, H. Walke, K. Pertsch, K. Black, O. Mees, S. Dasari, J. Hejna, C. Xu, J. Luo, T. Kreiman, Y. Tan, L. Y. Chen, P. Sanketi, Q. Vuong, T. Xiao, D. Sadigh, C. Finn, and S. Levine · 2024
Closest in time.
Scalable Meta-Learning with Gaussian Processes
P. Tighineanu, L. Grossberger, P. Baireuther, K. Skubch, S. Falkner, J. Vinogradska, and F. Berkenkamp · 2024
Closest in time.
InCoRo: In-Context Learning for Robotics Control with Feedback Loops, Feb. 2024
J. Y. Zhu, C. G. Cano, D. V. Bermudez, and M. Drozdzal · 2024
Closest in time.
MOKA: Open-Vocabulary Robotic Manipulation through Mark-Based Visual Prompting
F. Liu, K. Fang, P. Abbeel, and S. Levine · 2024
Closest in time.
RT-H: Action Hierarchies Using Language, Mar. 2024
S. Belkhale, T. Ding, T. Xiao, P. Sermanet, Q. Vuong, J. Tompson, Y. Chebotar, D. Dwibedi, and D. Sadigh · 2024
Closest in time.
Yell At Your Robot: Improving On-the-Fly from Language Corrections
L. X. Shi, Z. Hu, T. Z. Zhao, A. Sharma, K. Pertsch, J. Luo, S. Levine, and C. Finn · 2024
Closest in time.
Open X-Embodiment: Robotic Learning Datasets and RT-X Models, May 2024
O. X.-E. Collaboration, A. O’Neill, A. Rehman, A. Maddukuri, A. Gupta, A. Padalkar, A. Lee, A. Pooley, A. Gupta, A. Mandlekar, et al · 2024
Closest in time.
Scaling cross-embodied learning: One policy for manipulation, navigation, locomotion and aviation
R. Doshi, H. Walke, O. Mees, S. Dasari, and S. Levine · 2024
Closest in time.
Vision-language models provide promptable representations for reinforcement learning
W. Chen, O. Mees, A. Kumar, and S. Levine · 2024
Closest in time.
Toward Grounded Commonsense Reasoning
M. Kwon, H. Hu, V. Myers, S. Karamcheti, A. Dragan, and D. Sadigh · 2024
Closest in time.