Fetching the paper…
Reading the bibliography…
We describe a framework for using natural language to design state abstractions for imitation learning.
Alvinn: An autonomous land vehicle in a neural network
Dean A Pomerleau · 1988
Earlier work this paper cites.
Simplicity: a unifying principle in cognitive science?
Nick Chater and Paul Vitányi · 2003
Earlier work this paper cites.
An object-oriented representation for efficient reinforcement learning
Carlos Diuk, Andre Cohen, and Michael L Littman · 2008
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Conjugate markov decision processes
Philip S Thomas and Andrew G Barto · 2011
Earlier work this paper cites.
Designing robot learners that ask good questions
Maya Cakmak and Andrea L Thomaz · 2012
Earlier work this paper cites.
Learning feature representations with k-means
Adam Coates and Andrew Y Ng · 2012
Earlier work this paper cites.
Markov decision processes: discrete stochastic dynamic programming
Martin L Puterman · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
Near optimal behavior via approximate state abstraction
David Abel, David Hershkowitz, and Michael Littman · 2016
Earlier work this paper cites.
beta-vae: Learning basic visual concepts with a constrained variational framework
Irina Higgins, Loic Matthey, Arka Pal, Christopher Burgess, Xavier Glorot, Matthew Botvinick, Shakir Mohamed, and Alexander Lerchner · 2017
Earlier work this paper cites.
Dart: Noise injection for robust imitation learning
Michael Laskey, Jonathan Lee, Roy Fox, Anca Dragan, and Ken Goldberg · 2017
Earlier work this paper cites.
State abstractions for lifelong reinforcement learning
David Abel, Dilip Arumugam, Lucas Lehnert, and Michael Littman · 2018
Earlier work this paper cites.
Guiding policies with language via meta-learning
John D Co-Reyes, Abhishek Gupta, Suvansh Sanjeev, Nick Altieri, Jacob Andreas, John DeNero, Pieter Abbeel, and Sergey Levine · 2018
Earlier work this paper cites.
Mapping instructions to actions in 3D environments with visual goal prediction
Dipendra Misra, Andrew Bennett, Valts Blukis, Eyvind Niklasson, Max Shatkhin, and Yoav Artzi · 2018
Earlier work this paper cites.
Using natural language for reward shaping in reinforcement learning
Prasoon Goyal, Scott Niekum, and Raymond J Mooney · 2019
Cited alongside, same era.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych · 2019
Cited alongside, same era.
Pragmatic inference and visual abstraction enable contextual flexibility during visual communication
Judith E Fan, Robert D Hawkins, Mike Wu, and Noah D Goodman · 2020
Cited alongside, same era.
Feature expansive reward learning: Rethinking human input
Andreea Bobu, Marius Wiggert, Claire Tomlin, and Anca D. Dragan · 2021
Cited alongside, same era.
Kimin Lee, Laura Smith, and Pieter Abbeel · 2021
Cited alongside, same era.
SIRL: similarity-based implicit representation learning
Andreea Bobu, Yi Liu, Rohin Shah, Daniel S. Brown, and Anca D. Dragan · 2023
Later among the works it cites.
A survey of demonstration learning
André Correia and Luís A Alexandre · 2023
Later among the works it cites.
Guiding pretraining in reinforcement learning with large language models
Yuqing Du, Olivia Watkins, Zihan Wang, Cédric Colas, Trevor Darrell, Pieter Abbeel, Abhishek Gupta, and Jacob Andreas · 2023
Later among the works it cites.
Rational simplification and rigidity in human planning
Mark K Ho, Jonathan D Cohen, and Thomas L Griffiths · 2023
Later among the works it cites.
Visual explanations prioritize functional properties at the expense of visual fidelity
Holly Huey, Xuanchen Lu, Caren M Walker, and Judith E Fan · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mapping language models to grounded conceptual spaces
Roma Patel and Ellie Pavlick · 2021
Cited alongside, same era.
Learning rewards from linguistic feedback
Theodore R Sumers, Mark K Ho, Robert D Hawkins, Karthik Narasimhan, and Thomas L Griffiths · 2021
Cited alongside, same era.
Do as I can, not as I say: Grounding language in robotic affordances
Michael Ahn, Anthony Brohan, Noah Brown, Yevgen Chebotar, Omar Cortes, Byron David, Chelsea Finn, Keerthana Gopalakrishnan, Karol Hausman, Alex Herzog, et al · 2022
Cited alongside, same era.
Eager: Asking and answering questions for automatic reward shaping in language-guided rl
Thomas Carta, Pierre-Yves Oudeyer, Olivier Sigaud, and Sylvain Lamprier · 2022
Cited alongside, same era.
People construct simplified mental representations to plan
Mark K Ho, David Abel, Carlos G Correa, Michael L Littman, Jonathan D Cohen, and Thomas L Griffiths · 2022
Cited alongside, same era.
Vima: General robot manipulation with multimodal prompts
Yunfan Jiang, Agrim Gupta, Zichen Zhang, Guanzhi Wang, Yongqiang Dou, Yanjun Chen, Li Fei-Fei, Anima Anandkumar, Yuke Zhu, and Linxi Fan · 2022
Cited alongside, same era.
Improving intrinsic exploration with language abstractions
Jesse Mu, Victor Zhong, Roberta Raileanu, Minqi Jiang, Noah Goodman, Tim Rocktäschel, and Edward Grefenstette · 2022
Cited alongside, same era.
Siddharth Karamcheti, Suraj Nair, Annie S Chen, Thomas Kollar, Chelsea Finn, Dorsa Sadigh, and Percy Liang · 2023
Later among the works it cites.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Reward design with language models
Minae Kwon, Sang Michael Xie, Kalesha Bullard, and Dorsa Sadigh · 2023
Later among the works it cites.
Lampp: Language models as probabilistic priors for perception and action
Belinda Z Li, William Chen, Pratyusha Sharma, and Jacob Andreas · 2023
Later among the works it cites.
GPT-4 technical report, 2023
OpenAI · 2023
Later among the works it cites.
Diagnosis, feedback, adaptation: A human-in-the-loop framework for test-time policy adaptation
Andi Peng, Aviv Netanyahu, Mark K Ho, Tianmin Shu, Andreea Bobu, Julie Shah, and Pulkit Agrawal · 2023
Later among the works it cites.
Reflexion: an autonomous agent with dynamic memory and self-reflection
Noah Shinn, Beck Labash, and Ashwin Gopinath · 2023
Later among the works it cites.
Voyager: An open-ended embodied agent with large language models
Guanzhi Wang, Yuqi Xie, Yunfan Jiang, Ajay Mandlekar, Chaowei Xiao, Yuke Zhu, Linxi Fan, and Anima Anandkumar · 2023
Later among the works it cites.
Socratic models: Composing zero-shot multimodal reasoning with language
Andy Zeng, Maria Attarian, brian ichter, Krzysztof Marcin Choromanski, Adrian Wong, Stefan Welker, Federico Tombari, Aveek Purohit, Michael S Ryoo, Vikas Sindhwani, Johnny Lee, Vincent Vanhoucke, and Pete Florence · 2023
Later among the works it cites.
Aligning robot and human representations
Andreea Bobu, Andi Peng, Pulkit Agrawal, Julie Shah, and Anca D Dragan · 2024
Closest in time.