Fetching the paper…
Reading the bibliography…
Large language models can encode a wealth of semantic knowledge about the world.
Strips: A new approach to the application of theorem proving to problem solving
R. E. Fikes and N. J. Nilsson · 1971
Earlier work this paper cites.
Understanding natural language
T. Winograd · 1972
Earlier work this paper cites.
A structure for plans and behavior
E. D. Sacerdoti · 1975
Earlier work this paper cites.
The theory of affordances
J. J. Gibson · 1977
Earlier work this paper cites.
Grounding language in perception
J. M. Siskind · 1994
Earlier work this paper cites.
Shop: Simple hierarchical ordered planner
D. Nau, Y. Cao, A. Lotem, and H. Munoz-Avila · 1999
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, and action in route instructions
M. MacMahon, B. Stankiewicz, and B. Kuipers · 2006
Earlier work this paper cites.
Planning algorithms
S. M. LaValle · 2006
Earlier work this paper cites.
Toward understanding natural language directions
T. Kollar, S. Tellex, D. Roy, and N. Roy · 2010
Earlier work this paper cites.
Hierarchical planning in the now
L. P. Kaelbling and T. Lozano-Pérez · 2010
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
S. Tellex, T. Kollar, S. Dickerson, M. Walter, A. Banerjee, S. Teller, and N. Roy · 2011
Earlier work this paper cites.
A joint model of language and perception for grounded attribute learning
C. Matuszek, N. FitzGerald, L. Zettlemoyer, L. Bo, and D. Fox · 2012
Earlier work this paper cites.
Combined task and motion planning through an extensible planner-independent interface layer
S. Srivastava, E. Fang, L. Riano, R. Chitnis, S. Russell, and P. Abbeel · 2014
Earlier work this paper cites.
Asking for help using inverse semantics
S. Tellex, R. Knepper, A. Li, D. Rus, and N. Roy · 2014
Earlier work this paper cites.
Logic-geometric programming: An optimization-based approach to combined task and motion planning
M. Toussaint · 2015
Earlier work this paper cites.
Listen, attend, and walk: Neural mapping of navigational instructions to action sequences
H. Mei, M. Bansal, and M. R. Walter · 2016
Earlier work this paper cites.
Tell me dave: Context-sensitive grounding of natural language to manipulation instructions
D. K. Misra, J. Sung, K. Lee, and A. Saxena · 2016
Earlier work this paper cites.
Learning language games through interaction
S. I. Wang, P. Liang, and C. D. Manning · 2016
Earlier work this paper cites.
Prioritized experience replay
T. Schaul, J. Quan, I. Antonoglou, and D. Silver · 2016
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Mapping instructions and visual observations to actions with reinforcement learning
D. K. Misra, J. Langford, and Y. Artzi · 2017
Earlier work this paper cites.
Grounded language learning in a simulated 3d world
K. Hermann, F. Hill, S. Green, F. Wang, R. Faulkner, H. Soyer, D. Szepesvari, W. Czarnecki, M. Jaderberg, D. Teplyashin, M. Wainwright, C. Apps, D. Hassabis, and P. Blunsom · 2017
Earlier work this paper cites.
Zero-shot task generalization with multi-task deep reinforcement learning
J. Oh, S. Singh, H. Lee, and P. Kohli · 2017
Earlier work this paper cites.
Modular multitask reinforcement learning with policy sketches
J. Andreas, D. Klein, and S. Levine · 2017
Earlier work this paper cites.
Visual semantic planning using deep successor representations
Y. Zhu, D. Gordon, E. Kolve, D. Fox, L. Fei-Fei, A. Gupta, R. Mottaghi, and A. Farhadi · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Earlier work this paper cites.
D. Cer, Y. Yang, S.-y. Kong, N. Hua, N. Limtiaco, R. S. John, N. Constant, M. Guajardo-Cespedes, S. Yuan, C. Tar, et al · 2018
Earlier work this paper cites.
Differentiable physics and stable modes for tool-use and manipulation planning
M. A. Toussaint, K. R. Allen, K. A. Smith, and J. B. Tenenbaum · 2018
Earlier work this paper cites.
Neural task programming: Learning to generalize across hierarchical tasks
D. Xu, S. Nair, Y. Zhu, J. Gao, A. Garg, L. Fei-Fei, and S. Savarese · 2018
Earlier work this paper cites.
Semi-parametric topological memory for navigation
N. Savinov, A. Dosovitskiy, and V. Koltun · 2018
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 2019
Cited alongside, same era.
Videobert: A joint model for video and language representation learning
C. Sun, A. Myers, C. Vondrick, K. Murphy, and C. Schmid · 2019
Cited alongside, same era.
Visualbert: A simple and performant baseline for vision and language
L. H. Li, M. Yatskar, D. Yin, C.-J. Hsieh, and K.-W. Chang · 2019
Cited alongside, same era.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
J. Lu, D. Batra, D. Parikh, and S. Lee · 2019
Cited alongside, same era.
A survey of reinforcement learning informed by natural language
J. Luketina, N. Nardelli, G. Farquhar, J. N. Foerster, J. Andreas, E. Grefenstette, S. Whiteson, and T. Rocktäschel · 2019
Cited alongside, same era.
Learning language-conditioned robot behavior from offline data and crowd-sourced annotation
S. Nair, E. Mitchell, K. Chen, B. Ichter, S. Savarese, and C. Finn · 2021
Later among the works it cites.
Open-vocabulary object detection via vision and language knowledge distillation
X. Gu, T.-Y. Lin, W. Kuo, and Y. Cui · 2021
Later among the works it cites.
Merlot: Multimodal neural script knowledge models
R. Zellers, X. Lu, J. Hessel, Y. Yu, J. S. Park, J. Cao, A. Farhadi, and Y. Choi · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Later among the works it cites.
Embodied bert: A transformer model for embodied, language-guided visual task completion
A. Suglia, Q. Gao, J. Thomason, G. Thattai, and G. Sukhatme · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Language as an abstraction for hierarchical deep reinforcement learning
Y. Jiang, S. Gu, K. Murphy, and C. Finn · 2019
Cited alongside, same era.
Self-educated language agent with hindsight experience replay for instruction following
G. Cideron, M. Seurin, F. Strub, and O. Pietquin · 2019
Cited alongside, same era.
Regression planning networks
D. Xu, R. Martín-Martín, D.-A. Huang, Y. Zhu, S. Savarese, and L. F. Fei-Fei · 2019
Cited alongside, same era.
Neural task graphs: Generalizing to unseen tasks from a single video demonstration
D.-A. Huang, S. Nair, D. Xu, Y. Zhu, A. Garg, L. Fei-Fei, S. Savarese, and J. C. Niebles · 2019
Cited alongside, same era.
Search on the replay buffer: Bridging planning and reinforcement learning
B. Eysenbach, R. R. Salakhutdinov, and S. Levine · 2019
Cited alongside, same era.
Thinking while moving: Deep reinforcement learning with concurrent control
T. Xiao, E. Jang, D. Kalashnikov, S. Levine, J. Ibarz, K. Hausman, and A. Herzog · 2019
Cited alongside, same era.
Climbing towards nlu: On meaning, form, and understanding in the age of data
E. M. Bender and A. Koller · 2020
Cited alongside, same era.
Later among the works it cites.
Episodic transformer for vision-and-language navigation
A. Pashevich, C. Schmid, and C. Sun · 2021
Later among the works it cites.
Skill induction and planning with latent language
P. Sharma, A. Torralba, and J. Andreas · 2021
Later among the works it cites.
Piglet: Language grounding through neuro-symbolic interaction in a 3d world
R. Zellers, A. Holtzman, M. Peters, R. Mottaghi, A. Kembhavi, A. Farhadi, and Y. Choi · 2021
Later among the works it cites.
Broadly-exploring, local-policy trees for long-horizon task planning
B. Ichter, P. Sermanet, and C. Lynch · 2021
Later among the works it cites.
Example-driven model-based reinforcement learning for solving long-horizon visuomotor tasks
B. Wu, S. Nair, L. Fei-Fei, and C. Finn · 2021
Later among the works it cites.
Relmogen: Integrating motion generation in reinforcement learning for mobile manipulation
F. Xia, C. Li, R. Martín-Martín, O. Litany, A. Toshev, and S. Savarese · 2021
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell · 2021
Later among the works it cites.
On the opportunities and risks of foundation models
R. Bommasani, D. A. Hudson, E. Adeli, R. Altman, S. Arora, S. von Arx, M. S. Bernstein, J. Bohg, A. Bosselut, E. Brunskill, et al · 2021
Later among the works it cites.
Actionable models: Unsupervised offline reinforcement learning of robotic skills
Y. Chebotar, K. Hausman, Y. Lu, T. Xiao, D. Kalashnikov, J. Varley, A. Irpan, B. Eysenbach, R. C. Julian, C. Finn, and S. Levine · 2021
Later among the works it cites.
Lamda: Language models for dialog applications
R. Thoppilan, D. De Freitas, J. Hall, N. Shazeer, A. Kulshreshtha, H.-T. Cheng, A. Jin, T. Bos, L. Baker, Y. Du, et al · 2022
Closest in time.
Palm: Scaling language modeling with pathways
A. Chowdhery, S. Narang, J. Devlin, et al · 2022
Closest in time.
Training language models to follow instructions with human feedback
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. L. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, et al · 2022
Closest in time.
Value function spaces: Skill-centric state abstractions for long-horizon reasoning
D. Shah, P. Xu, Y. Lu, T. Xiao, A. Toshev, S. Levine, and B. Ichter · 2022
Closest in time.
Behavior: Benchmark for everyday household activities in virtual, interactive, and ecological environments
S. Srivastava, C. Li, M. Lingelbach, R. Martín-Martín, F. Xia, K. E. Vainio, Z. Lian, C. Gokmen, S. Buch, K. Liu, et al · 2022
Closest in time.
Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
W. Huang, P. Abbeel, D. Pathak, and I. Mordatch · 2022
Closest in time.
Chain of thought prompting elicits reasoning in large language models
J. Wei, X. Wang, D. Schuurmans, M. Bosma, E. Chi, Q. Le, and D. Zhou · 2022
Closest in time.
Inner monologue: Embodied reasoning through planning with language models
W. Huang, F. Xia, T. Xiao, H. Chan, J. Liang, P. Florence, A. Zeng, J. Tompson, I. Mordatch, Y. Chebotar, et al · 2022
Closest in time.
Cliport: What and where pathways for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2022
Closest in time.
A persistent spatial semantic representation for high-level natural language instruction execution
V. Blukis, C. Paxton, D. Fox, A. Garg, and Y. Artzi · 2022
Closest in time.
R3m: A universal visual representation for robot manipulation
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta · 2022
Closest in time.
A data-driven approach for learning to control computers
P. C. Humphreys, D. Raposo, T. Pohlen, G. Thornton, R. Chhaparia, A. Muldal, J. Abramson, P. Georgiev, A. Goldin, A. Santoro, et al · 2022
Closest in time.
Can wikipedia help offline reinforcement learning
M. Reid, Y. Yamada, and S. S. Gu · 2022
Closest in time.
Pre-trained language models for interactive decision-making
S. Li, X. Puig, Y. Du, C. Wang, E. Akyurek, A. Torralba, J. Andreas, and I. Mordatch · 2022
Closest in time.
Inventing relational state and action abstractions for effective and efficient bilevel planning
T. Silver, R. Chitnis, N. Kumar, W. McClinton, T. Lozano-Perez, L. P. Kaelbling, and J. Tenenbaum · 2022
Closest in time.