Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated impressive planning abilities due to their vast "world knowledge".
Fine-Tuning Language Models from Human Preferences
Ziegler, D. M.; Stiennon, N.; Wu, J.; Brown, T. B.; Radford, A.; Amodei, D.; Christiano, P.; and Irving, G. 2020 · 1909
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
Kaelbling, L. P.; Littman, M. L.; and Cassandra, A. R. 1998 · 1998
Earlier work this paper cites.
Planning as heuristic search
Bonet, B.; and Geffner, H. 2001 · 2001
Earlier work this paper cites.
The fast downward planning system
Helmert, M. 2006 · 2006
Earlier work this paper cites.
Virtualhome: Simulating household activities via programs
Puig, X.; Ra, K.; Boben, M.; Li, J.; Wang, T.; Fidler, S.; and Torralba, A. 2018 · 2018
Earlier work this paper cites.
BabyAI: First Steps Towards Grounded Language Learning With a Human In the Loop
Chevalier-Boisvert, M.; Bahdanau, D.; Lahlou, S.; Willems, L.; Saharia, C.; Nguyen, T. H.; and Bengio, Y. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Earlier work this paper cites.
Synthesizing Environment-Aware Activities via Activity Sketches
Liao, Y.-H.; Puig, X.; Boben, M.; Torralba, A.; and Fidler, S. 2019 · 2019
Earlier work this paper cites.
Representation Learning with Contrastive Predictive Coding
van den Oord, A.; Li, Y.; and Vinyals, O. 2019 · 2019
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C.; Shazeer, N.; Roberts, A.; Lee, K.; Narang, S.; Matena, M.; Zhou, Y.; Li, W.; and Liu, P. J. 2020 · 2020
Earlier work this paper cites.
Transporter Networks: Rearranging the Visual World for Robotic Manipulation
Zeng, A.; Florence, P.; Tompson, J.; Welker, S.; Chien, J.; Attarian, M.; Armstrong, T.; Krasin, I.; Duong, D.; Sindhwani, V.; and Lee, J. 2021 · 2020
Earlier work this paper cites.
On Generative Spoken Language Modeling from Raw Audio
Lakhotia, K.; Kharitonov, E.; Hsu, W.-N.; Adi, Y.; Polyak, A.; Bolte, B.; Nguyen, T.-A.; Copet, J.; Baevski, A.; Mohamed, A.; and Dupoux, E. 2021 · 2021
Earlier work this paper cites.
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation
Wang, Y.; Wang, W.; Joty, S.; and Hoi, S. C. 2021 · 2021
Earlier work this paper cites.
Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Ahn, M.; Brohan, A.; Brown, N.; Chebotar, Y.; Cortes, O.; David, B.; Finn, C.; Fu, C.; Gopalakrishnan, K.; Hausman, K.; Herzog, A.; Ho, D.; Hsu, J.; Ibarz, J.; Ichter, B.; Irpan, A.; Jang, E.; Ruano, R. J.; Jeffrey, K.; Jesmonth, S.; Joshi, N. J.; Julian, R.; Kalashnikov, D.; Kuang, Y.; Lee, K.-H.; Levine, S.; Lu, Y.; Luu, L.; Parada, C.; Pastor, P.; Quiambao, J.; Rao, K.; Rettinghouse, J.; Reyes, D.; Sermanet, P.; Sievers, N.; Tan, C.; Toshev, A.; Vanhoucke, V.; Xia, F.; Xiao, T.; Xu, P.; Xu, S.; Yan, M.; and Zeng, A. 2022 · 2022
Cited alongside, same era.
Scaling Instruction-Finetuned Language Models
Chung, H. W.; Hou, L.; Longpre, S.; Zoph, B.; Tay, Y.; Fedus, W.; Li, Y.; Wang, X.; Dehghani, M.; Brahma, S.; Webson, A.; Gu, S. S.; Dai, Z.; Suzgun, M.; Chen, X.; Chowdhery, A.; Castro-Ros, A.; Pellat, M.; Robinson, K.; Valter, D.; Narang, S.; Mishra, G.; Yu, A.; Zhao, V.; Huang, Y.; Dai, A.; Yu, H.; Petrov, S.; Chi, E. H.; Dean, J.; Devlin, J.; Roberts, A.; Zhou, D.; Le, Q. V.; and Wei, J. 2022 · 2022
Cited alongside, same era.
LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale
Dettmers, T.; Lewis, M.; Belkada, Y.; and Zettlemoyer, L. 2022 · 2022
Cited alongside, same era.
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Frantar, E.; Ashkboos, S.; Hoefler, T.; and Alistarh, D. 2023 · 2023
Closest in time.
Reasoning with Language Model is Planning with World Model
Hao, S.; Gu, Y.; Ma, H.; Hong, J. J.; Wang, Z.; Wang, D. Z.; and Hu, Z. 2023 · 2023
Closest in time.
Grounded Decoding: Guiding Text Generation with Grounded Models for Embodied Agents
Huang, W.; Xia, F.; Shah, D.; Driess, D.; Zeng, A.; Lu, Y.; Florence, P.; Mordatch, I.; Levine, S.; Hausman, K.; and Ichter, B. 2023 · 2023
Closest in time.
Reward Design with Language Models
Kwon, M.; Xie, S. M.; Bullard, K.; and Sadigh, D. 2023 · 2023
Closest in time.
Code as Policies: Language Model Programs for Embodied Control
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Du, Y.; Liu, Z.; Li, J.; and Zhao, W. X. 2022 · 2022
Cited alongside, same era.
Planning in Observable POMDPs in Quasipolynomial Time
Golowich, N.; Moitra, A.; and Rohatgi, D. 2022 · 2022
Cited alongside, same era.
Plansformer: Generating Symbolic Plans using Transformers
Pallagani, V.; Muppasani, B.; Murugesan, K.; Rossi, F.; Horesh, L.; Srivastava, B.; Fabiano, F.; and Loreggia, A. 2022 · 2022
Cited alongside, same era.
PDDL Planning with Pretrained Large Language Models
Silver, T.; Hariprasad, V.; Shuttleworth, R. S.; Kumar, N.; Lozano-Pérez, T.; and Kaelbling, L. P. 2022 · 2022
Cited alongside, same era.
Large Language Models Still Can’t Plan (A Benchmark for LLMs on Planning and Reasoning about Change)
Valmeekam, K.; Olmo, A.; Sreedharan, S.; and Kambhampati, S. 2022 · 2022
Cited alongside, same era.
RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Brohan, A.; Brown, N.; Carbajal, J.; Chebotar, Y.; Chen, X.; Choromanski, K.; Ding, T.; Driess, D.; Dubey, A.; Finn, C.; Florence, P.; Fu, C.; Arenas, M. G.; Gopalakrishnan, K.; Han, K.; Hausman, K.; Herzog, A.; Hsu, J.; Ichter, B.; Irpan, A.; Joshi, N.; Julian, R.; Kalashnikov, D.; Kuang, Y.; Leal, I.; Lee, L.; Lee, T.-W. E.; Levine, S.; Lu, Y.; Michalewski, H.; Mordatch, I.; Pertsch, K.; Rao, K.; Reymann, K.; Ryoo, M.; Salazar, G.; Sanketi, P.; Sermanet, P.; Singh, J.; Singh, A.; Soricut, R.; Tran, H.; Vanhoucke, V.; Vuong, Q.; Wahid, A.; Welker, S.; Wohlhart, P.; Wu, J.; Xia, F.; Xiao, T.; Xu, P.; Xu, S.; Yu, T.; and Zitkovich, B. 2023 · 2023
Cited alongside, same era.
Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality
Chiang, W.-L.; Li, Z.; Lin, Z.; Sheng, Y.; Wu, Z.; Zhang, H.; Zheng, L.; Zhuang, S.; Zhuang, Y.; Gonzalez, J. E.; Stoica, I.; and Xing, E. P. 2023 · 2023
Cited alongside, same era.
QLoRA: Efficient Finetuning of Quantized LLMs
Dettmers, T.; Pagnoni, A.; Holtzman, A.; and Zettlemoyer, L. 2023 · 2023
Cited alongside, same era.
Integrating action knowledge and LLMs for task planning and situation handling in open worlds
Ding, Y.; Zhang, X.; Amiri, S.; Cao, N.; Yang, H.; Kaminski, A.; Esselink, C.; and Zhang, S. 2023 · 2023
Cited alongside, same era.
Liang, J.; Huang, W.; Xia, F.; Xu, P.; Hausman, K.; Ichter, B.; Florence, P.; and Zeng, A. 2023 · 2023
Closest in time.
Text2Motion: from natural language instructions to feasible plans
Lin, K.; Agia, C.; Migimatsu, T.; Pavone, M.; and Bohg, J. 2023 · 2023
Closest in time.
LLM+P: Empowering Large Language Models with Optimal Planning Proficiency
Liu, B.; Jiang, Y.; Zhang, X.; Liu, Q.; Zhang, S.; Biswas, J.; and Stone, P. 2023 · 2023
Closest in time.
ProgPrompt: Generating Situated Robot Task Plans using Large Language Models
Singh, I.; Blukis, V.; Mousavian, A.; Goyal, A.; Xu, D.; Tremblay, J.; Fox, D.; Thomason, J.; and Garg, A. 2023 · 2023
Closest in time.
LLaMA: Open and Efficient Foundation Language Models
Touvron, H.; Lavril, T.; Izacard, G.; Martinet, X.; Lachaux, M.-A.; Lacroix, T.; Rozière, B.; Goyal, N.; Hambro, E.; Azhar, F.; Rodriguez, A.; Joulin, A.; Grave, E.; and Lample, G. 2023 · 2023
Closest in time.
Valmeekam, K.; Sreedharan, S.; Marquez, M.; Olmo, A.; and Kambhampati, S. 2023 · 2023
Closest in time.
Translating Natural Language to Planning Goals with Large-Language Models
Xie, Y.; Yu, C.; Zhu, T.; Bai, J.; Gong, Z.; and Soh, H. 2023 · 2023
Closest in time.
Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Yao, S.; Yu, D.; Zhao, J.; Shafran, I.; Griffiths, T. L.; Cao, Y.; and Narasimhan, K. 2023 · 2023
Closest in time.