Fetching the paper…
Reading the bibliography…
The recent surge in research interest in applying large language models (LLMs) to decision-making tasks has flourished by leveraging the extensive world knowledge embedded in LLMs.
Self-Improving Reactive Agents Based on Reinforcement Learning, Planning and Teaching
Lin, L.-J. 1992 · 1992
Earlier work this paper cites.
Q-learning
Watkins, C. J.; and Dayan, P. 1992 · 1992
Earlier work this paper cites.
Thinking, Fast and Slow
Kahneman, D. 2011 · 2011
Earlier work this paper cites.
Prioritized Experience Replay
Schaul, T.; Quan, J.; Antonoglou, I.; and Silver, D. 2015 · 2015
Earlier work this paper cites.
Reinforcement Learning: An Introduction
Sutton, R. S.; and Barto, A. G. 2018 · 2018
Earlier work this paper cites.
FEVER: a Large-scale Dataset for Fact Extraction and VERification
Thorne, J.; Vlachos, A.; Christodoulopoulos, C.; and Mittal, A. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Yang, Z.; Qi, P.; Zhang, S.; Bengio, Y.; Cohen, W.; Salakhutdinov, R.; and Manning, C. D. 2018 · 2018
Earlier work this paper cites.
Billion-scale Similarity Search with GPUs
Johnson, J.; Douze, M.; and Jégou, H. 2019 · 2019
Earlier work this paper cites.
Language Models as Knowledge Bases?
Petroni, F.; Rocktäschel, T.; Riedel, S.; Lewis, P.; Bakhtin, A.; Wu, Y.; and Miller, A. 2019 · 2019
Earlier work this paper cites.
Language Models are Few-Shot Learners
Brown, T.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J. D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 2020
Earlier work this paper cites.
MPNet: Masked and Permuted Pre-training for Language Understanding
Song, K.; Tan, X.; Qin, T.; Lu, J.; and Liu, T.-Y. 2020 · 2020
Earlier work this paper cites.
A Comprehensive Survey on Transfer Learning
Zhuang, F.; Qi, Z.; Duan, K.; Xi, D.; Zhu, Y.; Zhu, H.; Xiong, H.; and He, Q. 2020 · 2020
Earlier work this paper cites.
WebGPT: Browser-Assisted Question-Answering with Human Feedback
Nakano, R.; Hilton, J.; Balaji, S. A.; Wu, J.; Ouyang, L.; Kim, C.; Hesse, C.; Jain, S.; Kosaraju, V.; Saunders, W.; Jiang, X.; Cobbe, K.; Eloundou, T.; Krueger, G.; Button, K.; Knight, M.; Chess, B.; and Schulman, J. 2021 · 2021
Earlier work this paper cites.
ALFWorld: Aligning Text and Embodied Environments for Interactive Learning
Shridhar, M.; Yuan, X.; Côté, M.-A.; Bisk, Y.; Trischler, A.; and Hausknecht, M. 2021 · 2021
Earlier work this paper cites.
Scaling Instruction-Finetuned Language Models
Chung, H. W.; Hou, L.; Longpre, S.; Zoph, B.; Tay, Y.; Fedus, W.; Li, E.; Wang, X.; Dehghani, M.; Brahma, S.; et al. 2022 · 2022
Earlier work this paper cites.
Shortcut Learning of Large Language Models in Natural Language Understanding: A Survey
Du, M.; He, F.; Zou, N.; Tao, D.; and Hu, X. 2022 · 2022
Earlier work this paper cites.
Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents
Huang, W.; Abbeel, P.; Pathak, D.; and Mordatch, I. 2022 · 2022
Earlier work this paper cites.
Large-scale Retrieval for Reinforcement Learning
Humphreys, P.; Guez, A.; Tieleman, O.; Sifre, L.; Weber, T.; and Lillicrap, T. 2022 · 2022
Earlier work this paper cites.
Large Language Models are Zero-Shot Reasoners
Kojima, T.; Gu, S. S.; Reid, M.; Matsuo, Y.; and Iwasawa, Y. 2022 · 2022
Cited alongside, same era.
A Survey on Retrieval-Augmented Text Generation
Li, H.; Su, Y.; Cai, D.; Wang, Y.; and Liu, L. 2022 · 2022
Cited alongside, same era.
What Makes Good In-Context Examples for GPT-3?
Liu, J.; Shen, D.; Zhang, Y.; Dolan, B.; Carin, L.; and Chen, W. 2022 · 2022
Cited alongside, same era.
Training Language Models to Follow Instructions with Human Feedback
Ouyang, L.; Wu, J.; Jiang, X.; Almeida, D.; Wainwright, C.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; Schulman, J.; Hilton, J.; Kelton, F.; Miller, L.; Simens, M.; Askell, A.; Welinder, P.; Christiano, P. F.; Leike, J.; and Lowe, R. 2022 · 2022
Cited alongside, same era.
Learning To Retrieve Prompts for In-Context Learning
Rubin, O.; Herzig, J.; and Berant, J. 2022 · 2022
Cited alongside, same era.
LaMDA: Language Models for Dialog Applications
Thoppilan, R.; De Freitas, D.; Hall, J.; Shazeer, N.; Kulshreshtha, A.; Cheng, H.-T.; Jin, A.; Bos, T.; Baker, L.; Du, Y.; et al. 2022 · 2022
Large Language Models as General Pattern Machines
Mirchandani, S.; Xia, F.; Florence, P.; Ichter, B.; Driess, D.; Arenas, M. G.; Rao, K.; Sadigh, D.; and Zeng, A. 2023 · 2023
Closest in time.
EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought
Mu, Y.; Zhang, Q.; Hu, M.; Wang, W.; Ding, M.; Jin, J.; Wang, B.; Dai, J.; Qiao, Y.; and Luo, P. 2023 · 2023
Closest in time.
GPT-4 Technical Report
OpenAI. 2023 · 2023
Closest in time.
Generative Agents: Interactive Simulacra of Human Behavior
Park, J. S.; O’Brien, J.; Cai, C. J.; Morris, M. R.; Liang, P.; and Bernstein, M. S. 2023 · 2023
Closest in time.
Communicative Agents for Software Development
Qian, C.; Cong, X.; Yang, C.; Chen, W.; Su, Y.; Xu, J.; Liu, Z.; and Sun, M. 2023 · 2023
Closest in time.
From Pixels to UI Actions: Learning to Follow Instructions via Graphical User Interfaces
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
Yao, S.; Chen, H.; Yang, J.; and Narasimhan, K. 2022 · 2022
Cited alongside, same era.
Introducing Claude
Anthropic. 2023 · 2023
Cited alongside, same era.
Emergent Autonomous Scientific Research Capabilities of Large Language Models
Boiko, D. A.; MacKnight, R.; and Gomes, G. 2023 · 2023
Cited alongside, same era.
ChemCrow: Augmenting Large-Language Models with Chemistry Tools
Bran, A. M.; Cox, S.; White, A. D.; and Schwaller, P. 2023 · 2023
Cited alongside, same era.
Langchain
Chase, H. 2023 · 2023
Cited alongside, same era.
PaLM: Scaling Language Modeling with Pathways
Chowdhery, A.; Narang, S.; Devlin, J.; Bosma, M.; Mishra, G.; Roberts, A.; Barham, P.; Chung, H. W.; Sutton, C.; Gehrmann, S.; Schuh, P.; Shi, K.; Tsvyashchenko, S.; Maynez, J.; Rao, A.; Barnes, P.; Tay, Y.; Shazeer, N.; Prabhakaran, V.; Reif, E.; Du, N.; Hutchinson, B.; Pope, R.; Bradbury, J.; Austin, J.; Isard, M.; Gur-Ari, G.; Yin, P.; Duke, T.; Levskaya, A.; Ghemawat, S.; Dev, S.; Michalewski, H.; Garcia, X.; Misra, V.; Robinson, K.; Fedus, L.; Zhou, D.; Ippolito, D.; Luan, D.; Lim, H.; Zoph, B.; Spiridonov, A.; Sepassi, R.; Dohan, D.; Agrawal, S.; Omernick, M.; Dai, A. M.; Pillai, T. S.; Pellat, M.; Lewkowycz, A.; Moreira, E.; Child, R.; Polozov, O.; Lee, K.; Zhou, Z.; Wang, X.; Saeta, B.; Diaz, M.; Firat, O.; Catasta, M.; Wei, J.; Meier-Hellstern, K.; Eck, D.; Dean, J.; Petrov, S.; and Fiedel, N. 2023 · 2023
Cited alongside, same era.
Shaw, P.; Joshi, M.; Cohan, J.; Berant, J.; Pasupat, P.; Hu, H.; Khandelwal, U.; Lee, K.; and Toutanova, K. 2023 · 2023
Closest in time.
Reflexion: Language Agents with Verbal Reinforcement Learning
Shinn, N.; Cassano, F.; Gopinath, A.; Narasimhan, K. R.; and Yao, S. 2023 · 2023
Closest in time.
Cognitive Architectures for Language Agents
Sumers, T. R.; Yao, S.; Narasimhan, K.; and Griffiths, T. L. 2023 · 2023
Closest in time.
AdaPlanner: Adaptive Planning from Feedback with Language Models
Sun, H.; Zhuang, Y.; Kong, L.; Dai, B.; and Zhang, C. 2023 · 2023
Closest in time.
Stanford Alpaca: An Instruction-Following LLaMA Model
Taori, R.; Gulrajani, I.; Zhang, T.; Dubois, Y.; Li, X.; Guestrin, C.; Liang, P.; and Hashimoto, T. B. 2023 · 2023
Closest in time.
Focused Transformer: Contrastive Training for Context Scaling
Tworkowski, S.; Staniszewski, K.; Pacek, M.; Wu, Y.; Michalewski, H.; and Miłoś, P. 2023 · 2023
Closest in time.
Learning to Retrieve In-Context Examples for Large Language Models
Wang, L.; Yang, N.; and Wei, F. 2023 · 2023
Closest in time.
TidyBot: Personalized Robot Assistance with Large Language Models
Wu, J.; Antonova, R.; Kan, A.; Lepert, M.; Zeng, A.; Song, S.; Bohg, J.; Rusinkiewicz, S.; and Funkhouser, T. 2023 · 2023
Closest in time.
The Rise and Potential of Large Language Model Based Agents: A Survey
Xi, Z.; Chen, W.; Guo, X.; He, W.; Ding, Y.; Hong, B.; Zhang, M.; Wang, J.; Jin, S.; Zhou, E.; et al. 2023 · 2023
Closest in time.
Offline Prioritized Experience Replay
Yue, Y.; Kang, B.; Ma, X.; Huang, G.; Song, S.; and Yan, S. 2023 · 2023
Closest in time.
AgentTuning: Enabling Generalized Agent Abilities for LLMs
Zeng, A.; Liu, M.; Lu, R.; Wang, B.; Liu, X.; Dong, Y.; and Tang, J. 2023 · 2023
Closest in time.
Automatic Chain of Thought Prompting in Large Language Models
Zhang, Z.; Zhang, A.; Li, M.; and Smola, A. 2023 · 2023
Closest in time.
RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Zitkovich, B.; Yu, T.; Xu, S.; Xu, P.; Xiao, T.; Xia, F.; Wu, J.; Wohlhart, P.; Welker, S.; Wahid, A.; et al. 2023 · 2023
Closest in time.