Fetching the paper…
Reading the bibliography…
Augmenting large language models (LLMs) with external tools has emerged as a promising approach to extending the capability of LLMs.
Cumulated Gain-Based Evaluation of IR Techniques
Järvelin, K.; and Kekäläinen, J. 2002 · 2002
Earlier work this paper cites.
Product-Aware Answer Generation in E-Commerce Question-Answering
Gao, S.; Ren, Z.; Zhao, Y.; Zhao, D.; Yin, D.; and Yan, R. 2019 · 2019
Earlier work this paper cites.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Reimers, N.; and Gurevych, I. 2019 · 2019
Earlier work this paper cites.
DeepSpeed: System Optimizations Enable Training Deep Learning Models with Over 100 Billion Parameters
Rasley, J.; Rajbhandari, S.; Ruwase, O.; and He, Y. 2020 · 2020
Earlier work this paper cites.
Curriculum Learning for Natural Language Understanding
Xu, B.; Zhang, L.; Mao, Z.; Wang, Q.; Xie, H.; and Zhang, Y. 2020 · 2020
Earlier work this paper cites.
WebGPT: Browser-assisted question-answering with human feedback
Nakano, R.; Hilton, J.; Balaji, S. A.; Wu, J.; Ouyang, L.; Kim, C.; Hesse, C.; Jain, S.; Kosaraju, V.; Saunders, W.; Jiang, X.; Cobbe, K.; Eloundou, T.; Krueger, G.; Button, K.; Knight, M.; Chess, B.; and Schulman, J. 2021 · 2021
Earlier work this paper cites.
GLM: General Language Model Pretraining with Autoregressive Blank Infilling
Du, Z.; Qian, Y.; Liu, X.; Ding, M.; Qiu, J.; Yang, Z.; and Tang, J. 2022 · 2022
Earlier work this paper cites.
A Drop of Ink Makes a Million Think: The Spread of False Information in Large Language Models
Bian, N.; Liu, P.; Han, X.; Lin, H.; Lu, Y.; He, B.; and Sun, L. 2023 · 2023
Earlier work this paper cites.
Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality
Chiang, W.-L.; Li, Z.; Lin, Z.; Sheng, Y.; Wu, Z.; Zhang, H.; Zheng, L.; Zhuang, S.; Zhuang, Y.; Gonzalez, J. E.; Stoica, I.; and Xing, E. P. 2023 · 2023
Earlier work this paper cites.
PAL: Program-aided Language Models
Gao, L.; Madaan, A.; Zhou, S.; Alon, U.; Liu, P.; Yang, Y.; Callan, J.; and Neubig, G. 2023 · 2023
Cited alongside, same era.
The False Promise of Imitating Proprietary LLMs
Gudibande, A.; Wallace, E.; Snell, C. B.; Geng, X.; Liu, H.; Abbeel, P.; Levine, S.; and Song, D. 2023 · 2023
Cited alongside, same era.
ToolkenGPT: Augmenting Frozen Language Models with Massive Tools via Tool Embeddings
Hao, S.; Liu, T.; Wang, Z.; and Hu, Z. 2023 · 2023
Cited alongside, same era.
GeneGPT: Augmenting Large Language Models with Domain Tools for Improved Access to Biomedical Information
Jin, Q.; Yang, Y.; Chen, Q.; and Lu, Z. 2023 · 2023
Cited alongside, same era.
Language Models can Solve Computer Tasks
Kim, G.; Baldi, P.; and McAleer, S. 2023 · 2023
Gorilla: Large Language Model Connected with Massive APIs
Patil, S. G.; Zhang, T.; Wang, X.; and Gonzalez, J. E. 2023 · 2023
Closest in time.
Toolformer: Language Models Can Teach Themselves to Use Tools
Schick, T.; Dwivedi-Yu, J.; Dessì, R.; Raileanu, R.; Lomeli, M.; Zettlemoyer, L.; Cancedda, N.; and Scialom, T. 2023 · 2023
Closest in time.
HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in HuggingFace
Shen, Y.; Song, K.; Tan, X.; Li, D. S.; Lu, W.; and Zhuang, Y. T. 2023 · 2023
Closest in time.
RADE: Reference-Assisted Dialogue Evaluation for Open-Domain Dialogue
Shi, Z.; Sun, W.; Zhang, S.; Zhang, Z.; Ren, P.; and Ren, Z. 2023 · 2023
Closest in time.
ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
API-Bank: A Benchmark for Tool-Augmented LLMs
Li, M.; Song, F.; Yu, B.; Yu, H.; Li, Z.; Huang, F.; and Li, Y. 2023 · 2023
Cited alongside, same era.
Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models
Lu, P.; Peng, B.; Cheng, H.; Galley, M.; Chang, K.-W.; Wu, Y. N.; Zhu, S.-C.; and Gao, J. 2023 · 2023
Cited alongside, same era.
ART: Automatic multi-step reasoning and tool-use for large language models
Paranjape, B.; Lundberg, S. M.; Singh, S.; Hajishirzi, H.; Zettlemoyer, L.; and Ribeiro, M. T. 2023 · 2023
Cited alongside, same era.
Generative Agents: Interactive Simulacra of Human Behavior
Park, J. S.; O’Brien, J. C.; Cai, C. J.; Morris, M. R.; Liang, P.; and Bernstein, M. S. 2023 · 2023
Cited alongside, same era.
WebCPM: Interactive Web Search for Chinese Long-form Question Answering
Qin, Y.; Cai, Z.; Jin, D.; Yan, L.; Liang, S.; Zhu, K.; Lin, Y.; Han, X.; Ding, N.; Wang, H.; Xie, R.; Qi, F.; Liu, Z.; Sun, M.; and Zhou, J. 2023a
Cited in the paper.
Tool Learning with Foundation Models
Qin, Y.; Hu, S.; Lin, Y.; Chen, W.; Ding, N.; Cui, G.; Zeng, Z.; Huang, Y.; Xiao, C.; Han, C.; Fung, Y. R.; Su, Y.; Wang, H.; Qian, C.; Tian, R.; Zhu, K.; Liang, S.; Shen, X.; Xu, B.; Zhang, Z.; Ye, Y.; Li, B.; Tang, Z.; Yi, J.; Zhu, Y.; Dai, Z.; Yan, L.; Cong, X.; Lu, Y.-T.; Zhao, W.; Huang, Y.; Yan, J.-H.; Han, X.; Sun, X.; Li, D.; Phang, J.; Yang, C.; Wu, T.; Ji, H.; Liu, Z.; and Sun, M. 2023b
Cited in the paper.
WizardLM: Empowering Large Language Models to Follow Complex Instructions
Xu, C.; Sun, Q.; Zheng, K.; Geng, X.; Zhao, P.; Feng, J.; Tao, C.; and Jiang, D. 2023a
Cited in the paper.
Tang, Q.; Deng, Z.; Lin, H.; Han, X.; Liang, Q.; and Sun, L. 2023 · 2023
Closest in time.
Wang, Z.; Cai, S.; Liu, A.; Ma, X.; and Liang, Y. 2023 · 2023
Closest in time.
Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models
Wu, C.; Yin, S.-K.; Qi, W.; Wang, X.; Tang, Z.; and Duan, N. 2023 · 2023
Closest in time.
ReAct: Synergizing Reasoning and Acting in Language Models
Yao, S.; Zhao, J.; Yu, D.; Du, N.; Shafran, I.; Narasimhan, K. R.; and Cao, Y. 2023 · 2023
Closest in time.