Fetching the paper…
Reading the bibliography…
Instruction-tuned large language models (LLMs) excel at many tasks but often fail to use external tools due to complicated and unfamiliar syntax constraints.
Trie memory
Fredkin, E · 1960
Earlier work this paper cites.
Parameter estimation for probabilistic finite-state transducers
Eisner, J · 2002
Earlier work this paper cites.
Weighting finite-state transductions with neural context
Rastogi, P., Cotterell, R., and Eisner, J · 2016
Earlier work this paper cites.
Guided open vocabulary image captioning with constrained beam search
Anderson, P., Fernando, B., Johnson, M., and Gould, S · 2017
Earlier work this paper cites.
Lexically constrained decoding for sequence generation using grid beam search
Hokamp, C. and Liu, Q · 2017
Earlier work this paper cites.
A syntactic neural model for general-purpose code generation
Yin, P. and Neubig, G · 2017
Earlier work this paper cites.
Cgmh: Constrained sentence generation by metropolis-hastings sampling
Miao, N., Zhou, H., Mou, L., Yan, R., and Li, L · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Earlier work this paper cites.
Retrieval augmented language model pre-training
Guu, K., Lee, K., Tung, Z., Pasupat, P., and Chang, M · 2020
Earlier work this paper cites.
Program synthesis with large language models
Austin, J., Odena, A., Nye, M., Bosma, M., Michalewski, H., Dohan, D., Jiang, E., Cai, C., Terry, M., Le, Q., et al · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Chen, M., Tworek, J., Jun, H., Yuan, Q., Pinto, H. P. d. O., Kaplan, J., Edwards, H., Burda, Y., Joseph, N., Brockman, G., et al · 2021
Earlier work this paper cites.
Training verifiers to solve math word problems
Cobbe, K., Kosaraju, V., Bavarian, M., Chen, M., Jun, H., Kaiser, L., Plappert, M., Tworek, J., Hilton, J., Nakano, R., Hesse, C., and Schulman, J · 2021
Earlier work this paper cites.
Neurologic decoding:(un) supervised neural text generation with predicate logic constraints
Lu, X., West, P., Zellers, R., Le Bras, R., Bhagavatula, C., and Choi, Y · 2021
Earlier work this paper cites.
Webgpt: Browser-assisted question-answering with human feedback
Nakano, R., Hilton, J., Balaji, S., Wu, J., Ouyang, L., Kim, C., Hesse, C., Jain, S., Kosaraju, V., Saunders, W., et al · 2021
Cited alongside, same era.
Improving language models by retrieving from trillions of tokens
Borgeaud, S., Mensch, A., Hoffmann, J., Cai, T., Rutherford, E., Millican, K., Van Den Driessche, G. B., Lespiau, J.-B., Damoc, B., Clark, A., et al · 2022
Cited alongside, same era.
Program of thoughts prompting: Disentangling computation from reasoning for numerical reasoning tasks
Chen, W., Ma, X., Wang, X., and Cohen, W. W · 2022
Cited alongside, same era.
Visual programming: Compositional visual reasoning without training
Gupta, T. and Kembhavi, A · 2022
Cited alongside, same era.
Kamel : Knowledge analysis with multitoken entities in language models
Kalo, J.-C. and Fichtel, L · 2022
Toolkengpt: Augmenting frozen language models with massive tools via tool embeddings
Hao, S., Liu, T., Wang, Z., and Hu, Z · 2023
Closest in time.
Mistral 7b, 2023
Jiang, A. Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D. S., de las Casas, D., Bressand, F., Lengyel, G., Lample, G., Saulnier, L., Lavaud, L. R., Lachaux, M.-A., Stock, P., Scao, T. L., Lavril, T., Wang, T., Lacroix, T., and Sayed, W. E · 2023
Closest in time.
Augmented language models: a survey
Mialon, G., Dessì, R., Lomeli, M., Nalmpantis, C., Pasunuru, R., Raileanu, R., Rozière, B., Schick, T., Dwivedi-Yu, J., Celikyilmaz, A., et al · 2023
Closest in time.
Toolllm: Facilitating large language models to master 16000+ real-world apis
Qin, Y., Liang, S., Ye, Y., Zhu, K., Yan, L., Lu, Y., Lin, Y., Cong, X., Tang, X., Qian, B., et al · 2023
Closest in time.
Toolformer: Language models can teach themselves to use tools
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Neurologic a* esque decoding: Constrained text generation with lookahead heuristics
Lu, X., Welleck, S., West, P., Jiang, L., Kasai, J., Khashabi, D., Le Bras, R., Qin, L., Yu, Y., Zellers, R., et al · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., et al · 2022
Cited alongside, same era.
Talm: Tool augmented language models
Parisi, A., Zhao, Y., and Fiedel, N · 2022
Cited alongside, same era.
Challenging big-bench tasks and whether chain-of-thought can solve them
Suzgun, M., Scales, N., Schärli, N., Gehrmann, S., Tay, Y., Chung, H. W., Chowdhery, A., Le, Q. V., Chi, E. H., Zhou, D., , and Wei, J · 2022
Cited alongside, same era.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality, March 2023
Chiang, W.-L., Li, Z., Lin, Z., Sheng, Y., Wu, Z., Zhang, H., Zheng, L., Zhuang, S., Zhuang, Y., Gonzalez, J. E., Stoica, I., and Xing, E. P · 2023
Cited alongside, same era.
Pal: Program-aided language models
Gao, L., Madaan, A., Zhou, S., Alon, U., Liu, P., Yang, Y., Callan, J., and Neubig, G · 2023
Cited alongside, same era.
Grammar-constrained decoding for structured nlp tasks without finetuning
Geng, S., Josifoski, M., Peyrard, M., and West, R · 2023
Cited alongside, same era.
Schick, T., Dwivedi-Yu, J., Dessì, R., Raileanu, R., Lomeli, M., Zettlemoyer, L., Cancedda, N., and Scialom, T · 2023
Closest in time.
Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face, 2023
Shen, Y., Song, K., Tan, X., Li, D., Lu, W., and Zhuang, Y · 2023
Closest in time.
Restgpt: Connecting large language models with real-world restful apis, 2023
Song, Y., Xiong, W., Zhu, D., Wu, W., Qian, H., Song, M., Huang, H., Li, C., Wang, K., Yao, R., Tian, Y., and Li, S · 2023
Closest in time.
Llama: Open and efficient foundation language models, 2023
Touvron, H., Lavril, T., Izacard, G., Martinet, X., Lachaux, M.-A., Lacroix, T., Rozière, B., Goyal, N., Hambro, E., Azhar, F., Rodriguez, A., Joulin, A., Grave, E., and Lample, G · 2023
Closest in time.
Efficient guided generation for llms
Willard, B. T. and Louf, R · 2023
Closest in time.
React: Synergizing reasoning and acting in language models, 2023
Yao, S., Zhao, J., Yu, D., Du, N., Shafran, I., Narasimhan, K., and Cao, Y · 2023
Closest in time.
Zero and few-shot semantic parsing with ambiguous inputs, 2024
Stengel-Eskin, E., Rawlins, K., and Durme, B. V · 2024
Closest in time.
EASYTOOL: Enhancing LLM-based Agents with Concise Tool Instruction, January 2024
Yuan, S., Song, K., Chen, J., Tan, X., Shen, Y., Kan, R., Li, D., and Yang, D · 2024
Closest in time.