Fetching the paper…
Reading the bibliography…
Recently, tool learning with large language models (LLMs) has emerged as a promising paradigm for augmenting the capabilities of LLMs to tackle highly complex problems.
Towards completeness-oriented tool retrieval for large language models
Qu C, Dai S, Wei X, Cai H, Wang S, Yin D, Xu J, Wen J R · 1940
Earlier work this paper cites.
Tools and human evolution
Washburn S L · 1960
Earlier work this paper cites.
A statistical interpretation of term specificity and its application in retrieval
Sparck Jones K · 1972
Earlier work this paper cites.
Tools, language and cognition in human evolution
Gibson K R, Gibson K R, Ingold T · 1993
Earlier work this paper cites.
What is cognitive science?
Von Eckardt B · 1995
Earlier work this paper cites.
Cumulated gain-based evaluation of ir techniques
Järvelin K, Kekäläinen J · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni K, Roukos S, Ward T, Zhu W J · 2002
Earlier work this paper cites.
Recall, precision and average precision
Zhu M · 2004
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin C Y · 2004
Earlier work this paper cites.
The probabilistic relevance framework: Bm25 and beyond
Robertson S, Zaragoza H, others · 2009
Earlier work this paper cites.
cem: Coarsened exact matching in stata
Blackwell M, Iacus S, King G, Porro G · 2009
Earlier work this paper cites.
Animal tool behavior: the use and manufacture of tools by animals
Shumaker R W, Walkup K R, Beck B B · 2011
Earlier work this paper cites.
Hotpotqa: A dataset for diverse, explainable multi-hop question answering
Yang Z, Qi P, Zhang S, Bengio Y, Cohen W W, Salakhutdinov R, Manning C D · 2018
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
Kwiatkowski T, Palomaki J, Redfield O, Collins M, Parikh A, Alberti C, Epstein D, Polosukhin I, Devlin J, Lee K, others · 2019
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Reimers N, Gurevych I · 2019
Earlier work this paper cites.
Universal adversarial triggers for attacking and analyzing NLP
Wallace E, Feng S, Kandpal N, Gardner M, Singh S · 2019
Earlier work this paper cites.
Explainable ai: A review of machine learning interpretability methods
Linardatos P, Papastefanopoulos V, Kotsiantis S · 2020
Earlier work this paper cites.
Is bert really robust? a strong baseline for natural language attack on text classification and entailment
Jin D, Jin Z, Zhou J T, Szolovits P · 2020
Earlier work this paper cites.
Automatic text summarization: A comprehensive survey
El-Kassas W S, Salama C R, Rafea A A, Mohamed H K · 2021
Earlier work this paper cites.
Webgpt: Browser-assisted question-answering with human feedback
Nakano R, Hilton J, Balaji S, Wu J, Ouyang L, Kim C, Hesse C, Jain S, Kosaraju V, Saunders W, others · 2021
Earlier work this paper cites.
Training verifiers to solve math word problems
Cobbe K, Kosaraju V, Bavarian M, Chen M, Jun H, Kaiser L, Plappert M, Tworek J, Hilton J, Nakano R, others · 2021
Earlier work this paper cites.
Approximate nearest neighbor negative contrastive learning for dense text retrieval
Xiong L, Xiong C, Li Y, Tang K F, Liu J, Bennett P, Ahmed J, Overwijk A · 2021
Earlier work this paper cites.
Efficiently teaching an effective dense retriever with balanced topic aware sampling
Hofstätter S, Lin S C, Yang J H, Lin J, Hanbury A · 2021
Earlier work this paper cites.
Unsupervised dense information retrieval with contrastive learning
Izacard G, Caron M, Hosseini L, Riedel S, Bojanowski P, Joulin A, Grave E · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Chen M, Tworek J, Jun H, Yuan Q, Pinto H P d O, Kaplan J, Edwards H, Burda Y, Joseph N, Brockman G, others · 2021
Earlier work this paper cites.
Program synthesis with large language models
Austin J, Odena A, Nye M, Bosma M, Michalewski H, Dohan D, Jiang E, Cai C, Terry M, Le Q, others · 2021
Earlier work this paper cites.
Ethical and social risks of harm from language models
Weidinger L, Mellor J, Rauh M, Griffin C, Uesato J, Huang P S, Cheng M, Glaese M, Balle B, Kasirzadeh A, others · 2021
Earlier work this paper cites.
Mallen A, Asai A, Zhong V, Das R, Khashabi D, Hajishirzi H · 2022
Earlier work this paper cites.
Webshop: Towards scalable real-world web interaction with grounded language agents
Yao S, Chen H, Yang J, Narasimhan K · 2022
Earlier work this paper cites.
Internet-augmented language models through few-shot prompting for open-domain question answering
Lazaridou A, Gribovskaya E, Stokowiec W, Grigorev N · 2022
Earlier work this paper cites.
Talm: Tool augmented language models
Parisi A, Zhao Y, Fiedel N · 2022
Earlier work this paper cites.
Karpas E, Abend O, Belinkov Y, Lenz B, Lieber O, Ratner N, Shoham Y, Bata H, Levine Y, Leyton-Brown K, others · 2022
Earlier work this paper cites.
Reasoning with language model prompting: A survey
Qiao S, Ou Y, Zhang N, Chen X, Yao Y, Deng S, Tan C, Huang F, Chen H · 2022
Earlier work this paper cites.
Internet-augmented dialogue generation
Komeili M, Shuster K, Weston J · 2022
Earlier work this paper cites.
Lamda: Language models for dialog applications
Thoppilan R, De Freitas D, Hall J, Shazeer N, Kulshreshtha A, Cheng H T, Jin A, Bos T, Baker L, Du Y, others · 2022
Earlier work this paper cites.
Chaining simultaneous thoughts for numerical reasoning
Shao Z, Huang F, Huang M · 2022
Earlier work this paper cites.
Chen W, Ma X, Wang X, Cohen W W · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei J, Wang X, Schuurmans D, Bosma M, Xia F, Chi E, Le Q V, Zhou D, others · 2022
Earlier work this paper cites.
React: Synergizing reasoning and acting in language models
Yao S, Zhao J, Yu D, Du N, Shafran I, Narasimhan K, Cao Y · 2022
Earlier work this paper cites.
Unsupervised corpus aware language model pre-training for dense passage retrieval
Gao L, Callan J · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Ouyang L, Wu J, Jiang X, Almeida D, Wainwright C, Mishkin P, Zhang C, Agarwal S, Slama K, Ray A, others · 2022
Earlier work this paper cites.
Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
Huang W, Abbeel P, Pathak D, Mordatch I · 2022
Earlier work this paper cites.
Achiam J, Adler S, Agarwal S, Ahmad L, Akkaya I, Aleman F L, Almeida D, Altenschmidt J, Altman S, Anadkat S, others · 2023
Earlier work this paper cites.
Prompting large language model for machine translation: A case study
Zhang B, Haddow B, Birch A · 2023
Earlier work this paper cites.
Freshllms: Refreshing large language models with search engine augmentation
Vu T, Iyyer M, Wang X, Constant N, Wei J, Wei J, Tar C, Sung Y H, Zhou D, Le Q, others · 2023
Earlier work this paper cites.
Survey of hallucination in natural language generation
Ji Z, Lee N, Frieske R, Yu T, Su D, Xu Y, Ishii E, Bang Y J, Madotto A, Fung P · 2023
Earlier work this paper cites.
Siren’s song in the ai ocean: a survey on hallucination in large language models
Zhang Y, Li Y, Cui L, Cai D, Liu L, Fu T, Huang X, Zhao E, Zhang Y, Chen Y, others · 2023
Earlier work this paper cites.
Tool learning with foundation models
Qin Y, Hu S, Lin Y, Chen W, Ding N, Cui G, Zeng Z, Huang Y, Xiao C, Han C, others · 2023
Earlier work this paper cites.
Toolalpaca: Generalized tool learning for language models with 3000 simulated cases
Tang Q, Deng Z, Lin H, Han X, Liang Q, Sun L · 2023
Earlier work this paper cites.
Gear: Augmenting language models with generalizable and efficient tool resolution
Lu Y, Yu H, Khashabi D · 2023
Earlier work this paper cites.
Fact-checking complex claims with program-guided reasoning
Pan L, Wu X, Lu X, Luu A T, Wang W Y, Kan M Y, Nakov P · 2023
Earlier work this paper cites.
Vipergpt: Visual inference via python execution for reasoning
Surís D, Menon S, Vondrick C · 2023
Earlier work this paper cites.
Api-bank: A comprehensive benchmark for tool-augmented llms
Li M, Zhao Y, Yu B, Song F, Li H, Yu H, Li Z, Huang F, Li Y · 2023
Earlier work this paper cites.
T-eval: Evaluating the tool utilization capability step by step
Chen Z, Du W, Zhang W, Liu K, Liu J, Zheng M, Zhuo J, Zhang S, Lin D, Chen K, others · 2023
Earlier work this paper cites.
On the tool manipulation capability of open-source large language models
Xu Q, Hong F, Li B, Hu C, Chen Z, Zhang J · 2023
Earlier work this paper cites.
A survey of large language models
Zhao W X, Zhou K, Li J, Tang T, Wang X, Hou Y, Min Y, Zhang B, Zhang J, Dong Z, others · 2023
Earlier work this paper cites.
A survey of reasoning with foundation models
Sun J, Zheng C, Xie E, Liu Z, Chu R, Qiu J, Xu J, Ding M, Li H, Geng M, others · 2023
Earlier work this paper cites.
The rise and potential of large language model based agents: A survey
Xi Z, Chen W, Guo X, He W, Ding Y, Hong B, Zhang M, Wang J, Jin S, Zhou E, others · 2023
Earlier work this paper cites.
Retrieval-augmented generation for large language models: A survey
Gao Y, Xiong Y, Gao X, Jia K, Pan J, Bi Y, Dai Y, Sun J, Wang H · 2023
Earlier work this paper cites.
Augmented language models: a survey
Mialon G, Dessì R, Lomeli M, Nalmpantis C, Pasunuru R, Raileanu R, Rozière B, Schick T, Dwivedi-Yu J, Celikyilmaz A, others · 2023
Earlier work this paper cites.
Toolcoder: Teach code generation models to use api search tools
Zhang K, Zhang H, Li G, Li J, Li Z, Jin Z · 2023
Earlier work this paper cites.
Replug: Retrieval-augmented black-box language models
Shi W, Min S, Yasunaga M, Seo M, James R, Lewis M, Zettlemoyer L, Yih W t · 2023
Earlier work this paper cites.
Art: Automatic multi-step reasoning and tool-use for large language models
Paranjape B, Lundberg S, Singh S, Hajishirzi H, Zettlemoyer L, Ribeiro M T · 2023
Earlier work this paper cites.
Gorilla: Large language model connected with massive apis
Patil S G, Zhang T, Wang X, Gonzalez J E · 2023
Earlier work this paper cites.
Calc-x and calcformers: Empowering arithmetical chain-of-thought through interaction with symbolic systems
Kadlčík M, Štefánik M, Sotolar O, Martinek V · 2023
Earlier work this paper cites.
Solving math word problems by combining language models with symbolic solvers
He-Yueya J, Poesia G, Wang R E, Goodman N D · 2023
Earlier work this paper cites.
Pal: Program-aided language models
Gao L, Madaan A, Zhou S, Alon U, Liu P, Yang Y, Callan J, Neubig G · 2023
Earlier work this paper cites.
Leti: Learning to generate from textual interactions
Wang X, Peng H, Jabbarvand R, Ji H · 2023
Cited alongside, same era.
MultiTool-CoT: GPT-3 can use multiple external tools with chain of thought prompting
Inaba T, Kiyomaru H, Cheng F, Kurohashi S · 2023
Cited alongside, same era.
Mm-react: Prompting chatgpt for multimodal reasoning and action
Yang Z, Li L, Wang J, Lin K, Azarnasab E, Ahmed F, Liu Z, Liu C, Zeng M, Wang L · 2023
Cited alongside, same era.
Internchat: Solving vision-centric tasks by interacting with chatbots beyond language
Liu Z, He Y, Wang W, Wang W, Wang Y, Chen S, Zhang Q, Yang Y, Li Q, Yu J, others · 2023
Cited alongside, same era.
Assistgpt: A general multi-modal assistant that can plan, execute, inspect, and learn
Mmedagent: Learning to use medical tools with multi-modal agent
Li B, Yan T, Pan Y, Xu Z, Luo J, Ji R, Liu S, Dong H, Lin Z, Wang Y · 2024
Closest in time.
Clova: A closed-loop visual assistant with tool usage and update
Gao Z, Du Y, Zhang X, Ma X, Han W, Zhu S C, Li Q · 2024
Closest in time.
Diffagent: Fast and accurate text-to-image api selection with large language model
Zhao L, Yang Y, Zhang K, Shao W, Zhang Y, Qiao Y, Luo P, Ji R · 2024
Closest in time.
m&m’s: A benchmark to evaluate tool-use for multi-step multi-modal tasks
Ma Z, Huang W, Zhang J, Gupta T, Krishna R · 2024
Closest in time.
Tool-lmm: A large multi-modal model for tool agent learning
Wang C, Luo W, Chen Q, Mai H, Guo J, Dong S, Li Z, Ma L, Gao S, others · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gao D, Ji L, Zhou L, Lin K Q, Chen J, Fan Z, Shou M Z · 2023
Cited alongside, same era.
Gitagent: Facilitating autonomous agent with github by tool extension
Lyu B, Cong X, Yu H, Yang P, Qin Y, Ye Y, Lu Y, Zhang Z, Yan Y, Lin Y, others · 2023
Cited alongside, same era.
Restgpt: Connecting large language models with real-world applications via restful apis
Song Y, Xiong W, Zhu D, Li C, Wang K, Tian Y, Li S · 2023
Cited alongside, same era.
Tptu: Task planning and tool usage of large language model-based ai agents
Ruan J, Chen Y, Zhang B, Xu Z, Bao T, Du G, Shi S, Mao H, Zeng X, Zhao R · 2023
Cited alongside, same era.
Controlllm: Augment language models with tools by searching on graphs
Liu Z, Lai Z, Gao Z, Cui E, Li Z, Zhu X, Lu L, Chen Q, Qiao Y, Dai J, others · 2023
Cited alongside, same era.
Toolink: Linking toolkit creation and using through chain-of-solving on open-source model
Qian C, Xiong C, Liu Z, Liu Z · 2023
Cited alongside, same era.
Tptu-v2: Boosting task planning and tool usage of large language model-based agents in real-world systems, 2023
Kong Y, Ruan J, Chen Y, Zhang B, Bao T, Shi S, Du G, Hu X, Mao H, Li Z, Zeng X, Zhao R · 2023
Cited alongside, same era.
Protip: Progressive tool retrieval improves planning
Anantha R, Bandyopadhyay B, Kashi A, Mahinder S, Hill A W, Chappidi S · 2023
Cited alongside, same era.
Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face
Shen Y, Song K, Tan X, Li D, Lu W, Zhuang Y · 2024
Closest in time.
Toolchain*: Efficient action space navigation in large language models with a* search
Zhuang Y, Chen X, Yu T, Mitra S, Bursztyn V, Rossi R A, Sarkhel S, Zhang C · 2024
Closest in time.
Fortify the shortest stave in attention: Enhancing context awareness of large language models for effective tool use
Chen Y, Lv A, Lin T E, Chen C, Wu Y, Huang F, Li Y, Yan R · 2024
Closest in time.
Planning and editing what you retrieve for enhanced tool learning
Huang T, Jung D, Chen M · 2024
Closest in time.
Chain of tools: Large language model is an automatic multi-tool learner
Shi Z, Gao S, Chen X, Feng Y, Yan L, Shi H, Yin D, Chen Z, Verberne S, Ren Z · 2024
Closest in time.
Can graph learning improve task planning?
Wu X, Shen Y, Shan C, Song K, Wang S, Zhang B, Feng J, Cheng H, Chen W, Xiong Y, others · 2024
Closest in time.
From summary to action: Enhancing large language models for complex tasks with open world apis
Liu Y, Yuan Y, Wang C, Han J, Ma Y, Zhang L, Zheng N, Xu H · 2024
Closest in time.
Budget-constrained tool learning with planning
Zheng Y, Li P, Yan M, Zhang J, Huang F, Liu Y · 2024
Closest in time.
From exploration to mastery: Enabling llms to master tools via self-driven interactions
Qu C, Dai S, Wei X, Cai H, Wang S, Yin D, Xu J, Wen J R · 2024
Closest in time.
Taskmatrix. ai: Completing tasks by connecting foundation models with millions of apis
Liang Y, Wu C, Song T, Wu W, Xia Y, Liu Y, Ou Y, Lu S, Ji L, Mao S, others · 2024
Closest in time.
Small llms are weak tool learners: A multi-llm agent
Shen W, Li C, Chen H, Yan M, Quan X, Chen H, Zhang J, Huang F · 2024
Closest in time.
Efficient tool use with chain-of-abstraction reasoning
Gao S, Dwivedi-Yu J, Yu P, Tan X E, Pasunuru R, Golovneva O, Sinha K, Celikyilmaz A, Bosselut A, Wang T · 2024
Closest in time.
Look before you leap: Towards decision-aware and generalizable tool-usage for large language models
Gui A, Li J, Dai Y, Du N, Xiao H · 2024
Closest in time.
Openagi: When llm meets domain experts
Ge Y, Hua W, Mei K, Tan J, Xu S, Li Z, Zhang Y, others · 2024
Closest in time.
A solution-based llm api-using methodology for academic information seeking
Wang Y, Yu J, Yao Z, Zhang J, Xie Y, Tu S, Fu Y, Feng Y, Zhang J, Zhang J, others · 2024
Closest in time.
Advancing tool-augmented large language models: Integrating insights from errors in inference trees
Chen S, Wang Y, Wu Y F, Chen Q G, Xu Z, Luo W, Zhang K, Zhang L · 2024
Closest in time.
Apigen: Automated pipeline for generating verifiable and diverse function-calling datasets
Liu Z, Hoang T, Zhang J, Zhu M, Lan T, Kokane S, Tan J, Yao W, Liu Z, Feng Y, others · 2024
Closest in time.
Craft: Customizing llms by creating and retrieving from specialized toolsets
Yuan L, Chen Y, Wang X, Fung Y R, Peng H, Ji H · 2024
Closest in time.
Toolrerank: Adaptive and hierarchy-aware reranking for tool retrieval
Zheng Y, Li P, Liu W, Liu Y, Luan J, Wang B · 2024
Closest in time.
Chatcot: Tool-augmented chain-of-thought reasoning on chat-based large language models
Chen Z, Zhou K, Zhang B, Gong Z, Zhao X, Wen J R · 2024
Closest in time.
Toolnet: Connecting large language models with massive tools via tool graph
Liu X, Peng Z, Yi X, Xie X, Xiang L, Liu Y, Xu D · 2024
Closest in time.
Toolverifier: Generalization to new tools via self-verification
Mekala D, Weston J, Lanchantin J, Raileanu R, Lomeli M, Shang J, Dwivedi-Yu J · 2024
Closest in time.
Making language models better tool learners with execution feedback
Qiao S, Gui H, Chen H, Zhang N · 2024
Closest in time.
Anytool: Self-reflective, hierarchical agents for large-scale api calls
Du Y, Wei F, Zhang H · 2024
Closest in time.
Geckopt: Llm system efficiency via intent-based tool selection
Fore M, Singh S, Stamoulis D · 2024
Closest in time.
Easytool: Enhancing llm-based agents with concise tool instruction
Yuan S, Song K, Chen J, Tan X, Shen Y, Kan R, Li D, Yang D · 2024
Closest in time.
Learning to use tools via cooperative and interactive agents
Shi Z, Gao S, Chen X, Yan L, Shi H, Yin D, Chen Z, Ren P, Verberne S, Ren Z · 2024
Closest in time.
Gpt4tools: Teaching large language model to use tools via self-instruction
Yang R, Song L, Li Y, Zhao S, Ge Y, Li X, Shan Y · 2024
Closest in time.
Llms in the imaginarium: Tool learning through simulated trial and error
Wang B, Fang H, Eisner J, Van Durme B, Su Y · 2024
Closest in time.
Recomp: Improving retrieval-augmented lms with compression and selective augmentation
Xu F, Shi W, Choi E · 2024
Closest in time.
Ye J, Li G, Gao S, Huang C, Wu Y, Li S, Fan X, Dou S, Zhang Q, Gui T, others · 2024
Closest in time.
Huang S, Zhong W, Lu J, Zhu Q, Gao J, Liu W, Hou Y, Zeng X, Wang Y, Shang L, others · 2024
Closest in time.
Seal-tools: Self-instruct tool learning dataset for agent tuning and detailed benchmark
Wu M, Zhu T, Han H, Tan C, Zhang X, Chen W · 2024
Closest in time.
Api-blend: A comprehensive corpora for training and benchmarking api llms
Basu K, Abdelaziz I, Chaudhury S, Dan S, Crouse M, Munawar A, Kumaravel S, Muthusamy V, Kapanipathi P, Lastras L A · 2024
Closest in time.
Shortcutsbench: A large-scale real-world benchmark for api-based agents
Shen H, Li Y, Meng D, Cai D, Qi S, Zhang L, Xu M, Ma Y · 2024
Closest in time.
Gta: A benchmark for general tool agents
Wang J, Ma Z, Li Y, Zhang S, Chen C, Chen K, Le X · 2024
Closest in time.
Wtu-eval: A whether-or-not tool usage evaluation benchmark for large language models
Ning K, Su Y, Lv X, Zhang Y, Liu J, Liu K, Xu J · 2024
Closest in time.
Appworld: A controllable world of apps and people for benchmarking interactive coding agents
Trivedi H, Khot T, Hartmann M, Manku R, Dong V, Li E, Gupta S, Sabharwal A, Balasubramanian N · 2024
Closest in time.
Ye J, Wu Y, Gao S, Li S, Li G, Fan X, Zhang Q, Gui T, Huang X · 2024
Closest in time.
Toolsword: Unveiling safety issues of large language models in tool learning across three stages
Ye J, Li S, Li G, Huang C, Gao S, Wu Y, Zhang Q, Gui T, Huang X · 2024
Closest in time.
Sciagent: Tool-augmented language models for scientific reasoning
Ma Y, Gou Z, Hao J, Xu R, Wang S, Pan L, Yang Y, Cao Y, Sun A · 2024
Closest in time.
Stabletoolbench: Towards stable large-scale benchmarking on tool learning of large language models
Guo Z, Cheng S, Wang H, Liang S, Qin Y, Li P, Liu Z, Sun M, Liu Y · 2024
Closest in time.
Injecagent: Benchmarking indirect prompt injections in tool-integrated large language model agents
Zhan Q, Liang Z, Ying Z, Kang D · 2024
Closest in time.
Ctooleval: A chinese benchmark for llm-powered agent evaluation in real-world api interactions
Guo Z, Huang Y, Xiong D · 2024
Closest in time.
Lu J, Holleis T, Zhang Y, Aumayer B, Nan F, Bai F, Ma S, Ma S, Li M, Yin G, others · 2024
Closest in time.
Evaluating tool-augmented agents in remote sensing platforms
Singh S, Fore M, Stamoulis D · 2024
Closest in time.
Explainability for large language models: A survey
Zhao H, Chen H, Yang F, Liu N, Deng H, Cai H, Wang S, Yin D, Du M · 2024
Closest in time.
A new era in llm security: Exploring security concerns in real-world llm-based systems
Wu F, Zhang N, Jha S, McDaniel P, Xiao C · 2024
Closest in time.
Stride: A tool-assisted llm agent framework for strategic and interactive decision-making
Li C, Yang R, Li T, Bafarassat M, Sharifi K, Bergemann D, Yang Z · 2024
Closest in time.
Search-in-the-chain: Towards the accurate, credible and traceable content generation for complex knowledge-intensive tasks
Xu S, Pang L, Shen H, Cheng X, Chua T s · 2024
Closest in time.
Language models can solve computer tasks
Kim G, Baldi P, McAleer S · 2024
Closest in time.
Tool-planner: Dynamic solution tree planning for large language model with tool clustering
Liu Y, Peng X, Zhang Y, Cao J, Zhang X, Cheng S, Wang X, Yin J, Du T · 2024
Closest in time.
Erbacher P, Falissar L, Guigue V, Soulier L · 2024
Closest in time.
Enhancing tool retrieval with iterative feedback from large language models
Xu Q, Li Y, Xia H, Li W · 2024
Closest in time.
Gemini: A family of highly capable multimodal models, 2024
Team G, Anil R, Borgeaud S, Alayrac J B, Yu J, Soricut R, Schalkwyk J, Dai A M, Hauth A, others · 2024
Closest in time.
Concise and precise context compression for tool-using language models
Xu Y, Feng Y, Mu H, Hou Y, Li Y, Wang X, Zhong W, Li Z, Tu D, Zhu Q, others · 2024
Closest in time.
A survey on large language model (llm) security and privacy: The good, the bad, and the ugly
Yao Y, Duan J, Xu K, Cai Y, Sun Z, Zhang Y · 2024
Closest in time.
Risk taxonomy, mitigation, and assessment benchmarks of large language model systems
Cui T, Wang Y, Fu C, Xiao Y, Li S, Deng X, Liu Y, Zhang Q, Qiu Z, Li P, others · 2024
Closest in time.
Security and privacy challenges of large language models: A survey
Das B C, Amini M H, Wu Y · 2024
Closest in time.
Large language models as tool makers
Cai T, Wang X, Ma T, Chen X, Zhou D · 2024
Closest in time.
Trove: Inducing verifiable and efficient toolboxes for solving programmatic tasks
Wang Z, Fried D, Neubig G · 2024
Closest in time.