Fetching the paper…
Reading the bibliography…
Autonomous agents have long been a prominent research focus in both academic and industry communities.
Language models are few-shot learners
Brown T, Mann B, Ryder N, Subbiah M, Kaplan J D, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, others · 1901
Earlier work this paper cites.
Big five inventory
John O P, Donahue E M, Kentle R L · 1991
Earlier work this paper cites.
Dual-process theories of higher cognition: Advancing the debate
Evans J S B, Stanovich K E · 2013
Earlier work this paper cites.
Measuring thirty facets of the five factor model with a 120-item public domain inventory: Development of the ipip-neo-120
Johnson J A · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih V, Kavukcuoglu K, Silver D, Rusu A A, Veness J, Bellemare M G, Graves A, Riedmiller M, Fidjeland A K, Ostrovski G, others · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Lillicrap T P, Hunt J J, Pritzel A, Heess N, Erez T, Tassa Y, Silver D, Wierstra D · 2015
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman J, Wolski F, Dhariwal P, Radford A, Klimov O · 2017
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Haarnoja T, Zhou A, Abbeel P, Levine S · 2018
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford A, Wu J, Child R, Luan D, Amodei D, Sutskever I, others · 2019
Earlier work this paper cites.
Generative adversarial user model for reinforcement learning based recommendation system
Chen X, Li S, Li H, Jiang S, Qi Y, Song L · 2019
Earlier work this paper cites.
Webgpt: Browser-assisted question-answering with human feedback
Nakano R, Hilton J, Balaji S, Wu J, Ouyang L, Kim C, Hesse C, Jain S, Kosaraju V, Saunders W, others · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Chen M, Tworek J, Jun H, Yuan Q, Pinto H P d O, Kaplan J, Edwards H, Burda Y, Joseph N, Brockman G, others · 2021
Earlier work this paper cites.
Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
Huang W, Abbeel P, Pathak D, Mordatch I · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Wei J, Wang X, Schuurmans D, Bosma M, Xia F, Chi E, Le Q V, Zhou D, others · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
Kojima T, Gu S S, Reid M, Matsuo Y, Iwasawa Y · 2022
Earlier work this paper cites.
Planning with large language models via corrective re-prompting
Raman S S, Cohen V, Rosen E, Idrees I, Paulius D, Tellex S · 2022
Earlier work this paper cites.
Self-consistency improves chain of thought reasoning in language models
Wang X, Wei J, Schuurmans D, Le Q, Chi E, Narang S, Chowdhery A, Zhou D · 2022
Earlier work this paper cites.
Inner monologue: Embodied reasoning through planning with language models
Huang W, Xia F, Xiao T, Chan H, Liang J, Florence P, Zeng A, Tompson J, Mordatch I, Chebotar Y, others · 2022
Earlier work this paper cites.
Karpas E, Abend O, Belinkov Y, Lenz B, Lieber O, Ratner N, Shoham Y, Bata H, Levine Y, Leyton-Brown K, others · 2022
Earlier work this paper cites.
Do as i can, not as i say: Grounding language in robotic affordances
Ahn M, Brohan A, Brown N, Chebotar Y, Cortes O, David B, Finn C, Fu C, Gopalakrishnan K, Hausman K, others · 2022
Earlier work this paper cites.
Social simulacra: Creating populated prototypes for social computing systems
Park J S, Popowski L, Cai C, Morris M R, Liang P, Bernstein M S · 2022
Earlier work this paper cites.
Webshop: Towards scalable real-world web interaction with grounded language agents
Yao S, Chen H, Yang J, Narasimhan K · 2022
Earlier work this paper cites.
Memory-assisted prompt editing to improve GPT-3 after deployment
Madaan A, Tandon N, Clark P, Yang Y · 2022
Earlier work this paper cites.
A neural network solves, explains, and generates university math problems by program synthesis and few-shot learning at human level
Drori I, Zhang S, Shuttleworth R, Tang L, Lu A, Ke E, Liu K, Chen L, Tran S, Cheng N, others · 2022
Earlier work this paper cites.
Evaluating human-language model interaction
Lee M, Srivastava M, Hardy A, Thickstun J, Durmus E, Paranjape A, Gerard-Ursin I, Li X L, Ladhak F, Rong F, others · 2022
Earlier work this paper cites.
Towards reasoning in large language models: A survey
Huang J, Chang K C C · 2022
Earlier work this paper cites.
Achiam J, Adler S, Agarwal S, Ahmad L, Akkaya I, Aleman F L, Almeida D, Altenschmidt J, Altman S, Anadkat S, others · 2023
Earlier work this paper cites.
Model card and evaluations for claude models
Anthropic · 2023
Earlier work this paper cites.
Llama: Open and efficient foundation language models
Touvron H, Lavril T, Izacard G, Martinet X, Lachaux M A, Lacroix T, Rozière B, Goyal N, Hambro E, Azhar F, others · 2023
Earlier work this paper cites.
Llama 2: Open foundation and fine-tuned chat models
Touvron H, Martin L, Stone K, Albert P, Almahairi A, Babaei Y, Bashlykov N, Batra S, Bhargava P, Bhosale S, others · 2023
Earlier work this paper cites.
Toolllm: Facilitating large language models to master 16000+ real-world apis
Qin Y, Liang S, Ye Y, Zhu K, Yan L, Lu Y, Lin Y, Cong X, Tang X, Qian B, others · 2023
Earlier work this paper cites.
Zhu X, Chen Y, Tian H, Tao C, Su W, Yang C, Huang G, Li B, Lu L, Wang X, others · 2023
Earlier work this paper cites.
Minding language models’(lack of) theory of mind: A plug-and-play multi-character belief tracker
Sclar M, Kumar S, West P, Suhr A, Choi Y, Tsvetkov Y · 2023
Earlier work this paper cites.
Communicative agents for software development
Qian C, Cong X, Yang C, Chen W, Su Y, Xu J, Liu Z, Sun M · 2023
Earlier work this paper cites.
Agentverse
al. e C · 2023
Earlier work this paper cites.
Generative agents: Interactive simulacra of human behavior
Park J S, O’Brien J, Cai C J, Morris M R, Liang P, Bernstein M S · 2023
Earlier work this paper cites.
Recagent: A novel simulation paradigm for recommender systems
Wang L, Zhang J, Chen X, Lin Y, Song R, Zhao W X, Wen J R · 2023
Earlier work this paper cites.
Building cooperative embodied agents modularly with large language models
Zhang H, Du W, Shan J, Zhou Q, Du Y, Tenenbaum J B, Shu T, Gan C · 2023
Earlier work this paper cites.
Metagpt: Meta programming for multi-agent collaborative framework
Hong S, Zheng X, Chen J, Cheng Y, Wang J, Zhang C, Wang Z, Yau S K S, Lin Z, Zhou L, others · 2023
Earlier work this paper cites.
Self-collaboration code generation via chatgpt
Dong Y, Jiang X, Jin Z, Li G · 2023
Earlier work this paper cites.
Personality traits in large language models
Safdari M, Serapio-García G, Crepy C, Fitz S, Romero P, Sun L, Abdulhai M, Faust A, Matarić M · 2023
Earlier work this paper cites.
Toxicity in chatgpt: Analyzing persona-assigned language models
Deshpande A, Murahari V, Rajpurohit T, Kalyan A, Narasimhan K · 2023
Earlier work this paper cites.
Out of one, many: Using language models to simulate human samples
Argyle L P, Busby E C, Fulda N, Gubler J R, Rytting C, Wingate D · 2023
Earlier work this paper cites.
Reflective linguistic programming (rlp): A stepping stone in socially-aware agi (socialagi)
Fischer K A · 2023
Earlier work this paper cites.
Sayplan: Grounding large language models using 3d scene graphs for scalable robot task planning
Rana K, Haviland J, Garg S, Abou-Chakra J, Reid I, Suenderhauf N · 2023
Earlier work this paper cites.
Calypso: Llms as dungeon master’s assistants
Zhu A, Martin L, Head A, Callison-Burch C · 2023
Earlier work this paper cites.
Wang Z, Cai S, Chen G, Liu A, Ma X, Liang Y · 2023
Earlier work this paper cites.
Agentsims: An open-source sandbox for large language model evaluation
Lin J, Zhao H, Zhang A, Wu Y, Ping H, Chen Q · 2023
Earlier work this paper cites.
Liang X, Wang B, Huang H, Wu S, Wu P, Lu L, Ma Z, Li Z · 2023
Earlier work this paper cites.
Simplyretrieve: A private and lightweight retrieval-centric generative ai tool
Ng Y, Miyashita D, Hoshi Y, Morioka Y, Torii O, Kodama T, Deguchi J · 2023
Earlier work this paper cites.
Memory sandbox: Transparent and interactive memory management for conversational agents
Huang Z, Gutierrez S, Kamana H, MacNeil S · 2023
Earlier work this paper cites.
Voyager: An open-ended embodied agent with large language models
Wang G, Xie Y, Jiang Y, Mandlekar A, Xiao C, Zhu Y, Fan L, Anandkumar A · 2023
Earlier work this paper cites.
Memorybank: Enhancing large language models with long-term memory
Zhong W, Guo L, Gao Q, Wang Y · 2023
Earlier work this paper cites.
Chatdb: Augmenting llms with databases as their symbolic memory
Hu C, Fu J, Du C, Luo S, Zhao J, Zhao H · 2023
Earlier work this paper cites.
Ret-llm: Towards a general read-write memory for large language models
Modarressi A, Imani A, Fayyaz M, Schütze H · 2023
Earlier work this paper cites.
Memory augmented large language models are computationally universal
Schuurmans D · 2023
Earlier work this paper cites.
Expel: Llm agents are experiential learners
Zhao A, Huang D, Xu Q, Lin M, Liu Y J, Huang G · 2023
Earlier work this paper cites.
Rewoo: Decoupling reasoning from observations for efficient augmented language models
Xu B, Peng Z, Lei B, Mukherjee S, Liu Y, Xu D · 2023
Earlier work this paper cites.
Recmind: Large language model powered agent for recommendation
Wang Y, Jiang Z, Chen Z, Yang F, Zhou Y, Cho E, Fan X, Huang X, Lu Y, Yang Y · 2023
Earlier work this paper cites.
Graph of thoughts: Solving elaborate problems with large language models
Besta M, Blach N, Kubicek A, Gerstenberger R, Gianinazzi L, Gajda J, Lehmann T, Podstawski M, Niewiadomski H, Nyczyk P, others · 2023
Earlier work this paper cites.
Algorithm of thoughts: Enhancing exploration of ideas in large language models
Sel B, Al-Tawaha A, Khattar V, Wang L, Jia R, Jin M · 2023
Cited alongside, same era.
Generating executable action plans with environmentally-aware language models
Gramopadhye M, Szafir D · 2023
Cited alongside, same era.
Reasoning with language model is planning with world model
Hao S, Gu Y, Ma H, Hong J J, Wang Z, Wang D Z, Hu Z · 2023
Cited alongside, same era.
LLM+P: Empowering large language models with optimal planning proficiency
Liu B, Jiang Y, Zhang X, Liu Q, Zhang S, Biswas J, Stone P · 2023
Cited alongside, same era.
Dagan G, Keller F, Lascarides A · 2023
GPT engineer
al. e A O · 2023
Closest in time.
Xia Y, Shenoy M, Jazdi N, Weyrich M · 2023
Closest in time.
Ogundare O, Madasu S, Wiggins N · 2023
Closest in time.
Proagent: Building proactive cooperative ai with large language models
Zhang C, Yang K, Hu S, Wang Z, Li G, Sun Y, Zhang C, Zhang Z, Liu A, Zhu S C, others · 2023
Closest in time.
Enabling intelligent interactions between an agent and an llm: A reinforcement learning approach
Hu B, Zhao C, others · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
React: Synergizing reasoning and acting in language models
Yao S, Zhao J, Yu D, Du N, Shafran I, Narasimhan K, Cao Y · 2023
Cited alongside, same era.
Llm-planner: Few-shot grounded planning for embodied agents with large language models
Song C H, Wu J, Washington C, Sadler B M, Chao W L, Su Y · 2023
Cited alongside, same era.
Selfcheck: Using llms to zero-shot check their own step-by-step reasoning
Miao N, Teh Y W, Rainforth T · 2023
Cited alongside, same era.
Interact: Exploring the potentials of chatgpt as a cooperative agent
Chen P L, Chang C S · 2023
Cited alongside, same era.
Chen Z, Zhou K, Zhang B, Gong Z, Zhao W X, Wen J R · 2023
Cited alongside, same era.
TPTU: Task planning and tool usage of large language model-based AI agents
Ruan J, Chen Y, Zhang B, Xu Z, Bao T, Du G, Shi S, Mao H, Zeng X, Zhao R · 2023
Cited alongside, same era.
Gorilla: Large language model connected with massive apis
Patil S G, Zhang T, Wang X, Gonzalez J E · 2023
Cited alongside, same era.
Closest in time.
Plan, eliminate, and track–language models are good teachers for embodied agents
Wu Y, Min S Y, Bisk Y, Salakhutdinov R, Azaria A, Li Y, Mitchell T, Prabhumoye S · 2023
Closest in time.
Towards a unified agent with foundation models
Di Palo N, Byravan A, Hasenclever L, Wulfmeier M, Heess N, Riedmiller M · 2023
Closest in time.
Tidybot: Personalized robot assistance with large language models
Wu J, Antonova R, Kan A, Lepert M, Zeng A, Song S, Bohg J, Rusinkiewicz S, Funkhouser T · 2023
Closest in time.
Embodied task planning with large language models
Wu Z, Wang Z, Xu X, Lu J, Yan H · 2023
Closest in time.
Collaborating with language models for embodied reasoning
Dasgupta I, Kaeser-Chen C, Marino K, Ahuja A, Babayan S, Hill F, Fergus R · 2023
Closest in time.
Do embodied agents dream of pixelated sheep?: Embodied decision making using language guided world modelling
Nottingham K, Ammanabrolu P, Suhr A, Choi Y, Hajishirzi H, Singh S, Fox R · 2023
Closest in time.
Dialogue shaping: Empowering agents through npc interaction
Zhou W, Peng X, Riedl M · 2023
Closest in time.
The hitchhiker’s guide to program analysis: A journey with large language models
Li H, Hao Y, Zhai Y, Qian Z · 2023
Closest in time.
AgentGPT
al. e R · 2023
Closest in time.
Ai-legion
al. e E · 2023
Closest in time.
langchain
Chase H · 2023
Closest in time.
GPT-researcher
al. e A E · 2023
Closest in time.
Tool learning with foundation models
Qin Y, Hu S, Lin Y, Chen W, Ding N, Cui G, Zeng Z, Huang Y, Xiao C, Han C, others · 2023
Closest in time.
transformers-agent
Face H · 2023
Closest in time.
Superagi
al. e T · 2023
Closest in time.
Autogen: Enabling next-gen llm applications via multi-agent conversation framework
Wu Q, Bansal G, Zhang J, Wu Y, Zhang S, Zhu E, Li B, Jiang L, Zhang X, Wang C · 2023
Closest in time.
Agentverse: Facilitating multi-agent collaboration and exploring emergent behaviors in agents
Chen W, Su Y, Zuo J, Yang C, Yuan C, Qian C, Chan C M, Qin Y, Lu Y, Xie R, others · 2023
Closest in time.
Chateval: Towards better llm-based evaluators through multi-agent debate
Chan C M, Chen W, Su Y, Yu J, Xue W, Zhang S, Fu J, Liu Z · 2023
Closest in time.
Large language models are few-shot testers: Exploring llm-based general bug reproduction
Kang S, Yoon J, Yoo S · 2023
Closest in time.
Chatgpt and software testing education: Promises & perils
Jalil S, Rafi S, LaToza T D, Moran K, Lam W · 2023
Closest in time.
Mehta N, Teruel M, Sanz P F, Deng X, Awadallah A H, Kiseleva J · 2023
Closest in time.
Two failures of self-consistency in the multi-step reasoning of llms
Chen A, Phang J, Parrish A, Padmakumar V, Zhao C, Bowman S R, Cho K · 2023
Closest in time.
Choi M, Pei J, Kumar S, Shu C, Jurgens D · 2023
Closest in time.
Mobile-env: An evaluation platform and benchmark for interactive agents in llm era
Zhang D, Chen L, Zhao Z, Cao R, Yu K · 2023
Closest in time.
clembench: Using game play to evaluate chat-optimized language models as conversational agents
Chalamalasetti K, Götze J, Hakimov S, Madureira B, Sadler P, Schlangen D · 2023
Closest in time.
Decision-oriented dialogue for human-ai collaboration
Lin J, Tomlin N, Andreas J, Eisner J · 2023
Closest in time.
Towards autonomous testing agents via conversational large language models
Feldt R, Kang S, Yoon J, Yoo S · 2023
Closest in time.
Liang Y, Zhu L, Yang Y · 2023
Closest in time.
Agentbench: Evaluating llms as agents
Liu X, Yu H, Zhang H, Xu Y, Lei X, Lai H, Gu Y, Ding H, Men K, Yang K, others · 2023
Closest in time.
Bolaa: Benchmarking and orchestrating llm-augmented autonomous agents
Liu Z, Yao W, Zhang J, Xue L, Heinecke S, Murthy R, Feng Y, Chen Z, Niebles J C, Arpit D, others · 2023
Closest in time.
Gentopia. ai: A collaborative platform for tool-augmented llms
Xu B, Liu X, Shen H, Han Z, Li Y, Yue M, Peng Z, Liu Y, Yao Z, Xu D · 2023
Closest in time.
Emotionally numb or empathetic? evaluating how llms feel using emotionbench
Huang J t, Lam M H, Li E J, Ren S, Wang W, Jiao W, Tu Z, Lyu M R · 2023
Closest in time.
Webarena: A realistic web environment for building autonomous agents
Zhou S, Xu F F, Zhu H, Zhou X, Lo R, Sridhar A, Cheng X, Bisk Y, Fried D, Alon U, others · 2023
Closest in time.
Benchmarking llm powered chatbots: methods and metrics
Banerjee D, Singh P, Avadhanam A, Srivastava S · 2023
Closest in time.
A survey of large language models
Zhao W X, Zhou K, Li J, Tang T, Wang X, Hou Y, Min Y, Zhang B, Zhang J, Dong Z, others · 2023
Closest in time.
Harnessing the power of llms in practice: A survey on chatgpt and beyond
Yang J, Jin H, Tang R, Han X, Feng Q, Jiang H, Zhong S, Yin B, Hu X · 2023
Closest in time.
Aligning large language models with human: A survey
Wang Y, Zhong W, Li L, Mi F, Zeng X, Huang W, Shang L, Jiang X, Liu Q · 2023
Closest in time.
Augmented language models: a survey
Mialon G, Dessì R, Lomeli M, Nalmpantis C, Pasunuru R, Raileanu R, Rozière B, Schick T, Dwivedi-Yu J, Celikyilmaz A, others · 2023
Closest in time.
A survey on evaluation of large language models
Chang Y, Wang X, Wang J, Wu Y, Yang L, Zhu K, Chen H, Yi X, Wang C, Wang Y, others · 2023
Closest in time.
Emotionprompt: Leveraging psychology for large language models enhancement via emotional stimulus
Li C, Wang J, Zhu K, Zhang Y, Hou W, Lian J, Xie X · 2023
Closest in time.
Zhuo T Y, Li Z, Huang Y, Shiri F, Wang W, Haffari G, Li Y F · 2023
Closest in time.
On the robustness of dialogue history representation in conversational question answering: a comprehensive study and a new prompt-based method
Gekhman Z, Oved N, Keller O, Szpektor I, Reichart R · 2023
Closest in time.
Survey of hallucination in natural language generation
Ji Z, Lee N, Frieske R, Yu T, Su D, Xu Y, Ishii E, Bang Y J, Madotto A, Fung P · 2023
Closest in time.
Reflexion: Language agents with verbal reinforcement learning
Shinn N, Cassano F, Gopinath A, Narasimhan K, Yao S · 2024
Closest in time.
Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face
Shen Y, Song K, Tan X, Li D, Lu W, Zhuang Y · 2024
Closest in time.
Toolformer: Language models can teach themselves to use tools
Schick T, Dwivedi-Yu J, Dessì R, Raileanu R, Lomeli M, Hambro E, Zettlemoyer L, Cancedda N, Scialom T · 2024
Closest in time.
Swiftsage: A generative agent with fast and slow thinking for complex interactive tasks
Lin B Y, Fu Y, Yang K, Brahman F, Huang S, Bhagavatula C, Ammanabrolu P, Choi Y, Ren X · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Yao S, Yu D, Zhao J, Shafran I, Griffiths T, Cao Y, Narasimhan K · 2024
Closest in time.
Self-refine: Iterative refinement with self-feedback
Madaan A, Tandon N, Gupta P, Hallinan S, Gao L, Wiegreffe S, Alon U, Dziri N, Prabhumoye S, Yang Y, others · 2024
Closest in time.
Taskmatrix. ai: Completing tasks by connecting foundation models with millions of apis
Liang Y, Wu C, Song T, Wu W, Xia Y, Liu Y, Ou Y, Lu S, Ji L, Mao S, others · 2024
Closest in time.
Openagi: When llm meets domain experts
Ge Y, Hua W, Mei K, Tan J, Xu S, Li Z, Zhang Y, others · 2024
Closest in time.
Mind2web: Towards a generalist agent for the web
Deng X, Gu Y, Zheng B, Chen S, Stevens S, Wang B, Sun H, Su Y · 2024
Closest in time.
Large language models are semi-parametric reinforcement learning agents
Zhang D, Chen L, Zhang S, Xu H, Zhao Z, Yu K · 2024
Closest in time.
Language model behavior: A comprehensive survey
Chang T A, Bergen B K · 2024
Closest in time.