Fetching the paper…
Reading the bibliography…
Due to the powerful capabilities demonstrated by large language model (LLM), there has been a recent surge in efforts to integrate them with AI agents to enhance their performance.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Dota 2 with large scale deep reinforcement learning
Christopher Berner and Brockman et al. 2019 · 1912
Earlier work this paper cites.
Memory: Facts and fallacies
Ian ML Hunter. 1957 · 1957
Earlier work this paper cites.
Programs with common sense
J McCarthy. 1959 · 1959
Earlier work this paper cites.
Episodic and semantic memory
Endel Tulving et al. 1972 · 1972
Earlier work this paper cites.
A universal modular actor formalism for artificial intelligence
Carl Hewitt, Peter Bishop, and Richard Steiger. 1973 · 1973
Earlier work this paper cites.
Working memory
Alan David Baddeley. 1983 · 1983
Earlier work this paper cites.
Elements of episodic memory
Endel Tulving. 1983 · 1983
Earlier work this paper cites.
Learning representations by back-propagating errors
David E Rumelhart, Geoffrey E Hinton, and Ronald J Williams. 1986 · 1986
Earlier work this paper cites.
The Society of Mind
Marvin L. Minsky. 1988 · 1988
Earlier work this paper cites.
A basic agent
Steven Vere and Timothy Bickmore. 1990 · 1990
Earlier work this paper cites.
Htn planning: complexity and expressivity
Kutluhan Erol, James Hendler, and Dana S Nau. 1994 · 1994
Earlier work this paper cites.
Human memory: Theory and practice
Alan D Baddeley. 1997 · 1997
Earlier work this paper cites.
Pddl— the planning domain definition language
Constructions Aeronautiques, Adele Howe, Craig Knoblock, ISI Drew McDermott, Ashwin Ram, Manuela Veloso, Daniel Weld, David Wilkins SRI, Anthony Barrett, Dave Christianson, et al. 1998 · 1998
Earlier work this paper cites.
Pddl2. 1: An extension to pddl for expressing temporal planning domains
Maria Fox and Derek Long. 2003 · 2003
Earlier work this paper cites.
Extending cognitive architecture with episodic memory
Andrew M Nuxoll and John E Laird. 2007 · 2007
Earlier work this paper cites.
Artificial Intelligence: A Modern Approach , 3 edition
Stuart Russell and Peter Norvig. 2010 · 2010
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller. 2013 · 2013
Earlier work this paper cites.
An emotion understanding framework for intelligent agents based on episodic and semantic memories
Mohammad Kazemifard, Nasser Ghasem-Aghaee, Bryan L Koenig, and Tuncer I Ören. 2014 · 2014
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al. 2016 · 2016
Earlier work this paper cites.
Deep reinforcement learning: A brief survey
Kai Arulkumaran, Marc Peter Deisenroth, Miles Brundage, and Anil Anthony Bharath. 2017 · 2017
Earlier work this paper cites.
The neuroanatomical, neurophysiological and psychological basis of memory: Current models and their origins
Eduardo Camina and Francisco Güell. 2017 · 2017
Earlier work this paper cites.
Deep reinforcement learning: An overview
Yuxi Li. 2017 · 2017
Earlier work this paper cites.
Petlon: planning efficiently for task-level-optimal navigation
Shih-Yun Lo, Shiqi Zhang, and Peter Stone. 2018 · 2018
Earlier work this paper cites.
Interleaving hierarchical task planning and motion constraint testing for dual-arm manipulation
Alejandro Suárez-Hernández, Guillem Alenyà, and Carme Torras. 2018 · 2018
Earlier work this paper cites.
Commonsense knowledge mining from pretrained models
Joe Davison, Joshua Feldman, and Alexander M Rush. 2019 · 2019
Earlier work this paper cites.
Task planning in robotics: an empirical comparison of pddl-and asp-based systems
Yu-qian Jiang, Shi-qi Zhang, Piyush Khandelwal, and Peter Stone. 2019 · 2019
Earlier work this paper cites.
Bioinspired electronics for artificial sensory systems
Yei Hwan Jung, Byeonghak Park, Jong Uk Kim, and Tae-il Kim. 2019 · 2019
Earlier work this paper cites.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Earlier work this paper cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al. 2019 · 2019
Earlier work this paper cites.
Meta reinforcement learning for sim-to-real domain adaptation
Karol Arndt, Murtaza Hazara, Ali Ghadirzadeh, and Ville Kyrki. 2020 · 2020
Cited alongside, same era.
Pretrained language model embryology: The birth of albert
Cheng-Han Chiang, Sung-Feng Huang, and Hung-Yi Lee. 2020 · 2020
Cited alongside, same era.
Artificial sensory memory
Changjin Wan, Pingqiang Cai, Ming Wang, Yan Qian, Wei Huang, and Xiaodong Chen. 2020 · 2020
Cited alongside, same era.
Optimal mixed discrete-continuous planning for linear hybrid systems
Jingkai Chen, Brian C Williams, and Chuchu Fan. 2021 · 2021
Cited alongside, same era.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, et al. 2021 · 2021
Cited alongside, same era.
Interact: Exploring the potentials of chatgpt as a cooperative agent
Po-Lin Chen and Cheng-Shang Chang. 2023 · 2023
Closest in time.
Palm-e: An embodied multimodal language model
Danny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch, Aakanksha Chowdhery, Brian Ichter, Ayzaan Wahid, Jonathan Tompson, Quan Vuong, Tianhe Yu, Wenlong Huang, Yevgen Chebotar, Pierre Sermanet, Daniel Duckworth, Sergey Levine, Vincent Vanhoucke, Karol Hausman, Marc Toussaint, Klaus Greff, Andy Zeng, Igor Mordatch, and Pete Florence. 2023 · 2023
Closest in time.
Lin Guan, Karthik Valmeekam, Sarath Sreedharan, and Subbarao Kambhampati. 2023 · 2023
Closest in time.
Recent trends in task and motion planning for robotics: A survey
Huihui Guo, Fan Wu, Yunchuan Qin, Ruihui Li, Keqin Li, and Kenli Li. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jeff Da, Ronan Le Bras, Ximing Lu, Yejin Choi, and Antoine Bosselut. 2021 · 2021
Cited alongside, same era.
Webgpt: Browser-assisted question-answering with human feedback
Reiichiro Nakano, Jacob Hilton, Suchir Balaji, Jeff Wu, Long Ouyang, Christina Kim, Christopher Hesse, Shantanu Jain, Vineet Kosaraju, William Saunders, et al. 2021 · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021 · 2021
Cited alongside, same era.
A primer in bertology: What we know about how bert works
Anna Rogers, Olga Kovaleva, and Anna Rumshisky. 2021 · 2021
Cited alongside, same era.
Relational world knowledge representation in contextual language models: A review
Tara Safavi and Danai Koutra. 2021 · 2021
Cited alongside, same era.
Clip-nav: Using clip for zero-shot vision-and-language navigation
Vishnu Sashank Dorbala, Gunnar Sigurdsson, Robinson Piramuthu, Jesse Thomason, and Gaurav S Sukhatme. 2022 · 2022
Cited alongside, same era.
Minedojo: Building open-ended embodied agents with internet-scale knowledge
Linxi Fan, Guanzhi Wang, Yunfan Jiang, Ajay Mandlekar, Yuncong Yang, Haoyi Zhu, Andrew Tang, De-An Huang, Yuke Zhu, and Anima Anandkumar. 2022 · 2022
Cited alongside, same era.
Bin Hu, Chenyang Zhao, Pu Zhang, Zihao Zhou, Yuanhang Yang, Zenglin Xu, and Bin Liu. 2023 · 2023
Closest in time.
Think before you act: Decision transformers with internal working memory
Jikun Kang, Romain Laroche, Xindi Yuan, Adam Trischler, Xue Liu, and Jie Fu. 2023 · 2023
Closest in time.
A machine with short-term, episodic, and semantic memory systems
Taewoon Kim, Michael Cochez, Vincent François-Lavet, Mark Neerincx, and Piek Vossen. 2023 · 2023
Closest in time.
Prompted llms as chatbot modules for long open-domain conversation
Gibbeum Lee, Volker Hartmann, Jongho Park, Dimitris Papailiopoulos, and Kangwook Lee. 2023 · 2023
Closest in time.
Adaptive and intelligent robot task planning for home service: A review
Haizhen Li and Xilun Ding. 2023 · 2023
Closest in time.
Large language model guided tree-of-thought
Jieyi Long. 2023 · 2023
Closest in time.
Chameleon: Plug-and-play compositional reasoning with large language models
Pan Lu, Baolin Peng, Hao Cheng, Michel Galley, Kai-Wei Chang, Ying Nian Wu, Song-Chun Zhu, and Jianfeng Gao. 2023 · 2023
Closest in time.
Video-chatgpt: Towards detailed video understanding via large vision and language models
Muhammad Maaz, Hanoona Rasheed, Salman Khan, and Fahad Shahbaz Khan. 2023 · 2023
Closest in time.
Empowering conversational agents using semantic in-context learning
Amin Omidvar and Aijun An. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Art: Automatic multi-step reasoning and tool-use for large language models
Bhargavi Paranjape, Scott Lundberg, Sameer Singh, Hannaneh Hajishirzi, Luke Zettlemoyer, and Marco Tulio Ribeiro. 2023 · 2023
Closest in time.
Generative agents: Interactive simulacra of human behavior
Joon Sung Park, Joseph C O’Brien, Carrie J Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein. 2023 · 2023
Closest in time.
Gorilla: Large language model connected with massive apis
Shishir G Patil, Tianjun Zhang, Xin Wang, and Joseph E Gonzalez. 2023 · 2023
Closest in time.
Baolin Peng, Michel Galley, Pengcheng He, Hao Cheng, Yujia Xie, Yu Hu, Qiuyuan Huang, Lars Liden, Zhou Yu, Weizhu Chen, et al. 2023 · 2023
Closest in time.
Toolformer: Language models can teach themselves to use tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom. 2023 · 2023
Closest in time.
Hugginggpt: Solving ai tasks with chatgpt and its friends in huggingface
Yongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li, Weiming Lu, and Yueting Zhuang. 2023 · 2023
Closest in time.
Reflexion: an autonomous agent with dynamic memory and self-reflection
Noah Shinn, Beck Labash, and Ashwin Gopinath. 2023 · 2023
Closest in time.
Progprompt: Generating situated robot task plans using large language models
Ishika Singh, Valts Blukis, Arsalan Mousavian, Ankit Goyal, Danfei Xu, Jonathan Tremblay, Dieter Fox, Jesse Thomason, and Animesh Garg. 2023 · 2023
Closest in time.
Adaplanner: Adaptive planning from feedback with language models
Haotian Sun, Yuchen Zhuang, Lingkai Kong, Bo Dai, and Chao Zhang. 2023 · 2023
Closest in time.
Graspgpt: Leveraging semantic knowledge from a large language model for task-oriented grasping
Chao Tang, Dehao Huang, Wenqi Ge, Weiyu Liu, and Hong Zhang. 2023 · 2023
Closest in time.
Karthik Valmeekam, Sarath Sreedharan, Matthew Marquez, Alberto Olmo, and Subbarao Kambhampati. 2023 · 2023
Closest in time.
Translating natural language to planning goals with large-language models
Yaqi Xie, Chen Yu, Tongyao Zhu, Jinbin Bai, Ze Gong, and Harold Soh. 2023 · 2023
Closest in time.
Gentopia: A collaborative platform for tool-augmented llms
Binfeng Xu, Xukun Liu, Hua Shen, Zeyu Han, Yuhan Li, Murong Yue, Zhiyuan Peng, Yuchen Liu, Ziyu Yao, and Dongkuan Xu. 2023 · 2023
Closest in time.
Memorybank: Enhancing large language models with long-term memory
Wanjun Zhong, Lianghong Guo, Qiqi Gao, and Yanlin Wang. 2023 · 2023
Closest in time.
Mindstorms in natural language-based societies of mind
Mingchen Zhuge and Haozhe Liu et al. 2023 · 2023
Closest in time.