Fetching the paper…
Reading the bibliography…
This paper explores the integration of two AI subdisciplines employed in the development of artificial agents that exhibit intelligent behavior: Large Language Models (LLMs) and Cognitive Architectures (CAs).
Language Models are Few-Shot Learners
Brown, T.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J. D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; Agarwal, S.; Herbert-Voss, A.; Krueger, G.; Henighan, T.; Child, R.; Ramesh, A.; Ziegler, D.; Wu, J.; Winter, C.; Hesse, C.; Chen, M.; Sigler, E.; Litwin, M.; Gray, S.; Chess, B.; Clark, J.; Berner, C.; McCandlish, S.; Radford, A.; Sutskever, I.; and Amodei, D. 2020 · 1901
Earlier work this paper cites.
Neural text generation with unlikelihood training
Welleck, S.; Kulikov, I.; Roller, S.; Dinan, E.; Cho, K.; and Weston, J. 2019 · 1908
Earlier work this paper cites.
Society of mind
Minsky, M. 1988 · 1988
Earlier work this paper cites.
The Next Decade in AI: Four Steps Towards Robust Artificial Intelligence
Marcus, G. 2020 · 2002
Earlier work this paper cites.
The LIDA architecture: Adding new modes of learning to an intelligent, autonomous, software agent
Franklin, S.; and Patterson, F. 2006 · 2006
Earlier work this paper cites.
A cognitive architecture that combines internal simulation with a global workspace
Shanahan, M. 2006 · 2006
Earlier work this paper cites.
The current status of the simulation theory of cognition
Hesslow, G. 2012 · 2012
Earlier work this paper cites.
The atomic components of thought
Anderson, J. R.; and Lebiere, C. J. 2014 · 2014
Earlier work this paper cites.
Anatomy of the mind: exploring psychological mechanisms and processes with the Clarion cognitive architecture
Sun, R. 2016 · 2016
Earlier work this paper cites.
A Standard Model of the Mind: Toward a Common Computational Framework across Artificial Intelligence, Cognitive Science, Neuroscience, and Robotics
Laird, J. E.; Lebiere, C.; and Rosenbloom, P. S. 2017 · 2017
Earlier work this paper cites.
The knowledge level in cognitive architectures: Current limitations and possible developments
Lieto, A.; Lebiere, C.; and Oltramari, A. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Earlier work this paper cites.
The Soar cognitive architecture
Laird, J. E. 2019 · 2019
Earlier work this paper cites.
40 years of cognitive architectures: core cognitive abilities and practical applications
Kotseruba, I.; and Tsotsos, J. K. 2020 · 2020
Earlier work this paper cites.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?
Bender, E. M.; Gebru, T.; McMillan-Major, A.; and Shmitchell, S. 2021 · 2021
Cited alongside, same era.
A Task-Oriented Dialogue Architecture via Transformer Neural Language Models and Symbolic Injection
Romero, O. J.; Wang, A.; Zimmerman, J.; Steinfeld, A.; and Tomasic, A. 2021 · 2021
Cited alongside, same era.
Propositional Reasoning via Neural Transformer Language Models
Tomasic, A.; Romero, O. J.; Zimmerman, J.; and Steinfeld, A. 2021 · 2021
Cited alongside, same era.
A path towards autonomous machine intelligence version 0.9. 2, 2022-06-27
LeCun, Y. 2022 · 2022
Cited alongside, same era.
Pre-Trained Language Models for Interactive Decision-Making
Li, S.; Puig, X.; Paxton, C.; Du, Y.; Wang, C.; Fan, L.; Chen, T.; Huang, D.; Akyürek, E.; Anandkumar, A.; Andreas, J.; Mordatch, I.; Torralba, A.; and Zhu, Y. 2022 · 2022
Cited alongside, same era.
Caufield, J. H.; Hegde, H.; Emonet, V.; Harris, N. L.; Joachimiak, M. P.; Matentzoglu, N.; Kim, H.; Moxon, S. A.; Reese, J. T.; Haendel, M. A.; et al. 2023 · 2023
Closest in time.
Active Prompting with Chain-of-Thought for Large Language Models
Diao, S.; Wang, P.; Lin, Y.; and Zhang, T. 2023 · 2023
Closest in time.
Improving Factuality and Reasoning in Language Models through Multiagent Debate
Du, Y.; Li, S.; Torralba, A.; Tenenbaum, J. B.; and Mordatch, I. 2023 · 2023
Closest in time.
PAL: Program-aided Language Models
Gao, L.; Madaan, A.; Zhou, S.; Alon, U.; Liu, P.; Yang, Y.; Callan, J.; and Neubig, G. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Qian, J.; Wang, H.; Li, Z.; Li, S.; and Yan, X. 2022 · 2022
Cited alongside, same era.
Fine-tuned language models are continual learners
Scialom et al., T. 2022 · 2022
Cited alongside, same era.
A Study of Implicit Bias in Pretrained Language Models against People with Disabilities
Venkit, P. N.; Srinath, M.; and Wilson, S. 2022 · 2022
Cited alongside, same era.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Wei, J.; Wang, X.; Schuurmans, D.; Bosma, M.; Chi, E. H.; Le, Q.; and Zhou, D. 2022 · 2022
Cited alongside, same era.
Taxonomy of Risks Posed by Language Models
Weidinger, L.; Uesato, J.; Rauh, M.; Griffin, C.; Huang, P.-S.; Mellor, J.; Glaese, A.; Cheng, M.; Balle, B.; Kasirzadeh, A.; Biles, C.; Brown, S.; Kenton, Z.; Hawkins, W.; Stepleton, T.; Birhane, A.; Hendricks, L. A.; Rimell, L.; Isaac, W.; Haas, J.; Legassick, S.; Irving, G.; and Gabriel, I. 2022 · 2022
Cited alongside, same era.
Automatic Chain of Thought Prompting in Large Language Models
Zhang, Z.; Zhang, A.; Li, M.; and Smola, A. 2022 · 2022
Cited alongside, same era.
Unit 1: Understanding Production Systems
ACT-R Website. 2015 · 2023
Cited alongside, same era.
Huang, X.; Ruan, W.; Huang, W.; Jin, G.; Dong, Y.; Wu, C.; Bensalem, S.; Mu, R.; Qi, Y.; Zhao, X.; Cai, K.; Zhang, Y.; Wu, S.; Xu, P.; Wu, D.; Freitas, A.; and Mustafa, M. A. 2023 · 2023
Closest in time.
Theory of Mind May Have Spontaneously Emerged in Large Language Models
Kosinski, M. 2023 · 2023
Closest in time.
Augmented Language Models: a Survey
Mialon, G.; Dessì, R.; Lomeli, M.; Nalmpantis, C.; Pasunuru, R.; Raileanu, R.; Rozière, B.; Schick, T.; Dwivedi-Yu, J.; Celikyilmaz, A.; Grave, E.; LeCun, Y.; and Scialom, T. 2023 · 2023
Closest in time.
Generative Agents: Interactive Simulacra of Human Behavior
Park, J. S.; O’Brien, J. C.; Cai, C. J.; Morris, M. R.; Liang, P.; and Bernstein, M. S. 2023 · 2023
Closest in time.
Symbols and grounding in large language models
Pavlick, E. 2023 · 2023
Closest in time.
Toolformer: Language Models Can Teach Themselves to Use Tools
Schick, T.; Dwivedi-Yu, J.; Dessì, R.; Raileanu, R.; Lomeli, M.; Zettlemoyer, L.; Cancedda, N.; and Scialom, T. 2023 · 2023
Closest in time.
Voyager: An open-ended embodied agent with large language models
Wang, G.; Xie, Y.; Jiang, Y.; Mandlekar, A.; Xiao, C.; Zhu, Y.; Fan, L.; and Anandkumar, A. 2023 · 2023
Closest in time.
OlaGPT: Empowering LLMs With Human-like Problem-Solving Abilities
Xie, Y.; Xie, T.; Lin, M.; Wei, W.; Li, C.; Kong, B.; Chen, L.; Zhuo, C.; Hu, B.; and Li, Z. 2023 · 2023
Closest in time.
Minigpt-4: Enhancing vision-language understanding with advanced large language models
Zhu, D.; Chen, J.; Shen, X.; Li, X.; and Elhoseiny, M. 2023 · 2023
Closest in time.