Fetching the paper…
Reading the bibliography…
For a long time, humanity has pursued artificial intelligence (AI) equivalent to or surpassing the human level, with AI agents considered a promising vehicle for this pursuit.
Diderot’s early philosophical works
Diderot, D · 1911
Earlier work this paper cites.
The inn of tranquillity: studies and essays
Galsworthy, J · 1912
Earlier work this paper cites.
The wealth of nations [1776] , vol. 11937
Smith, A · 1937
Earlier work this paper cites.
Three laws of robotics
Asimov, I · 1941
Earlier work this paper cites.
Steps toward artificial intelligence
Minsky, M · 1961
Earlier work this paper cites.
Receptive fields, binocular interaction and functional architecture in the cat’s visual cortex
Hubel, D. H., T. N. Wiesel · 1962
Earlier work this paper cites.
Actions, reasons, and causes
Davidson, D · 1963
Earlier work this paper cites.
Reasoning about a rule
Wason, P. C · 1968
Earlier work this paper cites.
STRIPS: A new approach to the application of theorem proving to problem solving
Fikes, R., N. J. Nilsson · 1971
Earlier work this paper cites.
I. agency
— · 1971
Earlier work this paper cites.
Psychology of reasoning: Structure and content , vol. 86
Wason, P. C., P. N. Johnson-Laird · 1972
Earlier work this paper cites.
Planning in a hierarchy of abstraction spaces
Sacerdoti, E. D · 1973
Earlier work this paper cites.
Becoming modern: Individual change in six developing countries
Inkeles, A., D. H. Smith · 1974
Earlier work this paper cites.
The nonlinear nature of plans
Sacerdoti, E. D · 1975
Earlier work this paper cites.
Computer science as empirical inquiry: Symbols and search
Newell, A., H. A. Simon · 1976
Earlier work this paper cites.
‘dynamics of growth in a finite world’ – comprehensive sensitivity analysis
Vermeulen, P., D. de Jongh · 1976
Earlier work this paper cites.
Ascribing mental qualities to machines
McCarthy, J · 1979
Earlier work this paper cites.
Learning and reasoning by analogy
Winston, P. H · 1980
Earlier work this paper cites.
STEAMER: an interactive inspectable simulation-based training system
Hollan, J. D., E. L. Hutchins, L. Weitzman · 1984
Earlier work this paper cites.
Actors: a Model of Concurrent Computation in Distributed Systems (Parallel Processing, Semantics, Open, Programming Languages, Artificial Intelligence)
Agha, G. A · 1985
Earlier work this paper cites.
The synthesis of digital machines with provable epistemic properties
Rosenschein, S. J., L. P. Kaelbling · 1986
Earlier work this paper cites.
An intelligent system for document retrieval in distributed office environments
Mukhopadhyay, U., L. M. Stephens, M. N. Huhns, et al · 1986
Earlier work this paper cites.
A robust layered control system for a mobile robot
Brooks, R · 1986
Earlier work this paper cites.
A robust layered control system for a mobile robot
Brooks, R. A · 1986
Earlier work this paper cites.
Mechanisms of memory
Squire, L. R · 1986
Earlier work this paper cites.
An architecture for intelligent reactive systems
Kaelbling, L. P., et al · 1987
Earlier work this paper cites.
Universal plans for reactive robots in unpredictable environments
Schoppers, M · 1987
Earlier work this paper cites.
Précis of the intentional stance
Dennett, D. C · 1988
Earlier work this paper cites.
Practical planning - extending the classical AI planning paradigm
Wilkins, D. E · 1988
Earlier work this paper cites.
Plans and resource-bounded practical reasoning
Bratman, M. E., D. J. Israel, M. E. Pollack · 1988
Earlier work this paper cites.
Society of mind
Minsky, M · 1988
Earlier work this paper cites.
Learning from delayed rewards, 1989
Watkins, C. J. C. H · 1989
Earlier work this paper cites.
Information extraction and text summarization using linguistic knowledge acquisition
Rau, L. F., P. S. Jacobs, U. Zernik · 1989
Earlier work this paper cites.
Approaches to studying formal and everyday reasoning
Galotti, K. M · 1989
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
McCloskey, M., N. J. Cohen · 1989
Earlier work this paper cites.
A new approach to the economic analysis of nonstationary time series and the business cycle
Hamilton, J. D · 1989
Earlier work this paper cites.
Designing autonomous agents: Theory and practice from biology to engineering and back
Maes, P · 1990
Earlier work this paper cites.
Situated agents can have goals
Maes, P · 1990
Earlier work this paper cites.
Action and agency in cognitive science
Shardlow, N · 1990
Earlier work this paper cites.
Intelligence without representation
Brooks, R. A · 1991
Earlier work this paper cites.
Do the right thing: studies in limited rationality
Russell, S. J., E. Wefald · 1991
Earlier work this paper cites.
Psychobiology of personality , vol. 10
Zuckerman, M · 1991
Earlier work this paper cites.
Agent oriented programming
Shoham, Y · 1992
Earlier work this paper cites.
Toward agent programs with circuit semantics
Nilsson, N. J · 1992
Earlier work this paper cites.
Essentials of Artificial Intelligence
Ginsberg, M. L · 1993
Earlier work this paper cites.
System dynamics and the lessons of 35 years
Forrester, J. W · 1993
Earlier work this paper cites.
Software agents
Genesereth, M. R., S. P. Ketchpel · 1994
Earlier work this paper cites.
Enabling agents to work together
Guha, R. V., D. B. Lenat · 1994
Earlier work this paper cites.
Modelling interacting agents in dynamic environments
Müller, J. P., M. Pischel · 1994
Earlier work this paper cites.
On-line Q-learning using connectionist systems , vol. 37
Rummery, G. A., M. Niranjan · 1994
Earlier work this paper cites.
Guarantees for autonomy in cognitive agent architecture
Castelfranchi, C · 1994
Earlier work this paper cites.
KQML as an agent communication language
Finin, T. W., R. Fritzson, D. P. McKay, et al · 1994
Earlier work this paper cites.
The role of emotion in believable agents
Bates, J · 1994
Earlier work this paper cites.
Intelligent agents: theory and practice
Wooldridge, M. J., N. R. Jennings · 1995
Earlier work this paper cites.
Formalizing properties of agents
Goodwin, R · 1995
Earlier work this paper cites.
Temporal difference learning and td-gammon
Tesauro, G., et al · 1995
Earlier work this paper cites.
Artificial life meets entertainment: Lifelike autonomous agents
Maes, P · 1995
Earlier work this paper cites.
Intelligent agents for interactive simulation environments
Tambe, M., W. L. Johnson, R. M. Jones, et al · 1995
Earlier work this paper cites.
Reinforcement learning: A survey
Kaelbling, L. P., M. L. Littman, A. W. Moore · 1996
Earlier work this paper cites.
Visual object recognition
Logothetis, N. K., D. L. Sheinberg · 1996
Earlier work this paper cites.
Progress in astronautics and aeronautics: Global positioning system: Theory and applications , vol. 164
Parkinson, B. W., J. J. Spilker · 1996
Earlier work this paper cites.
Social Science Microsimulation [Dagstuhl Seminar, May, 1995] . Springer, 1996
Troitzsch, K. G., U. Mueller, G. N. Gilbert, et al., eds · 1996
Earlier work this paper cites.
The media equation - how people treat computers, television, and new media like real people and places
Reeves, B., C. Nass · 1996
Earlier work this paper cites.
Software agents: A review
Green, S., L. Hurst, B. Nangle, et al · 1997
Earlier work this paper cites.
Intention
Anscombe, G. E. M · 2000
Earlier work this paper cites.
A theory of universal artificial intelligence based on algorithmic complexity
Hutter, M · 2000
Earlier work this paper cites.
A social reinforcement learning agent
Isbell, C., C. R. Shelton, M. Kearns, et al · 2001
Earlier work this paper cites.
Reinforcement learning agents
Ribeiro, C · 2002
Earlier work this paper cites.
How children learn the meanings of words
Bloom, P · 2002
Earlier work this paper cites.
Deep blue
Campbell, M., A. J. Hoane, F. hsiung Hsu · 2002
Earlier work this paper cites.
Artificial intelligence - a modern approach, 2nd Edition
Russell, S., P. Norvig · 2003
Earlier work this paper cites.
Time series forecasting using a hybrid ARIMA and neural network model
Zhang, G. P · 2003
Earlier work this paper cites.
Universal artificial intelligence: Sequential decisions based on algorithmic probability
Hutter, M · 2004
Earlier work this paper cites.
Planning and the brain
Grafman, J., L. Spector, M. J. Rattermann · 2004
Earlier work this paper cites.
Developing intelligent agent systems: A practical guide
Padgham, L., M. Winikoff · 2005
Earlier work this paper cites.
Constructing a language: A usage-based theory of language acquisition
Tomasello, M · 2005
Earlier work this paper cites.
Embodied sentence comprehension
Zwaan, R. A., C. J. Madden · 2005
Earlier work this paper cites.
The development of embodied cognition: Six lessons from babies
Smith, L., M. Gasser · 2005
Earlier work this paper cites.
Planning and problem solving: from neuropsychology to functional neuroimaging
Unterrainer, J. M., A. M. Owen · 2006
Earlier work this paper cites.
Decomposition of planning problems
Sebastia, L., E. Onaindia, E. Marzal · 2006
Earlier work this paper cites.
What is language: some preliminary remarks
Searle, J. R · 2007
Earlier work this paper cites.
Extending cognitive architecture with episodic memory
Nuxoll, A. M., J. E. Laird · 2007
Earlier work this paper cites.
Integrative literature review: Human capital planning: A review of literature and implications for human resource development
Zula, K. J., T. J. Chermack · 2007
Earlier work this paper cites.
Theory of Games and Economic Behavior (60th-Anniversary Edition)
von Neumann, J., O. Morgenstern · 2007
Earlier work this paper cites.
Innateness and culture in the evolution of language
Kirby, S., M. Dowman, T. L. Griffiths · 2007
Earlier work this paper cites.
Software as a service: An integration perspective
Sun, W., K. Zhang, S.-K. Chen, et al · 2007
Earlier work this paper cites.
Delivering software as a service
Dubey, A., D. Wagle · 2007
Earlier work this paper cites.
Developing software online with platform-as-a-service technology
Lawton, G · 2008
Earlier work this paper cites.
Computing machinery and intelligence
Turing, A. M · 2009
Earlier work this paper cites.
Defining agency: Individuality, normativity, asymmetry, and spatio-temporality in action
Barandiaran, X. E., E. Di Paolo, M. Rohde · 2009
Earlier work this paper cites.
Transfer learning for reinforcement learning domains: A survey
Taylor, M. E., P. Stone · 2009
Earlier work this paper cites.
Reference resolution challenges for intelligent agents: The need for knowledge
McShane, M · 2009
Earlier work this paper cites.
Curriculum learning
Bengio, Y., J. Louradour, R. Collobert, et al · 2009
Earlier work this paper cites.
Artificial intelligence a modern approach
Russell, S. J · 2010
Earlier work this paper cites.
Mapping the world in 3d
Schwarz, B · 2010
Earlier work this paper cites.
An introduction to multi-agent systems
Balaji, P. G., D. Srinivasan · 2010
Earlier work this paper cites.
Multiagent systems: algorithmic, game-theoretic, and logical foundations by y. shoham and k. leyton-brown cambridge university press, 2008
Aziz, H · 2010
Earlier work this paper cites.
Cellular automata models for the simulation of real-world urban processes: A review and analysis
Santé, I., A. M. García, D. Miranda, et al · 2010
Earlier work this paper cites.
Cloud computing: A study of infrastructure as a service (iaas)
Bhardwaj, S., L. Jain, S. Jain · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Collobert, R., J. Weston, L. Bottou, et al · 2011
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
Tellex, S., T. Kollar, S. Dickerson, et al · 2011
Earlier work this paper cites.
Interactional feedback and the impact of attitude and motivation on noticing l2 form
Bassiri, M. A · 2011
Earlier work this paper cites.
Approaching the symbol grounding problem with probabilistic graphical models
Tellex, S., T. Kollar, S. Dickerson, et al · 2011
Earlier work this paper cites.
The nist definition of cloud computing, 2011
Mell, P., T. Grance, et al · 2011
Earlier work this paper cites.
Learning to parse natural language commands to a robot control system
Matuszek, C., E. Herbst, L. Zettlemoyer, et al · 2012
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Mnih, V., K. Kavukcuoglu, D. Silver, et al · 2013
Earlier work this paper cites.
Beyond conflict monitoring: Cognitive control and the neural basis of thinking before you act
Brown, J. W · 2013
Earlier work this paper cites.
Discoveries in the human brain: neuroscience prehistory, brain structure, and function
Marshall, L. H., H. W. Magoun · 2013
Earlier work this paper cites.
Automated agent decomposition for classical planning
Crosby, M., M. Rovatsos, R. Petrick · 2013
Earlier work this paper cites.
Speech analysis synthesis and perception , vol. 3
Flanagan, J. L · 2013
Earlier work this paper cites.
Metacognition and the Use of Tools , pages 187–195
Clarebout, G., J. Elen, N. A. J. Collazo, et al · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Mnih, V., K. Kavukcuoglu, D. Silver, et al · 2013
Earlier work this paper cites.
PrismarineJS, 2013
2013
Earlier work this paper cites.
Complex systems and society: modeling and simulation , vol. 2
Bellomo, N., G. A. Marsan, A. Tosin · 2013
Earlier work this paper cites.
Reconsolidation of human memory: brain mechanisms and clinical relevance
Schwabe, L., K. Nader, J. C. Pruessner · 2014
Earlier work this paper cites.
Intriguing properties of neural networks
Szegedy, C., W. Zaremba, I. Sutskever, et al · 2014
Earlier work this paper cites.
Policy transfer using reward shaping
Brys, T., A. Harutyunyan, M. E. Taylor, et al · 2015
Earlier work this paper cites.
Actor-mimic: Deep multitask and transfer reinforcement learning
Parisotto, E., J. L. Ba, R. Salakhutdinov · 2015
Earlier work this paper cites.
Vinyals, O., Q. V. Le · 2015
Earlier work this paper cites.
Readings in planning theory
Fainstein, S. S., J. DeFilippis · 2015
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Goodfellow, I. J., J. Shlens, C. Szegedy · 2015
Earlier work this paper cites.
Everything as a service (xaas) on the cloud: Origins, current and future trends
Duan, Y., G. Fu, N. Zhou, et al · 2015
Earlier work this paper cites.
Infrastructure as a service and cloud technologies
Serrano, N., G. Gallardo, J. Hernantes · 2015
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., A. Huang, C. J. Maddison, et al · 2016
Earlier work this paper cites.
Rl$ˆ2$: Fast reinforcement learning via slow reinforcement learning
Duan, Y., J. Schulman, X. Chen, et al · 2016
Earlier work this paper cites.
Learning distributed representations of sentences from unlabelled data
Hill, F., K. Cho, A. Korhonen · 2016
Earlier work this paper cites.
Improved automatic keyword extraction given more semantic knowledge
Yang, K., Z. Chen, Y. Cai, et al · 2016
Earlier work this paper cites.
Generative deep neural networks for dialogue: A short review
Serban, I. V., R. Lowe, L. Charlin, et al · 2016
Earlier work this paper cites.
Squad: 100, 000+ questions for machine comprehension of text
Rajpurkar, P., J. Zhang, K. Lopyrev, et al · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., A. Huang, C. J. Maddison, et al · 2016
Earlier work this paper cites.
Dialog-based language learning
Weston, J · 2016
Earlier work this paper cites.
Accessorize to a crime: Real and stealthy attacks on state-of-the-art face recognition
Sharif, M., S. Bhagavatula, L. Bauer, et al · 2016
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
Bolukbasi, T., K. Chang, J. Y. Zou, et al · 2016
Earlier work this paper cites.
Learning to generate reviews and discovering sentiment
Radford, A., R. Józefowicz, I. Sutskever · 2017
Earlier work this paper cites.
Deep reinforcement learning: An overview
Li, Y · 2017
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., P. Abbeel, S. Levine · 2017
Earlier work this paper cites.
Commonsense knowledge in machine intelligence
Tandon, N., A. S. Varde, G. de Melo · 2017
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Christiano, P. F., J. Leike, T. B. Brown, et al · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
Kirkpatrick, J., R. Pascanu, N. Rabinowitz, et al · 2017
Earlier work this paper cites.
Learning without forgetting
Li, Z., D. Hoiem · 2017
Earlier work this paper cites.
Gradient episodic memory for continual learning
Lopez-Paz, D., M. Ranzato · 2017
Earlier work this paper cites.
Neural discrete representation learning
van den Oord, A., O. Vinyals, K. Kavukcuoglu · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., N. Shazeer, N. Parmar, et al · 2017
Earlier work this paper cites.
Imitation learning: A survey of learning methods
Hussein, A., M. M. Gaber, E. Elyan, et al · 2017
Earlier work this paper cites.
Mastering the game of go without human knowledge
Silver, D., J. Schrittwieser, K. Simonyan, et al · 2017
Earlier work this paper cites.
Deal or no deal? end-to-end learning of negotiation dialogues
Lewis, M., D. Yarats, Y. N. Dauphin, et al · 2017
Earlier work this paper cites.
Dialogue learning with human-in-the-loop
Li, J., A. H. Miller, S. Chopra, et al · 2017
Earlier work this paper cites.
Learning a neural semantic parser from user feedback
Iyer, S., I. Konstas, A. Cheung, et al · 2017
Earlier work this paper cites.
What might get in the way: Barriers to the use of apps for depression
Stiles-Shields, C., E. Montague, E. G. Lattie, et al · 2017
Earlier work this paper cites.
Deepstack: Expert-level artificial intelligence in no-limit poker
Moravcík, M., M. Schmid, N. Burch, et al · 2017
Earlier work this paper cites.
Open multi-agent systems: Gossiping with random arrivals and departures
Hendrickx, J. M., S. Martin · 2017
Earlier work this paper cites.
Simulation modelling for sustainability: a review of the literature
Moon, Y. B · 2017
Earlier work this paper cites.
Robust adversarial reinforcement learning
Pinto, L., J. Davidson, R. Sukthankar, et al · 2017
Earlier work this paper cites.
Badnets: Identifying vulnerabilities in the machine learning model supply chain
Gu, T., B. Dolan-Gavitt, S. Garg · 2017
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Caliskan, A., J. J. Bryson, A. Narayanan · 2017
Earlier work this paper cites.
A survey of artificial general intelligence projects for ethics, risk, and policy
Baum, S · 2017
Earlier work this paper cites.
Reinforcement learning: An introduction
Sutton, R. S., A. G. Barto · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Radford, A., K. Narasimhan, T. Salimans, et al · 2018
Earlier work this paper cites.
Intelligence without reason
Brooks, R. A · 2018
Earlier work this paper cites.
Generalization and regularization in DQN
Farebrother, J., M. C. Machado, M. Bowling · 2018
Earlier work this paper cites.
A study on overfitting in deep reinforcement learning
Zhang, C., O. Vinyals, R. Munos, et al · 2018
Earlier work this paper cites.
Illuminating generalization in deep reinforcement learning through procedural level generation
Justesen, N., R. R. Torrado, P. Bontrager, et al · 2018
Earlier work this paper cites.
Meta-reinforcement learning of structured exploration strategies
Gupta, A., R. Mendonca, Y. Liu, et al · 2018
Earlier work this paper cites.
Vanschoren, J · 2018
Earlier work this paper cites.
Importance weighted transfer of samples in reinforcement learning
Tirinzoni, A., A. Sessa, M. Pirotta, et al · 2018
Earlier work this paper cites.
Measuring catastrophic forgetting in neural networks
Kemker, R., M. McClure, A. Abitino, et al · 2018
Earlier work this paper cites.
Learning from richer human guidance: Augmenting comparison-based learning with feature queries
Basu, C., M. Singhal, A. D. Dragan · 2018
Earlier work this paper cites.
Deep contextualized word representations
Peters, M. E., M. Neumann, M. Iyyer, et al · 2018
Earlier work this paper cites.
Overcoming catastrophic forgetting with hard attention to the task
Serrà, J., D. Surís, M. Miron, et al · 2018
Earlier work this paper cites.
Imitation from observation: Learning to imitate behaviors from raw video via context translation
Liu, Y., A. Gupta, P. Abbeel, et al · 2018
Earlier work this paper cites.
Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection
Levine, S., P. Pastor, A. Krizhevsky, et al · 2018
Earlier work this paper cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation
Kalashnikov, D., A. Irpan, P. Pastor, et al · 2018
Earlier work this paper cites.
Virtualhome: Simulating household activities via programs
Puig, X., K. Ra, M. Boben, et al · 2018
Earlier work this paper cites.
Irving, G., P. F. Christiano, D. Amodei · 2018
Earlier work this paper cites.
Gated-attention architectures for task-oriented language grounding
Chaplot, D. S., K. M. Sathyendra, R. K. Pasumarthi, et al · 2018
Earlier work this paper cites.
Can neural machine translation be improved with user feedback?
Kreutzer, J., S. Khadivi, E. Matusov, et al · 2018
Earlier work this paper cites.
Dialsql: Dialogue based structured query generation
Gur, I., S. Yavuz, Y. Su, et al · 2018
Earlier work this paper cites.
Sentigan: Generating sentimental texts via mixture adversarial networks
Wang, K., X. Wan · 2018
Earlier work this paper cites.
Mojitalk: Generating emotional responses at scale
Zhou, X., W. Y. Wang · 2018
Earlier work this paper cites.
Should machines express sympathy and empathy? experiments with a health advice chatbot
Liu, B., S. S. Sundar · 2018
Earlier work this paper cites.
Textworld: A learning environment for text-based games
Côté, M., Á. Kádár, X. Yuan, et al · 2018
Earlier work this paper cites.
Personalizing dialogue agents: I have a dog, do you have pets too?
Zhang, S., E. Dinan, J. Urbanek, et al · 2018
Earlier work this paper cites.
Multi-agent systems: A survey
Dorri, A., S. S. Kanhere, R. Jurdak · 2018
Earlier work this paper cites.
Simulating Societies: The Computer Simulation of Social Phenomena
Gilbert, N., J. Doran · 2018
Earlier work this paper cites.
Assessing the role of social media and digital technology in violence reporting
Roberts, T., G. Marchais · 2018
Earlier work this paper cites.
Ethical challenges in data-driven dialogue systems
Henderson, P., K. Sinha, N. Angelard-Gontier, et al · 2018
Earlier work this paper cites.
Riemannian walk for incremental learning: Understanding forgetting and intransigence
Chaudhry, A., P. K. Dokania, T. Ajanthan, et al · 2018
Earlier work this paper cites.
Towards deep learning models resistant to adversarial attacks
Madry, A., A. Makelov, L. Schmidt, et al · 2018
Earlier work this paper cites.
Audio adversarial examples: Targeted attacks on speech-to-text
Carlini, N., D. A. Wagner · 2018
Earlier work this paper cites.
The malicious use of artificial intelligence: Forecasting, prevention, and mitigation
Brundage, M., S. Avin, J. Clark, et al · 2018
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford, A., J. Wu, R. Child, et al · 2019
Earlier work this paper cites.
Efficient off-policy meta-reinforcement learning via probabilistic context variables
Rakelly, K., A. Zhou, C. Finn, et al · 2019
Earlier work this paper cites.
Fakoor, R., P. Chaudhari, S. Soatto, et al · 2019
Earlier work this paper cites.
A structural probe for finding syntax in word representations
Hewitt, J., C. D. Manning · 2019
Earlier work this paper cites.
Do massively pretrained language models make better storytellers?
See, A., A. Pappu, R. Saxena, et al · 2019
Earlier work this paper cites.
Better language models and their implications
Radford, A., J. Wu, D. Amodei, et al · 2019
Earlier work this paper cites.
HIBERT: document level pre-training of hierarchical bidirectional transformers for document summarization
Zhang, X., F. Wei, M. Zhou · 2019
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Devlin, J., M. Chang, K. Lee, et al · 2019
Earlier work this paper cites.
Episodic memory in lifelong language learning
de Masson D’Autume, C., S. Ruder, L. Kong, et al · 2019
Earlier work this paper cites.
Experience replay for continual learning
Rolnick, D., A. Ahuja, J. Schwarz, et al · 2019
Earlier work this paper cites.
Attention on attention for image captioning
Huang, L., W. Wang, J. Chen, et al · 2019
Earlier work this paper cites.
M 2 {}^{\mbox{2}} : Meshed-memory transformer for image captioning
Cornia, M., M. Stefanini, L. Baraldi, et al · 2019
Earlier work this paper cites.
Fastspeech: Fast, robust and controllable text to speech
Ren, Y., Y. Ruan, X. Tan, et al · 2019
Earlier work this paper cites.
Review of deep reinforcement learning for robot manipulation
Nguyen, H., H. M. La · 2019
Earlier work this paper cites.
Reinforcement learning for market making in a multi-agent dealer market
Ganesh, S., N. Vadori, M. Xu, et al · 2019
Earlier work this paper cites.
Habitat: A platform for embodied AI research
Savva, M., J. Malik, D. Parikh, et al · 2019
Earlier work this paper cites.
HRL4IN: hierarchical reinforcement learning for interactive navigation with mobile manipulators
Li, C., F. Xia, R. Martín-Martín, et al · 2019
Earlier work this paper cites.
Model-based interactive semantic parsing: A unified framework and A text-to-sql case study
Yao, Z., Y. Su, H. Sun, et al · 2019
Earlier work this paper cites.
Improving natural language interaction with robots using advice
Mehta, N., D. Goldwasser · 2019
Earlier work this paper cites.
Learning from dialogue after deployment: Feed yourself, chatbot!
Hancock, B., A. Bordes, P. Mazaré, et al · 2019
Earlier work this paper cites.
Caire: An empathetic neural chatbot
Lin, Z., P. Xu, G. I. Winata, et al · 2019
Earlier work this paper cites.
Moel: Mixture of empathetic listeners
Lin, Z., A. Madotto, J. Shin, et al · 2019
Earlier work this paper cites.
On the utility of learning about humans for human-ai coordination
Carroll, M., R. Shah, M. K. Ho, et al · 2019
Earlier work this paper cites.
Persuasion for good: Towards a personalized persuasive dialogue system for social good
Wang, X., W. Shi, R. Kim, et al · 2019
Earlier work this paper cites.
Learning to speak and act in a fantasy text adventure game
Urbanek, J., A. Fan, S. Karamcheti, et al · 2019
Earlier work this paper cites.
A Variational Basis for the Regulation and Structuration Mechanisms of Agent Societies
da Rocha Costa, A. C · 2019
Earlier work this paper cites.
" if you catch my drift…": ability to infer implied meaning is distinct from vocabulary and grammar skills
Wilson, A. C., D. V. Bishop · 2019
Earlier work this paper cites.
Learning a unified classifier incrementally via rebalancing
Hou, S., X. Pan, C. C. Loy, et al · 2019
Earlier work this paper cites.
Benchmarking neural network robustness to common corruptions and perturbations
Hendrycks, D., T. G. Dietterich · 2019
Earlier work this paper cites.
Textbugger: Generating adversarial text against real-world applications
Li, J., S. Ji, T. Du, et al · 2019
Earlier work this paper cites.
Experimental security research of tesla autopilot
Lab, T. K. S · 2019
Earlier work this paper cites.
Generating natural language adversarial examples through probability weighted word saliency
Ren, S., Y. Deng, K. He, et al · 2019
Earlier work this paper cites.
Robustness may be at odds with accuracy
Tsipras, D., S. Santurkar, L. Engstrom, et al · 2019
Cited alongside, same era.
Theoretically principled trade-off between robustness and accuracy
Zhang, H., Y. Yu, J. Jiao, et al · 2019
Cited alongside, same era.
Experience grounds language
Bisk, Y., A. Holtzman, J. Thomason, et al · 2020
Cited alongside, same era.
Language models are few-shot learners
Brown, T. B., B. Mann, N. Ryder, et al · 2020
Cited alongside, same era.
Transfer learning in deep reinforcement learning: A survey
Zhu, Z., K. Lin, J. Zhou · 2020
Cited alongside, same era.
Scaling laws for neural language models
Kaplan, J., S. McCandlish, T. Henighan, et al · 2020
Palm-e: An embodied multimodal language model
Driess, D., F. Xia, M. S. M. Sajjadi, et al · 2023
Closest in time.
Embodiedgpt: Vision-language pre-training via embodied chain of thought
Mu, Y., Q. Zhang, M. Hu, et al · 2023
Closest in time.
Think before you act: Decision transformers with internal working memory
Kang, J., R. Laroche, X. Yuan, et al · 2023
Closest in time.
Valmeekam, K., S. Sreedharan, M. Marquez, et al · 2023
Closest in time.
LLM+P: empowering large language models with optimal planning proficiency
Liu, B., Y. Jiang, X. Zhang, et al · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
How much knowledge can you pack into the parameters of a language model?
Roberts, A., C. Raffel, N. Shazeer · 2020
Cited alongside, same era.
Probing pretrained language models for lexical semantics
Vulic, I., E. M. Ponti, R. Litschko, et al · 2020
Cited alongside, same era.
How can we know what language models know
Jiang, Z., F. F. Xu, J. Araki, et al · 2020
Cited alongside, same era.
BART: denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Lewis, M., Y. Liu, N. Goyal, et al · 2020
Cited alongside, same era.
Towards a human-like open-domain chatbot
Adiwardana, D., M. Luong, D. R. So, et al · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., N. Shazeer, A. Roberts, et al · 2020
Cited alongside, same era.
Closest in time.
Chain of hindsight aligns language models with feedback
Liu, H., C. Sferrazza, P. Abbeel · 2023
Closest in time.
Lin, Y., Y. Chen · 2023
Closest in time.
Improving language model negotiation with self-play and in-context learning from AI feedback
Fu, Y., H. Peng, T. Khot, et al · 2023
Closest in time.
Building cooperative embodied agents modularly with large language models
Zhang, H., W. Du, J. Shan, et al · 2023
Closest in time.
Bang, Y., S. Cahyawijaya, N. Lee, et al · 2023
Closest in time.
Is chatgpt a highly fluent grammatical error correction system? A comprehensive evaluation
Fang, T., S. Yang, K. Lan, et al · 2023
Closest in time.
Bounding the capabilities of large language models in open text generation with prompt constraints
Lu, A., H. Zhang, Y. Zhang, et al · 2023
Closest in time.
Clever hans or neural theory of mind? stress testing social reasoning in large language models
Shapira, N., M. Levy, S. H. Alavi, et al · 2023
Closest in time.
Large language models in medicine
Thirunavukarasu, A. J., D. S. J. Ting, K. Elangovan, et al · 2023
Closest in time.
DS-1000: A natural and reliable benchmark for data science code generation
Lai, Y., C. Li, Y. Wang, et al · 2023
Closest in time.
Editing large language models: Problems, methods, and opportunities
Yao, Y., P. Wang, B. Tian, et al · 2023
Closest in time.
Selfcheckgpt: Zero-resource black-box hallucination detection for generative large language models
Manakul, P., A. Liusie, M. J. F. Gales · 2023
Closest in time.
Self-checker: Plug-and-play modules for fact-checking with large language models
Li, M., B. Peng, Z. Zhang · 2023
Closest in time.
CRITIC: large language models can self-correct with tool-interactive critiquing
Gou, Z., Z. Shao, Y. Gong, et al · 2023
Closest in time.
Colt5: Faster long-range transformers with conditional computation
Ainslie, J., T. Lei, M. de Jong, et al · 2023
Closest in time.
Randomized positional encodings boost length generalization of transformers
Ruoss, A., G. Delétang, T. Genewein, et al · 2023
Closest in time.
Liang, X., B. Wang, H. Huang, et al · 2023
Closest in time.
Reflexion: an autonomous agent with dynamic memory and self-reflection
Shinn, N., B. Labash, A. Gopinath · 2023
Closest in time.
Memorybank: Enhancing large language models with long-term memory
Zhong, W., L. Guo, Q. Gao, et al · 2023
Closest in time.
Chateval: Towards better llm-based evaluators through multi-agent debate
Chan, C., W. Chen, Y. Su, et al · 2023
Closest in time.
Zhu, X., Y. Chen, H. Tian, et al · 2023
Closest in time.
RET-LLM: towards a general read-write memory for large language models
Modarressi, A., A. Imani, M. Fayyaz, et al · 2023
Closest in time.
Agentsims: An open-source sandbox for large language model evaluation
Lin, J., H. Zhao, A. Zhang, et al · 2023
Closest in time.
Chatdb: Augmenting llms with databases as their symbolic memory
Hu, C., J. Fu, C. Du, et al · 2023
Closest in time.
Memory sandbox: Transparent and interactive memory management for conversational agents
Huang, Z., S. Gutierrez, H. Kamana, et al · 2023
Closest in time.
Selection-inference: Exploiting large language models for interpretable logical reasoning
Creswell, A., M. Shanahan, I. Higgins · 2023
Closest in time.
Self-refine: Iterative refinement with self-feedback
Madaan, A., N. Tandon, P. Gupta, et al · 2023
Closest in time.
Hugginggpt: Solving AI tasks with chatgpt and its friends in huggingface
Shen, Y., K. Song, X. Tan, et al · 2023
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Yao, S., D. Yu, J. Zhao, et al · 2023
Closest in time.
Plan, eliminate, and track - language models are good teachers for embodied agents
Wu, Y., S. Y. Min, Y. Bisk, et al · 2023
Closest in time.
Wang, Z., S. Cai, A. Liu, et al · 2023
Closest in time.
Reasoning with language model is planning with world model
Hao, S., Y. Gu, H. Ma, et al · 2023
Closest in time.
Swiftsage: A generative agent with fast and slow thinking for complex interactive tasks
Lin, B. Y., Y. Fu, K. Yang, et al · 2023
Closest in time.
Chatcot: Tool-augmented chain-of-thought reasoning on chat-based large language models
Chen, Z., K. Zhou, B. Zhang, et al · 2023
Closest in time.
Voyager: An open-ended embodied agent with large language models
Wang, G., Y. Xie, Y. Jiang, et al · 2023
Closest in time.
Chat with the environment: Interactive multimodal perception using large language models
Zhao, X., M. Li, C. Weber, et al · 2023
Closest in time.
Selfcheck: Using llms to zero-shot check their own step-by-step reasoning
Miao, N., Y. W. Teh, T. Rainforth · 2023
Closest in time.
Images speak in images: A generalist painter for in-context visual learning
Wang, X., W. Wang, Y. Cao, et al · 2023
Closest in time.
Neural codec language models are zero-shot text to speech synthesizers
Wang, C., S. Chen, Y. Wu, et al · 2023
Closest in time.
A survey for in-context learning
Dong, Q., L. Li, D. Dai, et al · 2023
Closest in time.
A comprehensive survey of continual learning: Theory, method and application
Wang, L., X. Zhang, H. Su, et al · 2023
Closest in time.
Progressive prompts: Continual learning for language models
Razdaibiedina, A., Y. Mao, R. Hou, et al · 2023
Closest in time.
Llama: Open and efficient foundation language models
Touvron, H., T. Lavril, G. Izacard, et al · 2023
Closest in time.
Falcon-40b: an open large language model with state-of-the-art performance, 2023
Almazrouei, E., H. Alobeidli, A. Alshamsi, et al · 2023
Closest in time.
Mindstorms in natural language-based societies of mind
Zhuge, M., H. Liu, F. Faccio, et al · 2023
Closest in time.
Stanford alpaca: An instruction-following llama model, 2023
Taori, R., I. Gulrajani, T. Zhang, et al · 2023
Closest in time.
Openagi: When LLM meets domain experts
Ge, Y., W. Hua, J. Ji, et al · 2023
Closest in time.
MEGA: multilingual evaluation of generative AI
Ahuja, K., R. Hada, M. Ochieng, et al · 2023
Closest in time.
Large language models can be easily distracted by irrelevant context
Shi, F., X. Chen, K. Misra, et al · 2023
Closest in time.
Siren’s song in the AI ocean: A survey on hallucination in large language models
Zhang, Y., Y. Li, L. Cui, et al · 2023
Closest in time.
Augmented language models: a survey
Mialon, G., R. Dessì, M. Lomeli, et al · 2023
Closest in time.
Investigating the factual knowledge boundary of large language models with retrieval augmentation
Ren, R., Y. Wang, Y. Qu, et al · 2023
Closest in time.
Landmark attention: Random-access infinite context length for transformers
Mohtashami, A., M. Jaggi · 2023
Closest in time.
Unlimiformer: Long-range transformers with unlimited length input
Bertsch, A., U. Alon, G. Neubig, et al · 2023
Closest in time.
Expel: LLM agents are experiential learners
Zhao, A., D. Huang, Q. Xu, et al · 2023
Closest in time.
Zhou, X., G. Li, Z. Liu · 2023
Closest in time.
Towards reasoning in large language models: A survey
Huang, J., K. C. Chang · 2023
Closest in time.
Towards revealing the mystery behind chain of thought: a theoretical perspective
Feng, G., B. Zhang, Y. Gu, et al · 2023
Closest in time.
Rewoo: Decoupling reasoning from observations for efficient augmented language models
Xu, B., Z. Peng, B. Lei, et al · 2023
Closest in time.
Faithful chain-of-thought reasoning
Lyu, Q., S. Havaldar, A. Stein, et al · 2023
Closest in time.
Dagan, G., F. Keller, A. Lascarides · 2023
Closest in time.
Sayplan: Grounding large language models using 3d scene graphs for scalable task planning
Rana, K., J. Haviland, S. Garg, et al · 2023
Closest in time.
Multilingual machine translation with large language models: Empirical results and analysis
Zhu, W., H. Liu, Q. Dong, et al · 2023
Closest in time.
Speak foreign languages with your own voice: Cross-lingual neural codec language modeling
Zhang, Z., L. Zhou, C. Wang, et al · 2023
Closest in time.
Bootstrap your own skills: Learning to solve new tasks with large language model guidance
Zhang, J., J. Zhang, K. Pertsch, et al · 2023
Closest in time.
Continual diffusion: Continual customization of text-to-image diffusion with c-lora
Smith, J. S., Y.-C. Hsu, L. Zhang, et al · 2023
Closest in time.
Language is not all you need: Aligning perception with language models
Huang, S., L. Dong, W. Wang, et al · 2023
Closest in time.
BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models
Li, J., D. Li, S. Savarese, et al · 2023
Closest in time.
Instructblip: Towards general-purpose vision-language models with instruction tuning
Dai, W., J. Li, D. Li, et al · 2023
Closest in time.
Multimodal-gpt: A vision and language model for dialogue with humans
Gong, T., C. Lyu, S. Zhang, et al · 2023
Closest in time.
Pandagpt: One model to instruction-follow them all
Su, Y., T. Lan, H. Li, et al · 2023
Closest in time.
Liu, H., C. Li, Q. Wu, et al · 2023
Closest in time.
Audiogpt: Understanding and generating speech, music, sound, and talking head
Huang, R., M. Li, D. Yang, et al · 2023
Closest in time.
Chen, F., M. Han, H. Zhao, et al · 2023
Closest in time.
Video-llama: An instruction-tuned audio-visual language model for video understanding
Zhang, H., X. Li, L. Bing · 2023
Closest in time.
Interngpt: Solving vision-centric tasks by interacting with chatbots beyond language
Liu, Z., Y. He, W. Wang, et al · 2023
Closest in time.
Videochat: Chat-centric video understanding
Li, K., Y. He, Y. Wang, et al · 2023
Closest in time.
Learning to model the world with language
Lin, J., Y. Du, O. Watkins, et al · 2023
Closest in time.
UNIFIED-IO: A unified model for vision, language, and multi-modal tasks
Lu, J., C. Clark, R. Zellers, et al · 2023
Closest in time.
Kosmos-2: Grounding multimodal large language models to the world
Peng, Z., W. Wang, L. Dong, et al · 2023
Closest in time.
Macaw-llm: Multi-modal language modeling with image, audio, video, and text integration
Lyu, C., M. Wu, L. Wang, et al · 2023
Closest in time.
Video-chatgpt: Towards detailed video understanding via large vision and language models
Maaz, M., H. A. Rasheed, S. H. Khan, et al · 2023
Closest in time.
Training-free layout control with cross-attention guidance
Chen, M., I. Laina, A. Vedaldi · 2023
Closest in time.
Robust speech recognition via large-scale weak supervision
Radford, A., J. W. Kim, T. Xu, et al · 2023
Closest in time.
Tf-gridnet: Integrating full- and sub-band modeling for speech separation
Wang, Z., S. Cornell, S. Choi, et al · 2023
Closest in time.
Visual chatgpt: Talking, drawing and editing with visual foundation models
Wu, C., S. Yin, W. Qi, et al · 2023
Closest in time.
Large language models as tool makers
Cai, T., X. Wang, T. Ma, et al · 2023
Closest in time.
Qian, C., C. Han, Y. R. Fung, et al · 2023
Closest in time.
Teaching large language models to self-debug
Chen, X., M. Lin, N. Schärli, et al · 2023
Closest in time.
Alphablock: Embodied finetuning for vision-language reasoning in robot manipulation
Jin, C., W. Tan, J. Yang, et al · 2023
Closest in time.
Navgpt: Explicit reasoning in vision-and-language navigation with large language models
Zhou, G., Y. Hong, Q. Wu · 2023
Closest in time.
Do embodied agents dream of pixelated sheep: Embodied decision making using language guided world modelling
Nottingham, K., P. Ammanabrolu, A. Suhr, et al · 2023
Closest in time.
Distilling internet-scale vision-language models into embodied agents
Sumers, T., K. Marino, A. Ahuja, et al · 2023
Closest in time.
Extracting training data from diffusion models
Carlini, N., J. Hayes, M. Nasr, et al · 2023
Closest in time.
Can GPT-4 support analysis of textual data in tasks requiring highly specialized domain expertise?
Savelka, J., K. D. Ashley, M. A. Gray, et al · 2023
Closest in time.
Domain specialization as the key to make large language models disruptive: A comprehensive survey, 2023
Ling, C., X. Zhao, J. Lu, et al · 2023
Closest in time.
Universal and transferable adversarial attacks on aligned language models
Zou, A., Z. Wang, J. Z. Kolter, et al · 2023
Closest in time.
Secrets of RLHF in large language models part I: PPO
Zheng, R., S. Dou, S. Gao, et al · 2023
Closest in time.
Unifying large language models and knowledge graphs: A roadmap
Pan, S., L. Luo, Y. Wang, et al · 2023
Closest in time.
Chemcrow: Augmenting large-language models with chemistry tools, 2023
Bran, A. M., S. Cox, A. D. White, et al · 2023
Closest in time.
TPTU: task planning and tool usage of large language model-based AI agents
Ruan, J., Y. Chen, B. Zhang, et al · 2023
Closest in time.
Industrial engineering with large language models: A case study of chatgpt’s performance on oil & gas problems, 2023
Ogundare, O., S. Madasu, N. Wiggins · 2023
Closest in time.
Collaborating with language models for embodied reasoning
Dasgupta, I., C. Kaeser-Chen, K. Marino, et al · 2023
Closest in time.
The flan collection: Designing data and methods for effective instruction tuning
Longpre, S., L. Hou, T. Vu, et al · 2023
Closest in time.
Code as policies: Language model programs for embodied control
Liang, J., W. Huang, F. Xia, et al · 2023
Closest in time.
Enabling intelligent interactions between an agent and an LLM: A reinforcement learning approach
Hu, B., C. Zhao, P. Zhang, et al · 2023
Closest in time.
Visual language maps for robot navigation
Huang, C., O. Mees, A. Zeng, et al · 2023
Closest in time.
Can an embodied agent find your "cat-shaped mug"? llm-based zero-shot object navigation
Dorbala, V. S., J. F. M. Jr., D. Manocha · 2023
Closest in time.
RT-2: vision-language-action models transfer web knowledge to robotic control
— · 2023
Closest in time.
A real-world webagent with planning, long context understanding, and program synthesis
Gur, I., H. Furuta, A. Huang, et al · 2023
Closest in time.
Mind2web: Towards a generalist agent for the web
Deng, X., Y. Gu, B. Zheng, et al · 2023
Closest in time.
Multimodal web navigation with instruction-finetuned foundation models
Furuta, H., O. Nachum, K. Lee, et al · 2023
Closest in time.
Webarena: A realistic web environment for building autonomous agents
Zhou, S., F. F. Xu, H. Zhu, et al · 2023
Closest in time.
Language models can solve computer tasks
Kim, G., P. Baldi, S. McAleer · 2023
Closest in time.
Synapse: Leveraging few-shot exemplars for human-level computer control
Zheng, L., R. Wang, B. An · 2023
Closest in time.
Interact: Exploring the potentials of chatgpt as a cooperative agent
Chen, P., C. Chang · 2023
Closest in time.
The hitchhiker’s guide to program analysis: A journey with large language models
Li, H., Y. Hao, Y. Zhai, et al · 2023
Closest in time.
Towards autonomous testing agents via conversational large language models
Feldt, R., S. Kang, J. Yoon, et al · 2023
Closest in time.
Chatmof: An autonomous AI system for predicting and generating metal-organic frameworks
Kang, Y., J. Kim · 2023
Closest in time.
Plan4mc: Skill reinforcement learning and planning for open-world minecraft tasks
Yuan, H., C. Zhang, H. Wang, et al · 2023
Closest in time.
Chatllm network: More brains, more intelligence
Hao, R., L. Hu, W. Qi, et al · 2023
Closest in time.
Roco: Dialectic multi-robot collaboration with large language models
Mandi, Z., S. Jain, S. Song · 2023
Closest in time.
Blind judgement: Agent-based supreme court modelling with GPT
Hamilton, S · 2023
Closest in time.
Metagpt: Meta programming for multi-agent collaborative framework
Hong, S., X. Zheng, J. Chen, et al · 2023
Closest in time.
Autogen: Enabling next-gen LLM applications via multi-agent conversation framework
Wu, Q., G. Bansal, J. Zhang, et al · 2023
Closest in time.
Proagent: Building proactive cooperative AI with large language models
Zhang, C., K. Yang, S. Hu, et al · 2023
Closest in time.
DERA: enhancing large language model completions with dialog-enabled resolving agents
Nair, V., E. Schumacher, G. J. Tso, et al · 2023
Closest in time.
Multi-agent collaboration: Harnessing the power of intelligent LLM agents
Talebirad, Y., A. Nadiri · 2023
Closest in time.
Agentverse: Facilitating multi-agent collaboration and exploring emergent behaviors in agents
Chen, W., Y. Su, J. Zuo, et al · 2023
Closest in time.
CGMI: configurable general multi-agent interaction framework
Shi, J., J. Zhao, Y. Wang, et al · 2023
Closest in time.
Examining the inter-consistency of large language models: An in-depth analysis via debate
Xiong, K., X. Ding, Y. Cao, et al · 2023
Closest in time.
Hey dona! can you help me with student course registration?
Kalvakurthi, V., A. S. Varde, J. Jenq · 2023
Closest in time.
Math agents: Computational infrastructure, mathematical embedding, and genomics
Swan, M., T. Kido, E. Roland, et al · 2023
Closest in time.
Helping the helper: Supporting peer counselors via ai-empowered practice and feedback
Hsu, S.-L., R. S. Shah, P. Senthil, et al · 2023
Closest in time.
Huatuogpt, towards taming language model to be a doctor
Zhang, H., J. Chen, F. Jiang, et al · 2023
Closest in time.
Yang, S., H. Zhao, S. Zhu, et al · 2023
Closest in time.
Multi-turn dialogue agent as sales’ assistant in telemarketing
Gao, W., X. Gao, Y. Tang · 2023
Closest in time.
PEER: A collaborative language model
Schick, T., J. A. Yu, Z. Jiang, et al · 2023
Closest in time.
Lu, B., N. Haduong, C. Lee, et al · 2023
Closest in time.
Assistgpt: A general multi-modal assistant that can plan, execute, inspect, and learn
Gao, D., L. Ji, L. Zhou, et al · 2023
Closest in time.
SAPIEN: affective virtual agents powered by large language models
Hasan, M., C. Özel, S. Potter, et al · 2023
Closest in time.
Mastering the game of no-press diplomacy via human-regularized reinforcement learning and planning
Bakhtin, A., D. J. Wu, A. Lerer, et al · 2023
Closest in time.
Decision-oriented dialogue for human-ai collaboration
Lin, J., N. Tomlin, J. Andreas, et al · 2023
Closest in time.
Quantifying the impact of large language models on collective opinion dynamics
Li, C., X. Su, C. Fan, et al · 2023
Closest in time.
Agent_GPT
Reworked · 2023
Closest in time.
GPT Engineer
AntonOsika · 2023
Closest in time.
Knowledge-enhanced agents for interactive text games
Chhikara, P., J. Zhang, F. Ilievski, et al · 2023
Closest in time.
Leandojo: Theorem proving with retrieval-augmented language models
Yang, K., A. M. Swope, A. Gu, et al · 2023
Closest in time.
Evolutionary-scale prediction of atomic-level protein structure with a language model
Lin, Z., H. Akin, R. Rao, et al · 2023
Closest in time.
Wang, Z., S. Mao, W. Wu, et al · 2023
Closest in time.
Chatgpt as your personal data scientist
Hassan, M. M., R. A. Knipper, S. K. K. Santu · 2023
Closest in time.
Digitization of healthcare sector: A study on privacy and security concerns
Paul, M., L. Maglaras, M. A. Ferrag, et al · 2023
Closest in time.
Training language models with language feedback at scale
Scheurer, J., J. A. Campos, T. Korbak, et al · 2023
Closest in time.
Learning new skills after deployment: Improving open-domain internet-driven dialogue with human feedback
Xu, J., M. Ung, M. Komeili, et al · 2023
Closest in time.
Human-in-the-loop through chain-of-thought
Cai, Z., B. Chang, W. Han · 2023
Closest in time.
Mehta, N., M. Teruel, P. F. Sanz, et al · 2023
Closest in time.
Exploring large language models for communication games: An empirical study on werewolf, 2023
Xu, Y., S. Wang, P. Li, et al · 2023
Closest in time.
Mind meets machine: Unravelling gpt-4’s cognitive psychology
Dhingra, S., M. Singh, V. S. B, et al · 2023
Closest in time.
Hagendorff, T · 2023
Closest in time.
Emotional intelligence of large language models
Wang, X., X. Li, Z. Yin, et al · 2023
Closest in time.
Computer says "no": The case against empathetic conversational AI
Curry, A., A. C. Curry · 2023
Closest in time.
Chatgpt outperforms humans in emotional awareness evaluations
Elyoseph, Z., D. Hadar-Shoval, K. Asraf, et al · 2023
Closest in time.
Empathetic AI for empowering resilience in games
Habibi, R., J. Pfau, J. Holmes, et al · 2023
Closest in time.
Do llms possess a personality? making the MBTI test an amazing evaluation for large language models
Pan, K., Y. Zeng · 2023
Closest in time.
Does gpt-3 demonstrate psychopathy? evaluating large language models from a psychological perspective, 2023
Li, X., Y. Li, S. Joty, et al · 2023
Closest in time.
Personality traits in large language models
Safdari, M., G. Serapio-García, C. Crepy, et al · 2023
Closest in time.
Hoodwinked: Deception and cooperation in a text-based game for language models
O’Gara, A · 2023
Closest in time.
Bharadhwaj, H., J. Vakil, M. Sharma, et al · 2023
Closest in time.
S 3 {}^{\mbox{3}} : Social-network simulation system with large language model-empowered agents
Gao, C., X. Lan, Z. Lu, et al · 2023
Closest in time.
Recagent: A novel simulation paradigm for recommender systems
Wang, L., J. Zhang, X. Chen, et al · 2023
Closest in time.
Epidemic modeling with generative agents
Williams, R., N. Hosseinichimeh, A. Majumdar, et al · 2023
Closest in time.
Heterogeneous value evaluation for large language models
Zhang, Z., N. Liu, S. Qi, et al · 2023
Closest in time.
Personhood and ai: Why large language models don’t understand us
Browning, J · 2023
Closest in time.
Theory of mind may have spontaneously emerged in large language models
Kosinski, M · 2023
Closest in time.
Inductive reasoning in humans and large language models
Han, S. J., K. Ransom, A. Perfors, et al · 2023
Closest in time.
Thinking fast and slow in large language models, 2023
Hagendorff, T., S. Fabi, M. Kosinski · 2023
Closest in time.
Hagendorff, T., S. Fabi · 2023
Closest in time.
Ma, Z., Y. Mei, Z. Su · 2023
Closest in time.
What, when, and how to ground: Designing user persona-aware conversational agents for engaging dialogue
Kwon, D. S., S. Lee, K. H. Kim, et al · 2023
Closest in time.
Ai and the transformation of social science research
Grossmann, I., M. Feinberg, D. C. Parker, et al · 2023
Closest in time.
Multi-party chat: Conversational agents in group settings with humans and models
Wei, J., K. Shuster, A. Szlam, et al · 2023
Closest in time.
Can large language models transform computational social science?
Ziems, C., W. Held, O. Shaikh, et al · 2023
Closest in time.
Playing the werewolf game with artificial intelligence for language understanding
Shibata, H., S. Miki, Y. Nakamura · 2023
Closest in time.
Exploring the intersection of large language models and agent-based modeling via prompt engineering
Junprung, E · 2023
Closest in time.
Investigating emergent goal-like behaviour in large language models using experimental economics
Phelps, S., Y. I. Russell · 2023
Closest in time.
Chatgpt and the AI act
Helberger, N., N. Diakopoulos · 2023
Closest in time.
Toxicity in chatgpt: Analyzing persona-assigned language models
Deshpande, A., V. Murahari, T. Rajpurohit, et al · 2023
Closest in time.
Large language models struggle to learn long-tail knowledge
Kandpal, N., H. Deng, A. Roberts, et al · 2023
Closest in time.
Should chatgpt be biased? challenges and risks of bias in large language models
Ferrara, E · 2023
Closest in time.
Opiniongpt: Modelling explicit biases in instruction-tuned llms, 2023
Haller, P., A. Aynetdinov, A. Akbik · 2023
Closest in time.
In-context impersonation reveals large language models’ strengths and biases
Salewski, L., S. Alaniz, I. Rio-Torto, et al · 2023
Closest in time.
Towards healthy AI: large language models need therapists too
Lin, B., D. Bouneffouf, G. A. Cecchi, et al · 2023
Closest in time.
Privacy and data protection in chatgpt and other ai chatbots: Strategies for securing user information
Sebastian, G · 2023
Closest in time.
A conversation with bing’s chatbot left me deeply unsettled, 2023
Roose, K · 2023
Closest in time.
Emergent world representations: Exploring a sequence model trained on a synthetic task
Li, K., A. K. Hopkins, D. Bau, et al · 2023
Closest in time.
Agentbench: Evaluating llms as agents
Liu, X., H. Yu, H. Zhang, et al · 2023
Closest in time.
Using large language models to simulate multiple humans and replicate human subject studies
Aher, G. V., R. I. Arriaga, A. T. Kalai · 2023
Closest in time.
Liang, Y., L. Zhu, Y. Yang · 2023
Closest in time.
Gentopia: A collaborative platform for tool-augmented llms
Xu, B., X. Liu, H. Shen, et al · 2023
Closest in time.
" help me help the ai": Understanding how explainability can support human-ai interaction
Kim, S. S., E. A. Watkins, O. Russakovsky, et al · 2023
Closest in time.
Choi, M., J. Pei, S. Kumar, et al · 2023
Closest in time.
Augmenting autotelic agents with large language models
Colas, C., L. Teodorescu, P. Oudeyer, et al · 2023
Closest in time.
Characterizing the impacts of instances on robustness
Zheng, R., Z. Xi, Q. Liu, et al · 2023
Closest in time.
Safety and ethical concerns of large language models
Zhiheng, X., Z. Rui, G. Tao · 2023
Closest in time.
Promptbench: Towards evaluating the robustness of large language models on adversarial prompts
Zhu, K., J. Wang, J. Zhou, et al · 2023
Closest in time.
How robust is GPT-3.5 to predecessors? A comprehensive study on language understanding tasks
Chen, X., J. Ye, C. Zu, et al · 2023
Closest in time.
Prompt injection attack against llm-integrated applications
Liu, Y., G. Deng, Y. Li, et al · 2023
Closest in time.
Huang, X., W. Ruan, W. Huang, et al · 2023
Closest in time.
A close look into the calibration of pre-trained language models
Chen, Y., L. Yuan, G. Cui, et al · 2023
Closest in time.
Survey of hallucination in natural language generation
Ji, Z., N. Lee, R. Frieske, et al · 2023
Closest in time.
Self-contradictory hallucinations of large language models: Evaluation, detection and mitigation
Mündler, N., J. He, S. Jenko, et al · 2023
Closest in time.
Varshney, N., W. Yao, H. Zhang, et al · 2023
Closest in time.
Lightman, H., V. Kosaraju, Y. Burda, et al · 2023
Closest in time.
Charan, P. V. S., H. Chunduri, P. M. Anand, et al · 2023
Closest in time.
Language agents in the digital world: Opportunities and risks
Yao, S., K. Narasimhan · 2023
Closest in time.
Is there any social principle for llm-based agents?
Bai, J., S. Zhang, Z. Chen · 2023
Closest in time.
Monotonic location attention for length generalization
Chowdhury, J. R., C. Caragea · 2023
Closest in time.