Fetching the paper…
Reading the bibliography…
Game-theoretic scenarios have become pivotal in evaluating the social intelligence of Large Language Model (LLM)-based social agents.
The measurement of social intelligence
Thelma Hunt · 1928
Earlier work this paper cites.
Prisoner’s dilemma: A study in conflict and cooperation , volume 165
Anatol Rapoport and Albert M Chammah · 1965
Earlier work this paper cites.
Generalization learning techniques for automating the learning of heuristics
Donald Arthur Waterman · 1970
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff · 1978
Earlier work this paper cites.
A further search for social intelligence
Martin E Ford and Marie S Tisak · 1983
Earlier work this paper cites.
Scorable games: A better way to teach negotiation
Lawrence E Susskind · 1985
Earlier work this paper cites.
The winner’s curse and public information in common value auctions
John H Kagel and Dan Levin · 1986
Earlier work this paper cites.
Children’s understanding of representational change and its relation to the understanding of false belief and the appearance-reality distinction
Alison Gopnik and Janet W Astington · 1988
Earlier work this paper cites.
Social intelligence and decoding of nonverbal cues
Michael L Barnes and Robert J Sternberg · 1989
Earlier work this paper cites.
The importance of the agenda in bargaining
Chaim Fershtman · 1990
Earlier work this paper cites.
Game theory
Drew Fudenberg and Jean Tirole · 1991
Earlier work this paper cites.
The belief-desire-intention model of agency
Michael Georgeff, Barney Pell, Martha Pollack, Milind Tambe, and Michael Wooldridge · 1999
Earlier work this paper cites.
Negotiation
Max H Bazerman, Jared R Curhan, Don A Moore, and Kathleen L Valley · 2000
Earlier work this paper cites.
Social intelligence
John F Kihlstrom and Nancy Cantor · 2000
Earlier work this paper cites.
A logic for strategic reasoning
Wiebe Van Der Hoek, Wojciech Jamroga, and Michael Wooldridge · 2005
Earlier work this paper cites.
Emoticons and social interaction on the internet: the importance of social context
Daantje Derks, Arjan ER Bos, and Jasper Von Grumbkow · 2007
Earlier work this paper cites.
Behavioral game theory: Experiments in strategic interaction
Colin F Camerer · 2011
Earlier work this paper cites.
Game theory
Guillermo Owen · 2013
Earlier work this paper cites.
Diplomacy
Henry Kissinger · 2014
Earlier work this paper cites.
Deepstack: Expert-level artificial intelligence in heads-up no-limit poker
Matej Moravčík, Martin Schmid, Neil Burch, Viliam Lisỳ, Dustin Morrill, Nolan Bard, Trevor Davis, Kevin Waugh, Michael Johanson, and Michael Bowling · 2017
Earlier work this paper cites.
The importance of modeling social factors of language: Theory and practice
Dirk Hovy and Diyi Yang · 2021
Earlier work this paper cites.
Using large language models to simulate multiple humans and replicate human subject studies
Gati Aher, RosaI. Arriaga, and Adam Tauman Kalai · 2022
Earlier work this paper cites.
Human-level play in the game of diplomacy by combining language models with strategic reasoning
Anton Bakhtin, Noam Brown, Emily Dinan, Gabriele Farina, Colin Flaherty, Daniel Fried, Andrew Goff, Jonathan Gray, Hengyuan Hu, Athul Paul Jacob, Mojtaba Komeili, Karthik Konath, Minae Kwon, Adam Lerer, Mike Lewis, Alexander H. Miller, Sandra Mitts, Adithya Renduchintala, Stephen Roller, Dirk Rowe, Weiyan Shi, Joe Spisak, Alexander Wei, David J. Wu, Hugh Zhang, and Markus Zijlstra · 2022
Earlier work this paper cites.
Werewolf among us: A multimodal dataset for modeling persuasion behaviors in social deduction games
Bolin Lai, Hongxin Zhang, Miao Liu, Aryan Pariani, Fiona Ryan, Wenqi Jia, Shirley Anugrah Hayati, James M Rehg, and Diyi Yang · 2022
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Earlier work this paper cites.
Llm-deliberation: Evaluating llms with interactive multi-agent negotiation games
Sahar Abdelnabi, Amr Gomaa, Sarath Sivaprasad, Lea Schönherr, and Mario Fritz · 2023
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Earlier work this paper cites.
Playing repeated games with large language models
Elif Akata, Lion Schulz, Julian Coda-Forno, Seong Joon Oh, Matthias Bethge, and Eric Schulz · 2023
Earlier work this paper cites.
Playing games with gpt: What can we learn about a large language model from canonical strategic games?
Philip Brookins and Jason Matthew DeBacker · 2023
Earlier work this paper cites.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al · 2023
Earlier work this paper cites.
Jiangjie Chen, Siyu Yuan, Rong Ye, Bodhisattwa Prasad Majumder, and Kyle Richardson · 2023
Earlier work this paper cites.
Opencompass: A universal evaluation platform for foundation models
OpenCompass Contributors · 2023
Earlier work this paper cites.
Can large language models serve as rational players in game theory? a systematic analysis
Caoyun Fan, Jindou Chen, Yaohui Jin, and Hao He · 2023
Earlier work this paper cites.
Improving language model negotiation with self-play and in-context learning from ai feedback
Yao Fu, Hao Peng, Tushar Khot, and Mirella Lapata · 2023
Earlier work this paper cites.
Strategic reasoning with language models
Kanishk Gandhi, Dorsa Sadigh, and Noah D Goodman · 2023
Cited alongside, same era.
Gpt in game theory experiments
Fulin Guo · 2023
Cited alongside, same era.
Suspicion-agent: Playing imperfect information games with theory of mind aware gpt-4
Jiaxian Guo, Bo Yang, Paul Yoo, Bill Yuchen Lin, Yusuke Iwasawa, and Yutaka Matsuo · 2023
Cited alongside, same era.
Are chatgpt and gpt-4 good poker players?–a pre-flop analysis
Akshat Gupta · 2023
Cited alongside, same era.
Large language models as simulated economic agents: What can we learn from homo silicus?
John J. Horton · 2023
Understanding social reasoning in language models with language models
Kanishk Gandhi, Jan-Philipp Fränken, Tobias Gerstenberg, and Noah Goodman · 2024
Closest in time.
Richelieu: Self-evolving llm-based agents for ai diplomacy
Zhenyu Guan, Xiangyu Kong, Fangwei Zhong, and Yizhou Wang · 2024
Closest in time.
Economics arena for large language models
Shangmin Guo, Haoran Bu, Haochuan Wang, Yi Ren, Dianbo Sui, Yuming Shang, and Siting Lu · 2024
Closest in time.
Fundamental problems with model editing: How should rational belief revision work in llms?
Peter Hase, Thomas Hofweber, Xiang Zhou, Elias Stengel-Eskin, and Mohit Bansal · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Albert Q Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, et al · 2023
Cited alongside, same era.
Language models with rationality
Nora Kassner, Oyvind Tafjord, Ashish Sabharwal, Kyle Richardson, Hinrich Schuetze, and Peter Clark · 2023
Cited alongside, same era.
Theory of mind might have spontaneously emerged in large language models
Michal Kosinski · 2023
Cited alongside, same era.
Llm-based agent society investigation: Collaboration and confrontation in avalon gameplay
Yihuai Lan, Zhiqiang Hu, Lei Wang, Yang Wang, De-Yong Ye, Peilin Zhao, Ee-Peng Lim, Hui Xiong, and Hao Wang · 2023
Cited alongside, same era.
Do llm agents exhibit social behavior?
Yan Leng and Yuan Yuan · 2023
Cited alongside, same era.
Avalonbench: Evaluating llms playing the game of avalon
Jonathan Light, Min Cai, Sheng Shen, and Ziniu Hu · 2023
Cited alongside, same era.
Large language models play starcraft ii: Benchmarks and a chain of summarization approach
Weiyu Ma, Qirui Mi, Xue Yan, Yuqiao Wu, Runji Lin, Haifeng Zhang, and Jun Wang · 2023
Cited alongside, same era.
Daniel A Herrmann and Benjamin A Levinstein · 2024
Closest in time.
Pokergpt: An end-to-end lightweight solver for multi-player texas hold’em via large language model
Chenghao Huang, Yanbo Cao, Yinlong Wen, Tao Zhou, and Yanru Zhang · 2024
Closest in time.
Learning to discuss strategically: A case study on one night ultimate werewolf
Xuanfa Jin, Ziyan Wang, Yali Du, Meng Fang, Haifeng Zhang, and Jun Wang · 2024
Closest in time.
Perceptions to beliefs: Exploring precursory inferences for theory of mind in large language models
Chani Jung, Dongkwan Kim, Jiho Jin, Jiseon Kim, Yeon Seonwoo, Yejin Choi, Alice Oh, and Hyunwoo Kim · 2024
Closest in time.
Still no lie detector for language models: Probing empirical and conceptual roadblocks
Benjamin A Levinstein and Daniel A Herrmann · 2024
Closest in time.
Efficacy of language model self-play in non-zero-sum games
Austen Liao, Nicholas Tomlin, and Dan Klein · 2024
Closest in time.
Strategist: Learning strategic skills by llms via bi-level tree search
Jonathan Light, Min Cai, Weiqin Chen, Guanzhi Wang, Xiusi Chen, Wei Cheng, Yisong Yue, and Ziniu Hu · 2024
Closest in time.
On llms-driven synthetic data generation, curation, and evaluation: A survey
Lin Long, Rui Wang, Ruixuan Xiao, Junbo Zhao, Xiao Ding, Gang Chen, and Haobo Wang · 2024
Closest in time.
Strategic behavior of large language models and the role of game structure versus contextual framing
Nunzio Lorè and Babak Heydari · 2024
Closest in time.
Computational experiments meet large language model based agents: A survey and perspective
Qun Ma, Xiao Xue, Deyu Zhou, Xiangning Yu, Donghua Liu, Xuwen Zhang, Zihan Zhao, Yifan Shen, Peilin Ji, Juanjuan Li, et al · 2024
Closest in time.
Advancing social intelligence in ai agents: Technical challenges and open questions
Leena Mathur, Paul Pu Liang, and Louis-Philippe Morency · 2024
Closest in time.
A turing test of whether ai chatbots are behaviorally similar to humans
Qiaozhu Mei, Yutong Xie, Walter Yuan, and Matthew O Jackson · 2024
Closest in time.
Ai emerges as the frontier in behavioral science
Juanjuan Meng · 2024
Closest in time.
Llms with personalities in multi-issue negotiation games
Sean Noh and Ho-Chun Herbert Chang · 2024
Closest in time.
Explaining decisions of agents in mixed-motive games
Maayan Orner, Oleg Maksimov, Akiva Kleinerman, Charles Ortiz, and Sarit Kraus · 2024
Closest in time.
Cooperate or collapse: Emergence of sustainability behaviors in a society of llm agents
Giorgio Piatti, Zhijing Jin, Max Kleiman-Weiner, Bernhard Schölkopf, Mrinmaya Sachan, and Rada Mihalcea · 2024
Closest in time.
Civrealm: A learning and reasoning odyssey in civilization for decision-making agents
Siyuan Qi, Shuo Chen, Yexin Li, Xiangyu Kong, Junqi Wang, Bangcheng Yang, Pring Wong, Yifan Zhong, Xiaoyuan Zhang, Zhaowei Zhang, Nian Liu, Wei Wang, Yaodong Yang, and Song-Chun Zhu · 2024
Closest in time.
Llm economicus? mapping the behavioral biases of llms via utility theory
Jillian Ross, Yoon Kim, and Andrew W Lo · 2024
Closest in time.
Evaluating the moral beliefs encoded in llms
Nino Scherrer, Claudia Shi, Amir Feder, and David Blei · 2024
Closest in time.
Truth-value judgment in language models: belief directions are context sensitive
Stefan F Schouten, Peter Bloem, Ilia Markov, and Piek Vossen · 2024
Closest in time.
Swarmbrain: Embodied agent for real-time strategy game starcraft ii via large language models
Xiao Shao, Weifu Jiang, Fei Zuo, and Mengqing Liu · 2024
Closest in time.
Glee: A unified framework and benchmark for language-based economic environments
Eilam Shapira, Omer Madmon, Itamar Reinman, Samuel Joseph Amouyal, Roi Reichart, and Moshe Tennenholtz · 2024
Closest in time.
Jen tse Huang, Eric Li, Man Ho Lam, Tian Liang, Wenxuan Wang, Youliang Yuan, Wenxiang Jiao, Xing Wang, Zhaopeng Tu, and Michael R. Lyu · 2024
Closest in time.
Deciphering digital detectives: Understanding llm behaviors and capabilities in multi-agent mystery games
Dekun Wu, Haochen Shi, Zhiyuan Sun, and Bang Liu · 2024
Closest in time.
Measuring bargaining abilities of llms: A benchmark and a buyer-enhancement method
Tian Xia, Zhiwei He, Tong Ren, Yibo Miao, Zhuosheng Zhang, Yang Yang, and Rui Wang · 2024
Closest in time.
Magic: Investigation of large language model powered multi-agent in cognition, adaptability, rationality and collaboration
Lin Xu, Zhiyuan Hu, Daquan Zhou, Hongyu Ren, Zhen Dong, Kurt Keutzer, See-Kiong Ng, and Jiashi Feng · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan · 2024
Closest in time.
Yauwai Yim, Chunkit Chan, Tianyu Shi, Zheye Deng, Wei Fan, Tianshi Zheng, and Yangqiu Song · 2024
Closest in time.
Let’s negotiate! a survey of negotiation dialogue systems
Haolan Zhan, Yufei Wang, Tao Feng, Yuncheng Hua, Suraj Sharma, Zhuang Li, Lizhen Qu, Zhaleh Semnani Azad, Ingrid Zukerman, and Gholamreza Haffari · 2024
Closest in time.