Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have been increasingly used in real-world settings, yet their strategic decision-making abilities remain largely unexplored.
The mathematics of Tucker: A sampler
Tucker, A. W.; and Straffin Jr, P. D. 1983 · 1983
Earlier work this paper cites.
A course in game theory
Osborne, M. J.; and Rubinstein, A. 1994 · 1994
Earlier work this paper cites.
The Prisoner’s Dilemma and the Prisoners of the Prisoner’s Dilemma
Gilbert Jr, D. R. 1996 · 1996
Earlier work this paper cites.
The stag hunt
Skyrms, B. 2001 · 2001
Earlier work this paper cites.
Preference relations, social decision rules, singlepeakedness, and social welfare functions
Albouy, D. 2004 · 2004
Earlier work this paper cites.
Two-player Nonzero-sum-regular Games
Chatterjee, K. 2004 · 2004
Earlier work this paper cites.
Handbook of biological statistics , volume 2
McDonald, J. H. 2009 · 2009
Earlier work this paper cites.
Evolutionary dynamics of collective action in N-person stag hunt dilemmas
Pacheco, J. M.; Santos, F. C.; Souza, M. O.; and Skyrms, B. 2009 · 2009
Earlier work this paper cites.
Algorithmic game theory
Roughgarden, T. 2010 · 2010
Earlier work this paper cites.
Biostatistics for medical and biomedical practitioners
Hoffman, J. I. 2015 · 2015
Earlier work this paper cites.
Statistical notes for clinical researchers: Chi-squared test and Fisher’s exact test
Kim, H.-Y. 2017 · 2017
Earlier work this paper cites.
COURSE IN GAME THEORY
Martin, J. O. 2017 · 2017
Earlier work this paper cites.
Intuition and deliberation in the stag hunt game
Belloc, M.; Bilancini, E.; Boncinelli, L.; and D’Alessandro, S. 2019 · 2019
Earlier work this paper cites.
Calibrate before use: Improving few-shot performance of language models
Zhao, Z.; Wallace, E.; Feng, S.; Klein, D.; and Singh, S. 2021 · 2021
Earlier work this paper cites.
Large language models are zero-shot reasoners
Kojima, T.; Gu, S. S.; Reid, M.; Matsuo, Y.; and Iwasawa, Y. 2022 · 2022
Cited alongside, same era.
Factors of influence in prisoner’s dilemma task: A review of medical literature
Mantas, V.; Pehlivanidis, A.; Kotoula, V.; Papanikolaou, K.; Vassiliou, G.; Papaiakovou, A.; and Papageorgiou, C. 2022 · 2022
Cited alongside, same era.
Automatic chain of thought prompting in large language models
Zhang, Z.; Zhang, A.; Li, M.; and Smola, A. 2022 · 2022
Cited alongside, same era.
Using Large Language Models to Simulate Multiple Humans and Replicate Human Subject Studies
Aher, G.; Arriaga, R. I.; and Kalai, A. T. 2023 · 2023
Cited alongside, same era.
The Reversal Curse: LLMs trained on” A is B” fail to learn” B is A”
Berglund, L.; Tong, M.; Kaufmann, M.; Balesni, M.; Stickland, A. C.; Korbak, T.; and Evans, O. 2023 · 2023
Large language models are not fair evaluators
Wang, P.; Li, L.; Chen, L.; Zhu, D.; Lin, B.; Cao, Y.; Liu, Q.; Liu, T.; and Sui, Z. 2023 · 2023
Later among the works it cites.
Xu, L.; Hu, Z.; Zhou, D.; Ren, H.; Dong, Z.; Keutzer, K.; Ng, S. K.; and Feng, J. 2023 · 2023
Later among the works it cites.
Large language models are not robust multiple choice selectors
Zheng, C.; Zhou, H.; Meng, F.; Zhou, J.; and Huang, M. 2023 · 2023
Later among the works it cites.
Premise Order Matters in Reasoning with Large Language Models
Chen, X.; Chi, R. A.; Wang, X.; and Zhou, D. 2024 · 2024
Closest in time.
Gtbench: Uncovering the strategic reasoning limitations of llms via game-theoretic evaluations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Playing games with GPT: What can we learn about a large language model from canonical strategic games?
Brookins, P.; and DeBacker, J. M. 2023 · 2023
Cited alongside, same era.
Chen, J.; Yuan, S.; Ye, R.; Majumder, B. P.; and Richardson, K. 2023 · 2023
Cited alongside, same era.
Strategic reasoning with language models
Gandhi, K.; Sadigh, D.; and Goodman, N. D. 2023 · 2023
Cited alongside, same era.
GPT in Game Theory Experiments
Guo, F. 2023 · 2023
Cited alongside, same era.
A survey on large language models: Applications, challenges, limitations, and practical usage
Hadi, M. U.; Qureshi, R.; Shah, A.; Irfan, M.; Zafar, A.; Shaikh, M. B.; Akhtar, N.; Wu, J.; Mirjalili, S.; et al. 2023 · 2023
Cited alongside, same era.
Beyond Static Datasets: A Deep Interaction Approach to LLM Evaluation
Li, J.; Li, R.; and Liu, Q. 2023 · 2023
Cited alongside, same era.
Strategic behavior of large language models: Game structure vs. contextual framing
Lorè, N.; and Heydari, B. 2023 · 2023
Cited alongside, same era.
Duan, J.; Zhang, R.; Diffenderfer, J.; Kailkhura, B.; Sun, L.; Stengel-Eskin, E.; Bansal, M.; Chen, T.; and Xu, K. 2024 · 2024
Closest in time.
Can large language models serve as rational players in game theory? a systematic analysis
Fan, C.; Chen, J.; Jin, Y.; and He, H. 2024 · 2024
Closest in time.
Reverse training to nurse the reversal curse
Golovneva, O.; Allen-Zhu, Z.; Weston, J.; and Sukhbaatar, S. 2024 · 2024
Closest in time.
Large language model based multi-agents: A survey of progress and challenges
Guo, T.; Chen, X.; Wang, Y.; Chang, R.; Pei, S.; Chawla, N. V.; Wiest, O.; and Zhang, X. 2024 · 2024
Closest in time.
Huang, J.-t.; Li, E. J.; Lam, M. H.; Liang, T.; Wang, W.; Yuan, Y.; Jiao, W.; Wang, X.; Tu, Z.; and Lyu, M. R. 2024 · 2024
Closest in time.
OpenAI. 2024
2024
Closest in time.
Your LLM judge may be biased - ai alignment forum
Papadatos, H.; and Freedman, R. 2024 · 2024
Closest in time.
Game Theory
Ross, D. 2024 · 2024
Closest in time.
A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications
Sahoo, P.; Singh, A. K.; Saha, S.; Jain, V.; Mondal, S.; and Chadha, A. 2024 · 2024
Closest in time.
Judging llm-as-a-judge with mt-bench and chatbot arena
Zheng, L.; Chiang, W.-L.; Sheng, Y.; Zhuang, S.; Wu, Z.; Zhuang, Y.; Lin, Z.; Li, Z.; Li, D.; Xing, E.; et al. 2024 · 2024
Closest in time.