Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) reasoning abilities are increasingly being applied to classical board and card games, but the dominant approach -- involving prompting for direct move generation -- has significant drawbacks.
OpenSpiel: A framework for reinforcement learning in games
Marc Lanctot, Edward Lockhart, Jean-Baptiste Lespiau, Vinicius Zambaldi, Satyaki Upadhyay, Julien Pérolat, Sriram Srinivasan, Finbarr Timbers, Karl Tuyls, Shayegan Omidshafiei, Daniel Hennes, Dustin Morrill, Paul Muller, Timo Ewalds, Ryan Faulkner, János Kramár, Bart De Vylder, Brennan Saeta, James Bradbury, David Ding, Sebastian Borgeaud, Matthew Lai, Julian Schrittwieser, Thomas Anthony, Edward Hughes, Ivo Danihelka, and Jonah Ryan-Davis · 1908
Earlier work this paper cites.
Extensive games and the problem of information
H. W. Kuhn · 1953
Earlier work this paper cites.
A perspective on judgement and choice
Daniel Kahneman · 2003
Earlier work this paper cites.
Bandit based monte-carlo planning
Levente Kocsis and Csaba Szepesvári · 2006
Earlier work this paper cites.
Efficient selectivity and backup operators in monte-carlo tree search
Rémi Coulom · 2007
Earlier work this paper cites.
Multiagent Systems: Algorithmic, Game-Theoretic, and Logical Foundations
Y. Shoham and K. Leyton-Brown · 2009
Earlier work this paper cites.
Monte-carlo planning in large pomdps
David Silver and Joel Veness · 2010
Earlier work this paper cites.
Information set Monte Carlo tree search
Peter I. Cowling, Edward J. Powley, and Daniel Whitehouse · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
React: Synergizing reasoning and acting in language models., 2022
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao · 2022
Earlier work this paper cites.
Multi-Agent Reinforcement Learning: Foundations and Modern Approaches
Stefano V. Albrecht, Filippos Christianos, and Lukas Schäfer · 2024
Cited alongside, same era.
Gamebench: Evaluating strategic reasoning abilities of llm agents, 2024
Anthony Costarelli, Mat Allen, Roman Hauksson, Grace Sodunke, Suhas Hariharan, Carlson Cheng, Wenjie Li, Joshua Clymer, and Arjun Yadav · 2024
Cited alongside, same era.
Generating code world models with large language models guided by monte carlo tree search
Nicola Dainese, Matteo Merler, Minttu Alakuijala, and Pekka Marttinen · 2024
Cited alongside, same era.
Charting the shapes of stories with game theory
Constantinos Daskalakis, Ian Gemp, Yanchen Jiang, Renato Paes Leme, Christos Papadimitriou, and Georgios Piliouras · 2024
Cited alongside, same era.
Gtbench: Uncovering the strategic reasoning capabilities of llms via game-theoretic evaluations
Jinhao Duan, Renming Zhang, James Diffenderfer, Bhavya Kailkhura, Lichao Sun, Elias Stengel-Eskin, Mohit Bansal, Tianlong Chen, and Kaidi Xu · 2024
From natural language to extensive-form game representations
Shilong Deng, Yongzhao Wang, and Rahul Savani · 2025
Closest in time.
Are large language models reliable AI scientists? assessing reverse-engineering of black-box systems
Jiayi Geng, Howard Chen, Dilip Arumugam, and Thomas L Griffiths · 2025
Closest in time.
Leon Guertler, Bobby Cheng, Simon Yu, Bo Liu, Leshem Choshen, and Cheston Tan · 2025
Closest in time.
Yi Liao, Yu Gu, Yuan Sui, Zining Zhu, Yifan Lu, Guohua Tang, Zhongqian Sun, and Wei Yang · 2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Steering language models with game-theoretic solvers
Ian Gemp, Roma Patel, Yoram Bachrach, Marc Lanctot, Vibhavari Dasagi, Luke Marris, Georgios Piliouras, Siqi Liu, and Karl Tuyls · 2024
Cited alongside, same era.
Amortized planning with large-scale transformers: A case study on chess
Anian Ruoss, Grégoire Delétang, Sourabh Medapati, Jordi Grau-Moya, Li Kevin Wenliang, Elliot Catt, John Reid, Cannada A. Lewis, Joel Veness, and Tim Genewein · 2024
Cited alongside, same era.
Gambit: The package for computation in game theory , version 16.2.0 edition, 2024
Rahul Savani and Theodore L. Turocy · 2024
Cited alongside, same era.
Lucia Cipolina-Kun, Marianna Nezhurina, and Jenia Jitsev · 2025
Cited alongside, same era.
LLM-guided probabilistic program induction for POMDP model estimation
Aidan Curtis, Hao Tang, Thiago Veloso, Kevin Ellis, Joshua Tenenbaum, Tomás Lozano-Pérez, and Leslie Pack Kaelbling · 2025
Cited alongside, same era.
lmgame-bench: How good are llms at playing games?, 2025a
Lanxiang Hu, Mingjia Huo, Yuxuan Zhang, Haoyang Yu, Eric P. Xing, Ion Stoica, Tajana Rosing, Haojian Jin, and Hao Zhang
Cited in the paper.
A survey on large language model-based game agents, 2025b
Sihao Hu, Tiansheng Huang, Gaowen Liu, Ramana Rao Kompella, Fatih Ilhan, Selim Furkan Tekin, Yichang Xu, Zachary Yahn, and Ling Liu
Cited in the paper.
Kevin Murphy · 2025
Closest in time.
Mastering board games by external and internal planning with language models
John Schultz, Jakub Adamek, Matej Jusup, Marc Lanctot, Michael Kaisers, Sarah Perrin, Daniel Hennes, Jeremy Shar, Cannada Lewis, Anian Ruoss, Tom Zahavy, Petar Veličković, Laurel Prince, Satinder Singh, Eric Malmi, and Nenad Tomašev · 2025
Closest in time.
Game theory meets large language models: A systematic survey with taxonomy and new frontiers, 2025
Haoran Sun, Yusen Wu, Peng Wang, Wei Chen, Yukun Cheng, Xiaotie Deng, and Xu Chu · 2025
Closest in time.
Measuring general intelligence with generated games, 2025
Vivek Verma, David Huang, William Chen, Dan Klein, and Nicholas Tomlin · 2025
Closest in time.
Learning strategic language agents in the werewolf game with iterative latent space policy optimization
Zelai Xu, Wanjun Gu, Chao Yu, Yi Wu, and Yu Wang · 2025
Closest in time.
Assessing adaptive world models in machines with novel games
Lance Ying, Katherine M Collins, Prafull Sharma, Cedric Colas, Kaiya Ivy Zhao, Adrian Weller, Zenna Tavares, Phillip Isola, Samuel J Gershman, Jacob D Andreas, Thomas L Griffiths, Francois Chollet, Kelsey R Allen, and Joshua B Tenenbaum · 2025
Closest in time.