Fetching the paper…
Reading the bibliography…
Independent learners are agents that employ single-agent algorithms in multi-agent systems, intentionally ignoring the effect of other strategic agents.
Rational choice and the structure of the environment
Herbert A. Simon · 1956
Earlier work this paper cites.
Equilibrium in a stochastic n n -person game
Arlington M. Fink · 1964
Earlier work this paper cites.
Learning from Delayed Rewards
Christopher Watkins · 1989
Earlier work this paper cites.
On algorithms for simple stochastic games
Anne Condon · 1990
Earlier work this paper cites.
Rational learning leads to Nash equilibrium
Ehud Kalai and Ehud Lehrer · 1993
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents
Ming Tan · 1993
Earlier work this paper cites.
Learning to coordinate without sharing information
Sandip Sen, Mahendra Sekaran, and John Hale · 1994
Earlier work this paper cites.
Asynchronous stochastic approximation and Q-learning
J. N. Tsitsiklis · 1994
Earlier work this paper cites.
Subjective games and equilibria
Ehud Kalai and Ehud Lehrer · 1995
Earlier work this paper cites.
A generalized reinforcement-learning model: Convergence and applications
Michael L. Littman and Csaba Szepesvári · 1996
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
Caroline Claus and Craig Boutilier · 1998
Earlier work this paper cites.
Win–stay, lose–shift strategies for repeated games—Memory length, aspiration levels and noise
Martin Posch · 1999
Earlier work this paper cites.
Friend-or-foe Q-learning in general-sum games
Michael L. Littman · 2001
Earlier work this paper cites.
Nash Q-learning for general-sum stochastic games
Junling Hu and Michael P. Wellman · 2003
Earlier work this paper cites.
Beliefs in repeated games
John H. Nachbar · 2005
Earlier work this paper cites.
Oblivious equilibrium: A mean field approximation for large-scale dynamic games
Gabriel Y. Weintraub, C. Lanier Benkard, and Benjamin Van Roy · 2005
Earlier work this paper cites.
Regret testing: Learning to play Nash equilibrium without knowing you have an opponent
Dean Foster and H. Peyton Young · 2006
Earlier work this paper cites.
Large population stochastic dynamic games: Closed-loop McKean-Vlasov systems and the Nash certainty equivalence principle
Minyi Huang, Roland P. Malhamé, and Peter E. Caines · 2006
Earlier work this paper cites.
Large-population cost-coupled LQG problems with nonuniform agents: Individual-mass behavior and decentralized ϵ \epsilon -Nash equilibria
Minyi Huang, Peter E. Caines, and Roland P. Malhamé · 2007
Earlier work this paper cites.
Mean field games
Jean-Michel Lasry and Pierre-Louis Lions · 2007
Earlier work this paper cites.
Markov perfect industry dynamics with many firms
Gabriel Y. Weintraub, C. Lanier Benkard, and Benjamin Van Roy · 2008
Earlier work this paper cites.
Coordination of independent learners in cooperative Markov games
Laetitia Matignon, Guillaume J. Laurent, and Nadine Le Fort-Piat · 2009
Earlier work this paper cites.
Convergence to approximate Nash equilibria in congestion games
Steve Chien and Alistair Sinclair · 2011
Earlier work this paper cites.
Robust mean field games with application to production of an exhaustible resource
Dario Bauso, Hamidou Tembine, and Tamer Başar · 2012
Earlier work this paper cites.
Independent reinforcement learners in cooperative Markov games: A survey regarding coordination problems
Laetitia Matignon, Guillaume J. Laurent, and Nadine Le Fort-Piat · 2012
Earlier work this paper cites.
Near-potential games: Geometry and dynamics
Ozan Candogan, Asuman Ozdaglar, and Pablo A. Parrilo · 2013
Earlier work this paper cites.
Aspiration learning in coordination games
Georgios C. Chasparis, Ari Arapostathis, and Jeff S. Shamma · 2013
Earlier work this paper cites.
Opinion dynamics and stubbornness through mean-field games
Leonardo Stella, Fabio Bagagiolo, Dario Bauso, and Giacomo Como · 2013
Earlier work this paper cites.
Mean field games models: A brief survey
Diogo A. Gomes and João Saúde · 2014
Earlier work this paper cites.
Socio-economic applications of finite state mean field games
Diogo A. Gomes, Roberto M. Velho, and Marie-Therese Wolfram · 2014
Earlier work this paper cites.
Mean field equilibria of dynamic auctions with learning
Krishnamurthy Iyer, Ramesh Johari, and Mukund Sundararajan · 2014
Earlier work this paper cites.
Equilibria of dynamic games with many players: Existence, approximation, and market structure
Sachin Adlakha, Ramesh Johari, and Gabriel Y. Weintraub · 2015
Cited alongside, same era.
Mean field games with ergodic cost for discrete time Markov processes
Anup Biswas · 2015
Cited alongside, same era.
A micro-macro traffic model based on mean-field games
Geoffroy Chevalier, Jerome Le Ny, and Roland Malhamé · 2015
Cited alongside, same era.
Mean field games via controlled martingale problems: Existence of Markovian equilibria
Daniel Lacker · 2015
Cited alongside, same era.
Incentivizing sharing in realtime d2d streaming networks: A mean field game perspective
Jian Li, Rajarshi Bhattacharyya, Suman Paul, Srinivas Shakkottai, and Vijay Subramanian · 2016
Cited alongside, same era.
Decentralized Q-learning in zero-sum Markov games
Muhammed O. Sayin, Kaiqing Zhang, David Leslie, Tamer Başar, and Asuman Ozdaglar · 2021
Later among the works it cites.
Learning while playing in mean-field games: Convergence and optimality
Qiaomin Xie, Zhuoran Yang, Zhaoran Wang, and Andreea Minca · 2021
Later among the works it cites.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
Kaiqing Zhang, Zhuoran Yang, and Tamer Başar · 2021
Later among the works it cites.
Unified reinforcement Q-learning for mean field game and control problems
Andrea Angiuli, Jean-Pierre Fouque, and Mathieu Laurière · 2022
Closest in time.
Solving N-player dynamic routing games with congestion: A mean-field approach
Theophile Cabannes, Mathieu Laurière, Julien Perolat, Raphael Marinier, Sertan Girgin, Sarah Perrin, Olivier Pietquin, Alexandre M. Bayen, Eric Goubault, and Romuald Elie · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lenient learning in independent-learner stochastic cooperative games
Ermo Wei and Sean Luke · 2016
Cited alongside, same era.
On solutions of mean field games with ergodic cost
Ari Arapostathis, Anup Biswas, and Johnson Carroll · 2017
Cited alongside, same era.
Decentralized Q-learning for stochastic teams and games
Gürdal Arslan and Serdar Yüksel · 2017
Cited alongside, same era.
On the connection between symmetric n n -player games and mean field games
Markus Fischer · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, Pieter Abbeel, and Igor Mordatch · 2017
Cited alongside, same era.
Counterfactual multi-agent policy gradients
Jakob Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson · 2018
Cited alongside, same era.
Mean field games in nudge systems for societal networks
Jian Li, Bainan Xia, Xinbo Geng, Hao Ming, Srinivas Shakkottai, Vijay Subramanian, and Le Xie · 2018
Cited alongside, same era.
Luciano Campi and Markus Fischer · 2022
Closest in time.
Independent policy gradient for large-scale Markov potential games: Sharper rates, function approximation, and game-agnostic convergence
Dongsheng Ding, Chen-Yu Wei, Kaiqing Zhang, and Mihailo Jovanovic · 2022
Closest in time.
Simple agent, complex environment: Efficient reinforcement learning with agent states
Shi Dong, Benjamin Van Roy, and Zhengyuan Zhou · 2022
Closest in time.
Independent natural policy gradient always converges in Markov potential games
Roy Fox, Stephen M. McAleer, Will Overman, and Ioannis Panageas · 2022
Closest in time.
Convergence of finite memory Q-learning for POMDPs and near optimality of learned policies under filter stability
Ali D. Kara and Serdar Yüksel · 2022
Closest in time.
Learning mean field games: A survey
Mathieu Laurière, Sarah Perrin, Matthieu Geist, and Olivier Pietquin · 2022
Closest in time.
Global convergence of multi-agent policy gradient in Markov potential games
Stefanos Leonardos, Will Overman, Ioannis Panageas, and Georgios Piliouras · 2022
Closest in time.
Learning correlated equilibria in mean-field games
Paul Muller, Romuald Elie, Mark Rowland, Mathieu Laurière, Julien Perolat, Sarah Perrin, Matthieu Geist, Georgios Piliouras, Olivier Pietquin, and Karl Tuyls · 2022
Closest in time.
Generalization in mean field games by learning master policies
Sarah Perrin, Mathieu Laurière, Julien Pérolat, Romuald Élie, Matthieu Geist, and Olivier Pietquin · 2022
Closest in time.
Optimality of independently randomized symmetric policies for exchangeable stochastic teams with infinitely many decision makers
Sina Sanjari, Naci Saldi, and Serdar Yüksel · 2022
Closest in time.
Fictitious play in zero-sum stochastic games
Muhammed O. Sayin, Francesca Parise, and Asuman Ozdaglar · 2022
Closest in time.
Decentralized learning for optimality in stochastic dynamic teams and games with local control and global state information
Bora Yongacoglu, Gürdal Arslan, and Serdar Yüksel · 2022
Closest in time.
On the global convergence rates of decentralized softmax gradient play in Markov potential games
Runyu Zhang, Jincheng Mei, Bo Dai, Dale Schuurmans, and Na Li · 2022
Closest in time.
Subjective equilibria under beliefs of exogenous uncertainty for dynamic games
Gürdal Arslan and Serdar Yüksel · 2023
Closest in time.
Model-free reinforcement learning for mean field games
Rajesh Mishra, Sriram Vishwanath, and Deepanshu Vasal · 2023
Closest in time.
Episodic Logit-Q dynamics for efficient learning in stochastic teams
Onur Unlu and Muhammed O. Sayin · 2023
Closest in time.
Sequential decomposition of discrete-time mean-field games
Deepanshu Vasal · 2023
Closest in time.
Policy mirror ascent for efficient and independent learning in mean field games
Batuhan Yardim, Semih Cayci, Matthieu Geist, and Niao He · 2023
Closest in time.
Satisficing paths and independent multiagent reinforcement learning in stochastic games
Bora Yongacoglu, Gürdal Arslan, and Serdar Yüksel · 2023
Closest in time.
Oracle-free reinforcement learning in mean-field games along a single sample path
Muhammad Aneeq Uz Zaman, Alec Koppel, Sujay Bhatt, and Tamer Başar · 2023
Closest in time.
Reinforcement learning in non-Markovian environments
Siddharth Chandak, Pratik Shah, Vivek S. Borkar, and Parth Dodhia · 2024
Closest in time.
Q-learning for stochastic control under general information structures and non-Markovian environments
Ali D. Kara and Serdar Yüksel · 2024
Closest in time.
On learning history-based policies for controlling Markov decision processes
Gandharv Patil, Aditya Mahajan, and Doina Precup · 2024
Closest in time.
Periodic agent-state based Q-learning for POMDPs
Amit Sinha, Matthieu Geist, and Aditya Mahajan · 2024
Closest in time.
Generalizing better response paths and weakly acyclic games
Bora Yongacoglu, Gürdal Arslan, Lacra Pavel, and Serdar Yüksel · 2024
Closest in time.