Fetching the paper…
Reading the bibliography…
Empirical game-theoretic analysis (EGTA) has recently been applied successfully to analyze the behavior of large numbers of competing traders in a continuous double auction market.
Evolutionary stable strategies and game dynamics
Peter D Taylor and Leo B Jonker. 1978 · 1978
Earlier work this paper cites.
Integrated architectures for learning, planning, and reacting based on approximating dynamic programming. In Proceedings of the seventh International Conference on Machine Learning
Richard S Sutton. 1990 · 1990
Earlier work this paper cites.
Self-improving reactive agents based on reinforcement learning, planning and teaching
Long-Ji Lin. 1992 · 1992
Earlier work this paper cites.
Q-learning
Christopher JCH Watkins and Peter Dayan. 1992 · 1992
Earlier work this paper cites.
Allocative efficiency of markets with zero-intelligence traders: Market as a partial substitute for individual rationality
Dhananjay K Gode and Shyam Sunder. 1993 · 1993
Earlier work this paper cites.
Reinforcement learning for robots using neural networks
Long-Ji Lin. 1993 · 1993
Earlier work this paper cites.
Memoryless policies: Theoretical limitations and practical results. In From Animals to Animats 3: Proceedings of the Third International Conference on Simulation of Adaptive Behavior
Michael L Littman. 1994 · 1994
Earlier work this paper cites.
Learning without state-estimation in partially observable Markovian decision processes. In ICML
Satinder P Singh, Tommi Jaakkola, and Michael I Jordan. 1994 · 1994
Earlier work this paper cites.
Minimal-intelligence agents for bargaining behaviors in market-based environments
Dave Cliff. 1997 · 1997
Earlier work this paper cites.
Price formation in double auctions
Steven Gjerstad and John Dickhaut. 1998 · 1998
Earlier work this paper cites.
Using eligibility traces to find the best memoryless policy in partially observable Markov decision processes. In ICML
John Loch and Satinder P Singh. 1998 · 1998
Cited alongside, same era.
Reinforcement Learning: An Introduction
Richard S Sutton and Andrew G Barto. 1998 · 1998
Cited alongside, same era.
Methods for empirical game-theoretic analysis. In Proceedings of the National Conference on Artificial Intelligence
Michael P Wellman. 2006 · 1999
Cited alongside, same era.
High-performance bidding agents for the continuous double auction. In Proceedings of the 3rd ACM Conference on Electronic Commerce
Gerald Tesauro and Rajarshi Das. 2001 · 2001
Cited alongside, same era.
Finite-time analysis of the multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer. 2002 · 2002
Cited alongside, same era.
Stronger CDA strategies through empirical game-theoretic analysis and reinforcement learning. In Proceedings of The 8th International Conference on Autonomous Agents and Multiagent Systems-Volume 1
L Julian Schvartzman and Michael P Wellman. 2009 · 2009
Later among the works it cites.
Learning improved entertainment trading strategies for the TAC travel game
L Julian Schvartzman and Michael P Wellman. 2010 · 2010
Later among the works it cites.
Monte-Carlo planning in large POMDPs. In Advances in Neural Information Processing Systems
David Silver and Joel Veness. 2010 · 2010
Later among the works it cites.
Experience replay for real-time reinforcement learning control
Sander Adam, Lucian Buşoniu, and Robert Babuška. 2012 · 2012
Later among the works it cites.
A survey of Monte Carlo tree search methods
Cameron B Browne, Edward Powley, Daniel Whitehouse, Simon M Lucas, Peter Cowling, Philipp Rohlfshagen, Stephen Tavener, Diego Perez, Spyridon Samothrakis, Simon Colton, and others. 2012 · 2012
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Laurent Péret and Frédérick Garcia. 2004 · 2004
Cited alongside, same era.
Heuristic search value iteration for POMDPs. In Proceedings of the 20th Conference on Uncertainty in Artificial Intelligence
Trey Smith and Reid Simmons. 2004 · 2004
Cited alongside, same era.
Function approximation via tile coding: Automating parameter choice
Alexander A Sherstov and Peter Stone. 2005 · 2005
Cited alongside, same era.
Bandit based Monte-Carlo planning
Levente Kocsis and Csaba Szepesvári. 2006 · 2006
Cited alongside, same era.
Improved Monte-Carlo search
Levente Kocsis, Csaba Szepesvári, and Jan Willemson. 2006 · 2006
Cited alongside, same era.
Bryce Wiedenbeck and Michael P Wellman. 2012 · 2012
Later among the works it cites.
Playing Atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller. 2013 · 2013
Later among the works it cites.
Latency arbitrage, market fragmentation, and efficiency: A two-market model. In Proceedings of the fourteenth ACM conference on Electronic commerce
Elaine Wah and Michael P Wellman. 2013 · 2013
Later among the works it cites.
Welfare effects of market making in continuous double auctions. In Proceedings of the 2015 International Conference on Autonomous Agents and Multiagent Systems
Elaine Wah and Michael P Wellman. 2015 · 2015
Later among the works it cites.