Fetching the paper…
Reading the bibliography…
In recent years, Artificial Intelligence (AI) systems have surpassed human intelligence in a variety of computational tasks.
Intelligent machinery
A. Turing · 1948
Earlier work this paper cites.
Iterative solution of games by fictitious play
G. W. Brown · 1951
Earlier work this paper cites.
Neural network ensembles
L. K. Hansen and P. Salamon · 1990
Earlier work this paper cites.
The emperor’s new mind: Concerning computers, minds, and the laws of physics, 1990
R. Penrose and N. D. Mermin · 1990
Earlier work this paper cites.
From tools to theories: A heuristic of discovery in cognitive psychology
G. Gigerenzer · 1991
Earlier work this paper cites.
Shadows of the Mind , volume 4
R. Penrose · 1994
Earlier work this paper cites.
On combining classifiers
J. Kittler, M. Hatef, R. P. Duin, and J. Matas · 1998
Earlier work this paper cites.
Kasparov against the world: the story of the greatest online challenge
G. K. Kasparov and D. King · 2000
Earlier work this paper cites.
Creativity from constraints: The psychology of breakthrough
P. D. Stokes · 2005
Earlier work this paper cites.
Generalised weakened fictitious play
D. S. Leslie and E. J. Collins · 2006
Earlier work this paper cites.
Ensemble based systems in decision making
R. Polikar · 2006
Earlier work this paper cites.
Dvoretsky’s endgame manual
M. Dvoretsky · 2010
Earlier work this paper cites.
Evolving a diversity of virtual creatures through novelty search and local competition
J. Lehman and K. O. Stanley · 2011
Earlier work this paper cites.
Multi-armed bandits with episode context
C. D. Rosin · 2011
Earlier work this paper cites.
Detecting fortresses in chess
M. Guid and I. Bratko · 2012
Earlier work this paper cites.
Computational rationality: Linking mechanism and behavior through bounded utility maximization
R. L. Lewis, A. Howes, and S. Singh · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Earlier work this paper cites.
Illuminating search spaces by mapping elites
J. Mouret and J. Clune · 2015
Earlier work this paper cites.
Alphago versus lee sedol, 2016
C. Metz · 2016
Earlier work this paper cites.
Deep exploration via bootstrapped dqn
I. Osband, C. Blundell, A. Pritzel, and B. Van Roy · 2016
Earlier work this paper cites.
Quality diversity: A new frontier for evolutionary computation
J. K. Pugh, L. B. Soros, and K. O. Stanley · 2016
Earlier work this paper cites.
Deep reinforcement learning with double q-learning
H. Van Hasselt, A. Guez, and D. Silver · 2016
Earlier work this paper cites.
Averaged-dqn: Variance reduction and stabilization for deep reinforcement learning
O. Anschel, N. Baram, and N. Shimkin · 2017
Earlier work this paper cites.
Successor features for transfer in reinforcement learning
A. Barreto, W. Dabney, R. Munos, J. J. Hunt, T. Schaul, H. P. van Hasselt, and D. Silver · 2017
Earlier work this paper cites.
Quality and diversity optimization: A unifying modular framework
A. Cully and Y. Demiris · 2017
Earlier work this paper cites.
Will this position help understand human consciousness?, 2017
P. Doggers · 2017
Earlier work this paper cites.
Variational intrinsic control
K. Gregor, D. J. Rezende, and D. Wierstra · 2017
Cited alongside, same era.
Population based training of neural networks, 2017
M. Jaderberg, V. Dalibard, S. Osindero, W. M. Czarnecki, J. Donahue, A. Razavi, O. Vinyals, T. Green, I. Dunning, K. Simonyan, C. Fernando, and K. Kavukcuoglu · 2017
Cited alongside, same era.
Deep Thinking: Where Machine Intelligence Ends and Human Creativity Begins
G. Kasparov · 2017
Cited alongside, same era.
A unified game-theoretic approach to multiagent reinforcement learning
M. Lanctot, V. Zambaldi, A. Gruslys, A. Lazaridou, K. Tuyls, J. Pérolat, D. Silver, and T. Graepel · 2017
Cited alongside, same era.
Stein variational policy gradient
Y. Liu, P. Ramachandran, Q. Liu, and J. Peng · 2017
Cited alongside, same era.
Deepstack: Expert-level artificial intelligence in heads-up no-limit poker
M. Moravčík, M. Schmid, N. Burch, V. Lisỳ, D. Morrill, N. Bard, T. Davis, K. Waugh, M. Johanson, and M. Bowling · 2017
How deepmind boss demis hassabis used chess to get billionaire peter thiel to ‘take notice’ of his ai lab, 2020
S. Shead · 2020
Later among the works it cites.
Assessing game balance with alphazero: Exploring alternative rule sets in chess, 2020
N. Tomašev, U. Paquet, D. Hassabis, and V. Kramnik · 2020
Later among the works it cites.
Planning in hierarchical reinforcement learning: Guarantees for using local policies
T. Zahavy, A. Hasidim, H. Kaplan, and Y. Mansour · 2020
Later among the works it cites.
10 positions chess engines just don’t understand, 2021
S. Copeland · 2021
Later among the works it cites.
Adversarially guided actor-critic
Y. Flet-Berliac, J. Ferret, O. Pietquin, P. Preux, and M. Geist · 2021
Later among the works it cites.
Towards unifying behavioral and response diversity for open-ended learning in zero-sum games
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Mastering the game of go without human knowledge
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al · 2017
Cited alongside, same era.
Using centroidal voronoi tessellations to scale up the multi-dimensional archive of phenotypic elites algorithm, 2017
V. Vassiliades, K. Chatzilygeroudis, and J.-B. Mouret · 2017
Cited alongside, same era.
Frank-wolfe splitting via augmented lagrangian method
G. Gidel, F. Pedregosa, and S. Lacoste-Julien · 2018
Cited alongside, same era.
Chess, a <i>drosophila</i> of reasoning
G. Kasparov · 2018
Cited alongside, same era.
Film: Visual reasoning with a general conditioning layer
E. Perez, F. Strub, H. De Vries, V. Dumoulin, and A. Courville · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Cited alongside, same era.
X. Liu, H. Jia, Y. Wen, Y. Hu, Y. Chen, C. Fan, Z. Hu, and Y. Yang · 2021
Later among the works it cites.
Trajectory diversity for zero-shot coordination
A. Lupu, B. Cui, H. Hu, and J. Foerster · 2021
Later among the works it cites.
Ensemble bootstrapping for q-learning
O. Peer, C. Tessler, N. Merlis, and R. Meir · 2021
Later among the works it cites.
Chess fortresses, a causal test for state of the art symbolic [neuro] architectures
H. Steingrimsson · 2021
Later among the works it cites.
Collaborating with humans without human data
D. Strouse, K. McKee, M. Botvinick, E. Hughes, and R. Everett · 2021
Later among the works it cites.
Discovering faster matrix multiplication algorithms with reinforcement learning
A. Fawzi, M. Balog, A. Huang, T. Hubert, B. Romera-Paredes, M. Barekatain, A. Novikov, F. J. R Ruiz, J. Schrittwieser, G. Swirszcz, et al · 2022
Later among the works it cites.
Alphazero ideas, 2022
J. González-Díaz and I. Palacios-Huerta · 2022
Later among the works it cites.
Muzero with self-competition for rate control in vp9 video compression, 2022
A. Mandhane, A. Zhernov, M. Rauh, C. Gu, M. Wang, F. Xue, W. Shang, D. Pang, R. Claus, C.-H. Chiang, C. Chen, J. Han, A. Chen, D. J. Mankowitz, J. Broshear, J. Schrittwieser, T. Hubert, O. Vinyals, and T. Mann · 2022
Later among the works it cites.
Acquisition of chess knowledge in alphazero
T. McGrath, A. Kapishnikov, N. Tomašev, A. Pearce, M. Wattenberg, D. Hassabis, B. Kim, U. Paquet, and V. Kramnik · 2022
Later among the works it cites.
Measuring the non-transitivity in chess
R. Sanjaya, J. Wang, and Y. Yang · 2022
Later among the works it cites.
Approximate exploitability: learning a best response
F. Timbers, N. Bard, E. Lockhart, M. Lanctot, M. Schmid, N. Burch, J. Schrittwieser, T. Hubert, and M. Bowling · 2022
Later among the works it cites.
Adversarial policies beat professional-level go ais
T. T. Wang, A. Gleave, N. Belrose, T. Tseng, J. Miller, M. D. Dennis, Y. Duan, V. Pogrebniak, S. Levine, and S. Russell · 2022
Later among the works it cites.
Gpt-4 passes the bar exam, 2023
D. M. Katz, M. J. Bommarito, S. Gao, and P. Arredondo · 2023
Closest in time.
Beyond games: A systematic review of neural monte carlo tree search applications, 2023
M. Kemmerling, D. Lütticke, and R. H. Schmitt · 2023
Closest in time.
Deep laplacian-based options for temporally-extended exploration
M. Klissarov and M. C. Machado · 2023
Closest in time.
Temporal abstraction in reinforcement learning with the successor representation
M. C. Machado, A. Barreto, D. Precup, and M. Bowling · 2023
Closest in time.
Gpt-4 technical report, 2023
OpenAI · 2023
Closest in time.
Superhuman artificial intelligence can improve human decision-making by increasing novelty
M. Shin, J. Kim, B. van Opheusden, and T. L. Griffiths · 2023
Closest in time.
Towards expert-level medical question answering with large language models, 2023
K. Singhal, T. Tu, J. Gottweis, R. Sayres, E. Wulczyn, L. Hou, K. Clark, S. Pfohl, H. Cole-Lewis, D. Neal, M. Schaekermann, A. Wang, M. Amin, S. Lachgar, P. Mansfield, S. Prakash, B. Green, E. Dominowska, B. A. y Arcas, N. Tomasev, Y. Liu, R. Wong, C. Semturs, S. S. Mahdavi, J. Barral, D. Webster, G. S. Corrado, Y. Matias, S. Azizi, A. Karthikesalingam, and V. Natarajan · 2023
Closest in time.
Discovering policies with DOMiNO: Diversity optimization maintaining near optimality
T. Zahavy, Y. Schroecker, F. Behbahani, K. Baumli, S. Flennerhag, S. Hou, and S. Singh · 2023
Closest in time.