Fetching the paper…
Reading the bibliography…
Here we explore a new algorithmic framework for multi-agent reinforcement learning, called Malthusian reinforcement learning, which extends self-play to include fitness-linked population size dynamics that drive ongoing innovation.
Asynchronous methods for deep reinforcement learning. In International conference on machine learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
Soft Selection, Hard Selection, Kin Selection, and Group Selection
Michael J. Wade. 1985 · 1985
Earlier work this paper cites.
Evolution of Dispersal: Theoretical Models and Empirical Tests Using Birds and Mammals
M L Johnson and M S Gaines. 1990 · 1990
Earlier work this paper cites.
TD-Gammon, A Self-Teaching Backgammon Program, Achieves Master-Level Play
Gerald Tesauro. 1995 · 1995
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Cooperation and Competition Between Relatives
Stuart A. West, Ido Pen, and Ashleigh S. Griffin. 2002 · 2002
Earlier work this paper cites.
Inclusion of facilitation into ecological theory
John F Bruno, John J Stachowicz, and Mark D Bertness. 2003 · 2003
Earlier work this paper cites.
Cultural group selection, coevolutionary processes and large-scale cooperation
Joseph Henrich. 2004a · 2004
Earlier work this paper cites.
Demography and cultural evolution: how adaptive cultural processes can produce maladaptive losses—the Tasmanian case
Joseph Henrich. 2004b · 2004
Earlier work this paper cites.
Intrinsically motivated reinforcement learning. In Advances in neural information processing systems
Nuttapong Chentanez, Andrew G Barto, and Satinder P Singh. 2005 · 2005
Earlier work this paper cites.
Empowerment: A universal agent-centric measure of control. In Evolutionary Computation, 2005. The 2005 IEEE Congress on
Alexander S Klyubin, Daniel Polani, and Chrystopher L Nehaniv. 2005 · 2005
Earlier work this paper cites.
Why did modern human populations disperse from Africa ca. 60,000 years ago? A new model
Paul Mellars. 2006 · 2006
Earlier work this paper cites.
Universal intelligence: A definition of machine intelligence
Shane Legg and Marcus Hutter. 2007 · 2007
Earlier work this paper cites.
A farewell to alms: a brief economic history of the world
Gregory Clark. 2008 · 2008
Cited alongside, same era.
The late Pleistocene dispersal of modern humans in the Americas
Ted Goebel, Michael R Waters, and Dennis H O’rourke. 2008 · 2008
Cited alongside, same era.
Late Pleistocene demography and the appearance of modern human behavior
Adam Powell, Stephen Shennan, and Mark G Thomas. 2009 · 2009
Cited alongside, same era.
Cultural innovations and demographic change
Peter J Richerson, Robert Boyd, and Robert L Bettinger. 2009 · 2009
Cited alongside, same era.
Markets, religion, community size, and the evolution of fairness and punishment
Joseph Henrich, Jean Ensminger, Richard McElreath, Abigail Barr, Clark Barrett, Alexander Bolyanatz, Juan Camilo Cardenas, Michael Gurven, Edwins Gwako, Natalie Henrich, et al · 2010
Cited alongside, same era.
Formal theory of creativity, fun, and intrinsic motivation (1990–2010)
Mastering the Game of Go with Deep Neural Networks and Tree Search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis. 2016 · 2016
Later among the works it cites.
Emergent complexity via multi-agent competition
Trapit Bansal, Jakub Pachocki, Szymon Sidor, Ilya Sutskever, and Igor Mordatch. 2017 · 2017
Later among the works it cites.
Edoardo Conti, Vashisht Madhavan, Felipe Petroski Such, Joel Lehman, Kenneth O Stanley, and Jeff Clune. 2017 · 2017
Later among the works it cites.
Why are there so many explanations for primate brain evolution?
RIM Dunbar and Susanne Shultz. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jürgen Schmidhuber. 2010 · 2010
Cited alongside, same era.
The cultural niche: Why social learning is essential for human adaptation
Robert Boyd, Peter J Richerson, and Joseph Henrich. 2011 · 2011
Cited alongside, same era.
Late Pleistocene climate change and the global expansion of anatomically modern humans
Anders Eriksson, Lia Betti, Andrew D Friend, Stephen J Lycett, Joy S Singarayer, Noreen von Cramon-Taubadel, Paul J Valdes, Francois Balloux, and Andrea Manica. 2012 · 2012
Cited alongside, same era.
Human evolution out of Africa: the role of refugia and climate change
John R Stewart and Chris B Stringer. 2012 · 2012
Cited alongside, same era.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling. 2013 · 2013
Cited alongside, same era.
From cultural traditions to cumulative culture: parameterizing the differences between human and nonhuman culture
Marius Kempe, Stephen J Lycett, and Alex Mesoudi. 2014 · 2014
Cited alongside, same era.
Unifying count-based exploration and intrinsic motivation. In Advances in Neural Information Processing Systems
Marc Bellemare, Sriram Srinivasan, Georg Ostrovski, Tom Schaul, David Saxton, and Remi Munos. 2016 · 2016
Cited alongside, same era.
Jarryd Martin, Suraj Narayanan Sasikumar, Tom Everitt, and Marcus Hutter. 2017 · 2017
Later among the works it cites.
Count-based exploration with neural density models
Georg Ostrovski, Marc G Bellemare, Aaron van den Oord, and Rémi Munos. 2017 · 2017
Later among the works it cites.
Curiosity-driven exploration by self-supervised prediction. In International Conference on Machine Learning (ICML)
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell. 2017 · 2017
Later among the works it cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy Lillicrap, Karen Simonyan, and Demis Hassabis. 2017 · 2017
Later among the works it cites.
Large-scale study of curiosity-driven learning
Yuri Burda, Harri Edwards, Deepak Pathak, Amos Storkey, Trevor Darrell, and Alexei A Efros. 2018 · 2018
Closest in time.
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures. In Proceedings of the 35th International Conference on Machine Learning
Lasse Espeholt, Hubert Soyer, Remi Munos, Karen Simonyan, Vlad Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, Shane Legg, and Koray Kavukcuoglu. 2018 · 2018
Closest in time.
Max Jaderberg, Wojciech M Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C Rabinowitz, Ari S Morcos, Avraham Ruderman, et al · 2018
Closest in time.
Intrinsic Social Motivation via Causal Influence in Multi-Agent RL
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Caglar Gulcehre, Pedro A Ortega, DJ Strouse, Joel Z Leibo, and Nando de Freitas. 2018 · 2018
Closest in time.
Yaodong Yang, Lantao Yu, Yiwei Bai, Ying Wen, Weinan Zhang, and Jun Wang. 2018 · 2018
Closest in time.