Fetching the paper…
Reading the bibliography…
Critical sectors of human society are progressing toward the adoption of powerful artificial intelligence (AI) agents, which are trained individually on behalf of self-interested principals but deployed in a shared environment.
Potential games design using local information. In 2018 IEEE Conference on Decision and Control (CDC) . IEEE, 1911–1916
Changxi Li, Fenghua He, Hongsheng Qi, Daizhan Cheng, Longbiao Ma, Yonghong Wu, and Shuo Chen. 2018 · 1916
Earlier work this paper cites.
Marktform und Gleichgewicht
H. von Stackelberg. 1934 · 1934
Earlier work this paper cites.
Information structure, Stackelberg games, and incentive controllability
Yu-Chi Ho, P Luh, and Ramal Muralidharan. 1981 · 1981
Earlier work this paper cites.
Adapting bias by gradient descent: An incremental version of delta-bar-delta. In AAAI . 171–176
Richard S Sutton. 1992 · 1992
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Michael L Littman. 1994 · 1994
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation.. In NIPs , Vol. 99. Citeseer, 1057–1063
Richard S Sutton, David A McAllester, Satinder P Singh, Yishay Mansour, et al · 1999
Earlier work this paper cites.
Cooperative multi-agent learning: The state of the art
Liviu Panait and Sean Luke. 2005 · 2005
Earlier work this paper cites.
The topology of the 2x2 games: a new periodic table . Vol. 3
David Robinson and David Goforth. 2005 · 2005
Earlier work this paper cites.
Adaptive mechanism design: a metalearning approach. In Proceedings of the 8th international conference on Electronic commerce . 92–102
David Pardoe, Peter Stone, Maytal Saar-Tsechansky, and Kerem Tomak. 2006 · 2006
Earlier work this paper cites.
Handbook of computational economics: agent-based computational economics
Leigh Tesfatsion and Kenneth L Judd. 2006 · 2006
Earlier work this paper cites.
Reinforcement learning in the brain
Yael Niv. 2009 · 2009
Earlier work this paper cites.
Traffic congestion pricing methodologies and technologies
André de Palma and Robin Lindsey. 2011 · 2011
Earlier work this paper cites.
Reverse stackelberg games, part i: Basic framework. In 2012 IEEE International Conference on Control Applications . IEEE, 421–426
Noortje Groot, Bart De Schutter, and Hans Hellendoorn. 2012 · 2012
Earlier work this paper cites.
Designing games for distributed optimization
Na Li and Jason R Marden. 2013 · 2013
Earlier work this paper cites.
Potential games are necessary to ensure pure nash equilibria in cost sharing games
Ragavendran Gopalakrishnan, Jason R Marden, and Adam Wierman. 2014 · 2014
Earlier work this paper cites.
A survey on demand response programs in smart grids: Pricing methods and optimization algorithms
John S Vardakas, Nizar Zorba, and Christos V Verikoukis. 2014 · 2014
Earlier work this paper cites.
Recommender system application developments: a survey
Jie Lu, Dianshuang Wu, Mingsong Mao, Wei Wang, and Guangquan Zhang. 2015 · 2015
Earlier work this paper cites.
Trust region policy optimization. In International conference on machine learning . PMLR, 1889–1897
John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz. 2015 · 2015
Earlier work this paper cites.
On learning intrinsic rewards for policy gradient methods. In Advances in Neural Information Processing Systems . 4644–4654
Zeyu Zheng, Junhyuk Oh, and Satinder Singh. 2018 · 2015
Earlier work this paper cites.
Tensorflow: A system for large-scale machine learning. In 12th { \{ USENIX } \} symposium on operating systems design and implementation ( { \{ OSDI } \} 16) . 265–283
Martín Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, Michael Isard, et al · 2016
Earlier work this paper cites.
The social dilemma of autonomous vehicles
Jean-François Bonnefon, Azim Shariff, and Iyad Rahwan. 2016 · 2016
Earlier work this paper cites.
Bridging the divide in financial market forecasting: machine learners vs. financial economists
Ming-Wei Hsu, Stefan Lessmann, Ming-Chien Sung, Tiejun Ma, and Johnnie EV Johnson. 2016 · 2016
Cited alongside, same era.
Taxi dispatch with real-time sensing data in metropolitan areas: A receding horizon control approach
Fei Miao, Shuo Han, Shan Lin, John A Stankovic, Desheng Zhang, Sirajum Munir, Hua Huang, Tian He, and George J Pappas. 2016 · 2016
Cited alongside, same era.
A concise introduction to decentralized POMDPs
Frans A Oliehoek and Christopher Amato. 2016 · 2016
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Cited alongside, same era.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Cited alongside, same era.
A Perspective on Incentive Design: Challenges and Opportunities
Lillian J Ratliff, Roy Dong, Shreyas Sekar, and Tanner Fiez. 2019 · 2019
Later among the works it cites.
Automated Mechanism Design via Neural Networks. In Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems . 215–223
Weiran Shen, Pingzhong Tang, and Song Zuo. 2019 · 2019
Later among the works it cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Later among the works it cites.
Thriving in the era of pervasive AI: Deloitte’s State of AI in the Enterprise, 3rd Edition
Beena Ammanath, David Jarvis, and Susanne Hupfer. 2020 · 2020
Later among the works it cites.
Caring for the future can turn tragedy into comedy for long-term collective action under risk of collapse
Wolfram Barfuss, Jonathan F Donges, Vítor V Vasconcelos, Jürgen Kurths, and Simon A Levin. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Reinforcement mechanism design.. In IJCAI . 5146–5150
Pingzhong Tang. 2017 · 2017
Cited alongside, same era.
Learning with Opponent-Learning Awareness. In Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems . 122–130
Jakob Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch. 2018 · 2018
Cited alongside, same era.
Algorithms, bots, and political communication in the US 2016 election: The challenge of automated political communication for election law and administration
Philip N Howard, Samuel Woolley, and Ryan Calo. 2018 · 2018
Cited alongside, same era.
Inequity aversion improves cooperation in intertemporal social dilemmas. In Advances in neural information processing systems . 3326–3336
Edward Hughes, Joel Z Leibo, Matthew Phillips, Karl Tuyls, Edgar Dueñez-Guzman, Antonio García Castañeda, Iain Dunning, Tina Zhu, Kevin McKee, Raphael Koster, et al · 2018
Cited alongside, same era.
Data center cooling using model-predictive control. In NeurIPS
Nevena Lazic, Craig Boutilier, Tyler Lu, Eehern Wong, Binz Roy, MK Ryu, and Greg Imwalle. 2018 · 2018
Cited alongside, same era.
Stable Opponent Shaping in Differentiable Games. In International Conference on Learning Representations
Alistair Letcher, Jakob Foerster, David Balduzzi, Tim Rocktäschel, and Shimon Whiteson. 2018 · 2018
Cited alongside, same era.
Machine learning in agriculture: A review
Konstantinos G Liakos, Patrizia Busato, Dimitrios Moshou, Simon Pearson, and Dionysis Bochtis. 2018 · 2018
Cited alongside, same era.
Later among the works it cites.
Adaptive Mechanism Design: Learning to Promote Cooperation. In 2020 International Joint Conference on Neural Networks (IJCNN) . IEEE, 1–7
Tobias Baumann, Thore Graepel, and John Shawe-Taylor. 2020 · 2020
Later among the works it cites.
Autonomous automobilities: The social impacts of driverless vehicles
David Bissell, Thomas Birtchnell, Anthony Elliott, and Eric L Hsu. 2020 · 2020
Later among the works it cites.
Reinforcement Learning of Simple Indirect Mechanisms
Gianluca Brero, Alon Eden, Matthias Gerstgrasser, David C Parkes, and Duncan Rheingans-Yoo. 2020 · 2020
Later among the works it cites.
Open Problems in Cooperative AI
Allan Dafoe, Edward Hughes, Yoram Bachrach, Tantum Collins, Kevin R McKee, Joel Z Leibo, Kate Larson, and Thore Graepel. 2020 · 2020
Later among the works it cites.
Should I tear down this wall? Optimizing social metrics by evaluating novel actions
János Kramár, Neil Rabinowitz, Tom Eccles, and Andrea Tacchetti. 2020 · 2020
Later among the works it cites.
End-to-end learning and intervention in games
Jiayang Li, Jing Yu, Yu Nie, and Zhaoran Wang. 2020 · 2020
Later among the works it cites.
Optimizing millions of hyperparameters by implicit differentiation. In International Conference on Artificial Intelligence and Statistics . PMLR, 1540–1552
Jonathan Lorraine, Paul Vicol, and David Duvenaud. 2020 · 2020
Later among the works it cites.
Discovering Reinforcement Learning Algorithms
Junhyuk Oh, Matteo Hessel, Wojciech M Czarnecki, Zhongwen Xu, Hado P van Hasselt, Satinder Singh, and David Silver. 2020 · 2020
Later among the works it cites.
Adaptive incentive design
Lillian J Ratliff and Tanner Fiez. 2020 · 2020
Later among the works it cites.
Learning Intrinsic Rewards as a Bi-Level Optimization Problem. In Conference on Uncertainty in Artificial Intelligence . PMLR, 111–120
Bradly Stadie, Lunjun Zhang, and Jimmy Ba. 2020 · 2020
Later among the works it cites.
Sequential Social Dilemma Games
Eugene Vinitsky. 2020 · 2020
Later among the works it cites.
Meta-Gradient Reinforcement Learning with an Objective Discovered Online
Zhongwen Xu, Hado van Hasselt, Matteo Hessel, Junhyuk Oh, Satinder Singh, and David Silver. 2020 · 2020
Later among the works it cites.
Learning to Incentivize Other Learning Agents
Jiachen Yang, Ang Li, Mehrdad Farajtabar, Peter Sunehag, Edward Hughes, and Hongyuan Zha. 2020 · 2020
Later among the works it cites.
The AI Economist: Improving Equality and Productivity with AI-Driven Tax Policies
Stephan Zheng, Alexander Trott, Sunil Srinivasa, Nikhil Naik, Melvin Gruesbeck, David C Parkes, and Richard Socher. 2020 · 2020
Later among the works it cites.
Foundations of complexity economics
W Brian Arthur. 2021 · 2021
Closest in time.
Panayiotis Danassis, Aris Filos-Ratsikas, and Boi Faltings. 2021 · 2021
Closest in time.