Fetching the paper…
Reading the bibliography…
In multi-agent systems, complex interacting behaviors arise due to the high correlations among agents.
Non-Cooperative games
John Nash · 1951
Earlier work this paper cites.
Remarks on some nonparametric estimates of a density function
Murray Rosenblatt · 1956
Earlier work this paper cites.
Efficient training of artificial neural networks for autonomous navigation
Dean A Pomerleau · 1991
Earlier work this paper cites.
Markov games as a framework for multi-Agent reinforcement learning
Michael L Littman · 1994
Earlier work this paper cites.
Opponent modeling in poker
Darse Billings, Denis Papp, Jonathan Schaeffer, and Duane Szafron · 1998
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
Caroline Claus and Craig Boutilier · 1998
Earlier work this paper cites.
Learning agents for uncertain environments
Stuart J Russell · 1998
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng, Stuart J Russell, et al · 2000
Earlier work this paper cites.
Correlated q-Learning
Amy Greenwald, Keith Hall, and Roberto Serrano · 2003
Earlier work this paper cites.
A comprehensive survey of multiagent reinforcement learning
Lucian Bu, Robert Babu, Bart De Schutter, et al · 2008
Earlier work this paper cites.
Efficient reductions for imitation learning
Stéphane Ross and Drew Bagnell · 2010
Earlier work this paper cites.
Game theory-based opponent modeling in large imperfect-information games
Sam Ganzfried and Tuomas Sandholm · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-Regret online learning
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Inverse reinforcement learning for decentralized non-Cooperative multiagent systems
Tummalapalli Sudhamsh Reddy, Vamsikrishna Gopikrishna, Gergely Zaruba, and Manfred Huber · 2012
Earlier work this paper cites.
Computational rationalization: The inverse equilibrium problem
Kevin Waugh, Brian D Ziebart, and J Andrew Bagnell · 2013
Earlier work this paper cites.
Infinite time horizon maximum causal entropy inverse reinforcement learning
Michael Bloem and Nicholas Bambos · 2014
Earlier work this paper cites.
Multi-robot inverse reinforcement learning under occlusion with interactions
Kenneth Bogert and Prashant Doshi · 2014
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Multi-Agent inverse reinforcement learning for zero-Sum games
Xiaomin Lin, Peter A Beling, and Randy Cogill · 2014
Cited alongside, same era.
Learning driver behavior models from traffic observations for decision making and planning
Tobias Gindele, Sebastian Brechtel, and Rudiger Dillmann · 2015
Cited alongside, same era.
Optimizing neural networks with kronecker-factored approximate curvature
James Martens and Roger Grosse · 2015
Cited alongside, same era.
Opponent modeling in deep reinforcement learning
Autonomous agents modelling other agents: A comprehensive survey and open problems
Stefano V Albrecht and Peter Stone · 2018
Later among the works it cites.
Multi-Agent imitation learning for driving simulation
Raunak P Bhattacharyya, Derek J Phillips, Blake Wulfe, Jeremy Morton, Alex Kuefler, and Mykel J Kochenderfer · 2018
Later among the works it cites.
Counterfactual multi-Agent policy gradients
Jakob N Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson · 2018
Later among the works it cites.
Learning policy representations in multiagent systems
Aditya Grover, Maruan Al-Shedivat, Jayesh Gupta, Yuri Burda, and Harrison Edwards · 2018
Later among the works it cites.
Max Jaderberg, Wojciech M Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C Rabinowitz, Ari S Morcos, Avraham Ruderman, et al · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
He He, Jordan Boyd-Graber, Kevin Kwok, and Hal Daumé III · 2016
Cited alongside, same era.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
Understanding interactions between traffic participants based on learned behaviors
Florian Kuhnt, Jens Schulz, Thomas Schamm, and J Marius Zöllner · 2016
Cited alongside, same era.
Inverse reinforcement learning in swarm systems
Adrian Šošic, Wasiur R KhudaBukhsh, Abdelhak M Zoubir, and Heinz Koeppl · 2016
Cited alongside, same era.
Making friends on the fly: Cooperating with new teammates
Samuel Barrett, Avi Rosenfeld, Sarit Kraus, and Peter Stone · 2017
Cited alongside, same era.
Reinforcement learning with deep energy-based policies
Tuomas Haarnoja, Haoran Tang, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Coordinated multi-Agent imitation learning
Hoang M Le, Yisong Yue, Peter Carr, and Patrick Lucey · 2017
Cited alongside, same era.
Discriminator-Actor-Critic: Addressing sample inefficiency and reward bias in adversarial imitation learning
Ilya Kostrikov, Kumar Krishna Agrawal, Debidatta Dwibedi, Sergey Levine, and Jonathan Tompson · 2018
Later among the works it cites.
Multi-Agent inverse reinforcement learning for general-sum stochastic games
Xiaomin Lin, Stephen C Adams, and Peter A Beling · 2018
Later among the works it cites.
Openai five
OpenAI · 2018
Later among the works it cites.
Multi-Agent generative adversarial imitation learning
Jiaming Song, Hongyu Ren, Dorsa Sadigh, and Stefano Ermon · 2018
Later among the works it cites.
Multiagent soft q-Learning
Ermo Wei, Drew Wicke, David Freelan, and Sean Luke · 2018
Later among the works it cites.
Multi-Agent deep reinforcement learning for large-scale traffic signal control
Tianshu Chu, Jie Wang, Lara Codecà, and Zhaojian Li · 2019
Later among the works it cites.
Efficient ridesharing order dispatching with mean field multi-Agent reinforcement learning
Minne Li, Zhiwei Qin, Yan Jiao, Yaodong Yang, Jun Wang, Chenxi Wang, Guobin Wu, and Jieping Ye · 2019
Later among the works it cites.
A regularized opponent model with maximum entropy objective
Zheng Tian, Ying Wen, Zhichen Gong, Faiz Punakkath, Shihao Zou, and Jun Wang · 2019
Later among the works it cites.
Probabilistic recursive reasoning for multi-Agent reinforcement learning
Ying Wen, Yaodong Yang, Rui Luo, Jun Wang, and Wei Pan · 2019
Later among the works it cites.
Multi-Agent adversarial inverse reinforcement learning
Lantao Yu, Jiaming Song, and Stefano Ermon · 2019
Later among the works it cites.