Fetching the paper…
Reading the bibliography…
While advances in multi-agent learning have enabled the training of increasingly complex agents, most existing techniques produce a final policy that is not designed to adapt to a new partner's strategy.
Some mathematical notes on three-mode factor analysis
Ledyard R Tucker · 1966
Earlier work this paper cites.
Analysis of individual differences in multidimensional scaling via an n-way generalization of “eckart-young” decomposition
J Douglas Carroll and Jih-Jie Chang · 1970
Earlier work this paper cites.
Foundations of the parafac procedure: Models and conditions for an ”explanatory” multi-model factor analysis
Richard A. Harshman · 1970
Earlier work this paper cites.
Forecasting the behavior of multivariate time series using neural networks
Kanad Chakraborty, Kishan Mehrotra, Chilukuri K Mohan, and Sanjay Ranka · 1992
Earlier work this paper cites.
Multitask learning
Rich Caruana · 1997
Earlier work this paper cites.
Dealing with non-stationary environments using context detection
Bruno C Da Silva, Eduardo W Basso, Ana LC Bazzan, and Paulo M Engel · 2006
Earlier work this paper cites.
Collaborative filtering recommender systems
J Ben Schafer, Dan Frankowski, Jon Herlocker, and Shilad Sen · 2007
Earlier work this paper cites.
Human-robot collaboration by intention recognition using probabilistic state machines
Muhammad Awais and Dominik Henrich · 2010
Earlier work this paper cites.
Multi-agent learning with policy prediction
Chongjie Zhang and Victor Lesser · 2010
Earlier work this paper cites.
Capir: Collaborative action planning with intention recognition
Truong-Huy Dinh Nguyen, David Hsu, Wee-Sun Lee, Tze-Yun Leong, Leslie Pack Kaelbling, Tomas Lozano-Perez, and Andrew Haydn Grant · 2011
Earlier work this paper cites.
Tensor-train decomposition
Ivan V Oseledets · 2011
Earlier work this paper cites.
Formalizing assistive teleoperation
Anca D. Dragan and Siddhartha S. Srinivasa · 2012
Earlier work this paper cites.
Shared autonomy via hindsight optimization
Shervin Javdani, Siddhartha Srinivasa, and J. Andrew (Drew) Bagnell · 2015
Earlier work this paper cites.
Efficient model learning from joint-action demonstrations for human-robot collaborative tasks
Stefanos Nikolaidis, Ramya Ramakrishnan, Keren Gu, and Julie Shah · 2015
Earlier work this paper cites.
Social LSTM: human trajectory prediction in crowded spaces
Alexandre Alahi, Kratarth Goel, Vignesh Ramanathan, Alexandre Robicquet, Fei-Fei Li, and Silvio Savarese · 2016
Earlier work this paper cites.
Spectral tensor-train decomposition
Daniele Bigoni, Allan P Engsig-Karup, and Youssef M Marzouk · 2016
Earlier work this paper cites.
Andrzej Cichocki, Namgil Lee, Ivan V Oseledets, A-H Phan, Qibin Zhao, and D Mandic · 2016
Earlier work this paper cites.
A bayesian approach for learning and tracking switching, non-stationary opponents
Pablo Hernandez-Leal, Benjamin Rosman, Matthew E Taylor, L Enrique Sucar, and Enrique Munoz de Cote · 2016
Earlier work this paper cites.
Information gathering actions over human internal state
Dorsa Sadigh, S. Shankar Sastry, Sanjit A. Seshia, and Anca D. Dragan · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Earlier work this paper cites.
Learning modular neural network policies for multi-task and multi-robot transfer
Coline Devin, Abhishek Gupta, Trevor Darrell, Pieter Abbeel, and Sergey Levine · 2017
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
One-shot visual imitation learning via meta-learning
Chelsea Finn, Tianhe Yu, Tianhao Zhang, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Neural collaborative filtering
Xiangnan He, Lizi Liao, Hanwang Zhang, Liqiang Nie, Xia Hu, and Tat-Seng Chua · 2017
Cited alongside, same era.
A survey of learning in multiagent environments: Dealing with non-stationarity
Pablo Hernandez-Leal, Michael Kaisers, Tim Baarslag, and Enrique Munoz de Cote · 2017
Cited alongside, same era.
Coordinated multi-agent imitation learning
Hoang M Le, Yisong Yue, Peter Carr, and Patrick Lucey · 2017
Cited alongside, same era.
Artificial cognition for social human–robot interaction: An implementation
Superhuman AI for multiplayer poker
Noam Brown and Tuomas Sandholm · 2019
Later among the works it cites.
On the utility of learning about humans for human-ai coordination
Micah Carroll, Rohin Shah, M. Ho, T. Griffiths, S. Seshia, P. Abbeel, and A. Dragan · 2019
Later among the works it cites.
Garage: A toolkit for reproducible reinforcement learning research
The garage contributors · 2019
Later among the works it cites.
A continuous analogue of the tensor-train decomposition
Alex Gorodetsky, Sertac Karaman, and Youssef Marzouk · 2019
Later among the works it cites.
Simplified action decoder for deep multi-agent reinforcement learning
Hengyuan Hu and Jakob N Foerster · 2019
Later among the works it cites.
No-press diplomacy: Modeling multi-agent gameplay
Philip Paquette, Yuchen Lu, Steven Bocco, Max O Smith, Satya Ortiz-Gagné, Jonathan K Kummerfeld, Satinder Singh, Joelle Pineau, and Aaron Courville · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Séverin Lemaignan, Mathieu Warnier, E Akin Sisbot, Aurélie Clodic, and Rachid Alami · 2017
Cited alongside, same era.
Infogail: Interpretable imitation learning from visual demonstrations
Yunzhu Li, Jiaming Song, and Stefano Ermon · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, Pieter Abbeel, and Igor Mordatch · 2017
Cited alongside, same era.
Human-robot mutual adaptation in collaborative tasks: Models and experiments
Stefanos Nikolaidis, David Hsu, and Siddhartha Srinivasa · 2017
Cited alongside, same era.
End-to-end driving via conditional imitation learning
Felipe Codevilla, Matthias Müller, Antonio López, Vladlen Koltun, and Alexey Dosovitskiy · 2018
Cited alongside, same era.
Learning with opponent-learning awareness
Jakob Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch · 2018
Cited alongside, same era.
Counterfactual multi-agent policy gradients
Jakob Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson · 2018
Cited alongside, same era.
Later among the works it cites.
Deep local trajectory replanning and control for robot navigation
Ashwini Pokle, Roberto Martín-Martín, Patrick Goebel, Vincent Chow, Hans M Ewald, Junwei Yang, Zhenkai Wang, Amir Sadeghian, Dorsa Sadigh, Silvio Savarese, et al · 2019
Later among the works it cites.
Multi-agent adversarial inverse reinforcement learning
Lantao Yu, Jiaming Song, and Stefano Ermon · 2019
Later among the works it cites.
The hanabi challenge: A new frontier for ai research
Nolan Bard, Jakob N Foerster, Sarath Chandar, Neil Burch, Marc Lanctot, H Francis Song, Emilio Parisotto, Vincent Dumoulin, Subhodeep Moitra, Edward Hughes, et al · 2020
Later among the works it cites.
Trust-aware decision making for human-robot collaboration: Model learning and planning
Min Chen, Stefanos Nikolaidis, Harold Soh, David Hsu, and Siddhartha Srinivasa · 2020
Later among the works it cites.
Shared autonomy with learned latent actions
Hong Jun Jeon, Dylan Losey, and Dorsa Sadigh · 2020
Later among the works it cites.
Trust in robots: Challenges and opportunities
Bing Cai Kok and Harold Soh · 2020
Later among the works it cites.
Interpretable and personalized apprenticeship scheduling: Learning interpretable scheduling policies from heterogeneous user demonstrations
Rohan R. Paleja, Andrew Silva, Letian Chen, and Matthew C. Gombolay · 2020
Later among the works it cites.
Too many cooks: Coordinating multi-agent collaboration through inverse planning
Rose E. Wang, Sarah A. Wu, James A. Evans, Joshua B. Tenenbaum, David C. Parkes, and Max Kleiman-Weiner · 2020
Later among the works it cites.
Learning latent representations to influence multi-agent interaction
Annie Xie, Dylan Losey, Ryan Tolsma, Chelsea Finn, and Dorsa Sadigh · 2020
Later among the works it cites.
Harnessing structures for value-based planning and reinforcement learning
Yuzhe Yang, Guo Zhang, Zhi Xu, and Dina Katabi · 2020
Later among the works it cites.
Human-level performance in no-press diplomacy via equilibrium search
Jonathan Gray, Adam Lerer, Anton Bakhtin, and Noam Brown · 2021
Later among the works it cites.
On the critical role of conventions in adaptive human-{ai} collaboration
Andy Shih, Arjun Sawhney, Jovana Kondic, Stefano Ermon, and Dorsa Sadigh · 2021
Later among the works it cites.
Influencing towards stable multi-agent interactions
Woodrow Zhouyuan Wang, Andy Shih, Annie Xie, and Dorsa Sadigh · 2021
Later among the works it cites.
Tt-rec: Tensor train compression for deep learning recommendation models
Chunxing Yin, Bilge Acun, Xing Liu, and Carole-Jean Wu · 2021
Later among the works it cites.