Fetching the paper…
Reading the bibliography…
Learning interpretable communication is essential for multi-agent and human-agent teams (HATs).
D. M. Green and J. A. Swets, Signal Detection Theory and Psychophysics . New York: Wiley, 1966
1966
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3, pp. 229–256, 1992
1992
Earlier work this paper cites.
J. J. Van Merrienboer and J. Sweller, “Cognitive load theory and complex learning: Recent developments and future directions,” Educational psychology review , vol. 17, no. 2, pp. 147–177, 2005
2005
Earlier work this paper cites.
D. Tse, R. F. Langston, M. Kakeyama, I. Bethus, P. A. Spooner, E. R. Wood, M. P. Witter, and R. G. Morris, “Schemas and memory consolidation,” Science , vol. 316, no. 5821, pp. 76–82, 2007
2007
Earlier work this paper cites.
X. Fan and J. Yen, “Modeling cognitive loads for evolving shared mental models in human–agent collaboration,” IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics) , vol. 41, no. 2, pp. 354–367, 2010
2010
Earlier work this paper cites.
S. Nikolaidis and J. Shah, “Human-robot cross-training: Computational formulation, modeling and evaluation of a human team training strategy,” in 2013 8th ACM/IEEE International Conference on Human-Robot Interaction (HRI) , 2013, pp. 33–40
2013
Earlier work this paper cites.
S. Li, W. Sun, and T. Miller, “Communication in human-agent teams for tasks with joint action,” in International Workshop on Coordination, Organizations, Institutions, and Norms in Agent Systems . Springer, 2015, pp. 224–241
2015
Earlier work this paper cites.
S. Sukhbaatar, R. Fergus et al. , “Learning multiagent communication with backpropagation,” Advances in neural information processing systems , vol. 29, pp. 2244–2252, 2016
2016
Earlier work this paper cites.
A. Soltani, P. Khorsand, C. Guo, and J. Liu, “Neural substrates of cognitive biases during probabilistic inference,” Nature Communications , vol. 7, no. 1, p. e1000858, 2016
2016
Earlier work this paper cites.
J. N. Foerster, Y. M. Assael, N. de Freitas, and S. Whiteson, “Learning to communicate with deep multi-agent reinforcement learning,” in Proceedings of the 30th International Conference on Neural Information Processing Systems , 2016, pp. 2145–2153
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Andreas, A. Dragan, and D. Klein, “Translating neuralese,” in Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2017, pp. 232–242
2017
Earlier work this paper cites.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, Y. Chen, T. Lillicrap, F. Hui, L. Sifre, G. van den Driessche, T. Graepel, and D. Hassabis, “Mastering the game of go without human knowledge,” Nature , vol. 550, pp. 354–, Oct. 2017. [Online]. Available: http://dx.doi.org/10.1038/nature24270
2017
Earlier work this paper cites.
D. Chan, “The ai that has nothing to learn from humans,” The Atlantic , vol. 7, no. 1, p. e1000858, 2017
2017
Earlier work this paper cites.
R. Lowe, Y. Wu, A. Tamar, J. Harb, P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in Proceedings of the 31st International Conference on Neural Information Processing Systems , 2017, pp. 6382–6393
2017
Earlier work this paper cites.
J. C. Peterson, J. T. Abbott, and T. L. Griffiths, “Adapting deep network features to capture psychological representations: An abridged report.” in IJCAI , 2017, pp. 4934–4938
2017
Cited alongside, same era.
A. Singh, T. Jain, and S. Sukhbaatar, “Learning when to communicate at scale in multiagent cooperative and competitive tasks,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
R. Iyer, Y. Li, H. Li, M. Lewis, R. Sundar, and K. Sycara, “Transparency and explanation in deep reinforcement learning neural networks,” in Proceedings of the 2018 AAAI/ACM Conference on AI, Ethics, and Society , 2018, pp. 144–150
2018
Cited alongside, same era.
A. R. Marathe, K. E. Schaefer, A. W. Evans, and J. S. Metcalfe, “Bidirectional communication for effective human-agent teaming,” in International Conference on Virtual, Augmented and Mixed Reality . Springer, 2018, pp. 338–350
2018
Cited alongside, same era.
R. Wang, X. He, R. Yu, W. Qiu, B. An, and Z. Rabinovich, “Learning efficient multi-agent communication: An information bottleneck approach,” in International Conference on Machine Learning . PMLR, 2020, pp. 9908–9918
2020
Later among the works it cites.
H. Mao, Z. Zhang, Z. Xiao, Z. Gong, and Y. Ni, “Learning agent communication under limited bandwidth by message pruning,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 04, 2020, pp. 5142–5149
2020
Later among the works it cites.
A. Agarwal, S. Kumar, K. Sycara, and M. Lewis, “Learning transferable cooperative behavior in multi-agent teams,” in Proceedings of the 19th International Conference on Autonomous Agents and MultiAgent Systems , 2020, pp. 1741–1743
2020
Later among the works it cites.
2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. L. Marlow, C. N. Lacerenza, J. Paoletti, C. S. Burke, and E. Salas, “Does team communication represent a one-size-fits-all approach?: A meta-analysis of team communication and performance,” Organizational behavior and human decision processes , vol. 144, pp. 145–170, 2018
2018
Cited alongside, same era.
T. Rashid, M. Samvelyan, C. Schroeder, G. Farquhar, J. Foerster, and S. Whiteson, “Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning,” in International Conference on Machine Learning . PMLR, 2018, pp. 4295–4304
2018
Cited alongside, same era.
I. Mordatch and P. Abbeel, “Emergence of grounded compositional language in multi-agent populations,” in Thirty-second AAAI conference on artificial intelligence , 2018
2018
Cited alongside, same era.
W. Van Winsum, “The effects of cognitive and visual workload on peripheral detection in the detection response task,” Human factors , vol. 60, no. 6, pp. 855–869, 2018
2018
Cited alongside, same era.
2019
Cited alongside, same era.
M. Carroll, R. Shah, M. K. Ho, T. Griffiths, S. Seshia, P. Abbeel, and A. Dragan, “On the utility of learning about humans for human-ai coordination,” Advances in Neural Information Processing Systems , vol. 32, pp. 5174–5185, 2019
2019
Cited alongside, same era.
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, J. Oh, D. Horgan, M. Kroiss, I. Danihelka, A. Huang, L. Sifre, T. Cai, J. P. Agapiou, M. Jaderberg, A. S. Vezhnevets, R. Leblond, T. Pohlen, V. Dalibard, D. Budden, Y. Sulsky, J. Molloy, T. L. Paine, C. Gulcehre, Z. Wang, T. Pfaff, Y. Wu, R. Ring, D. Yogatama, D. Wünsch, K. McKinney, O. Smith, T. Schaul, T. P. Lillicrap, K. Kavukcuoglu, D. Hassabis, C. Apps, and D. Silver, “Grandmaster level in starcraft ii using multi-agent reinforcement learning,” Nature , pp. 1–5, 2019
2019
Cited alongside, same era.
A. Anderson, J. Dodge, A. Sadarangani, Z. Juozapaitis, E. Newman, J. Irvine, S. Chattopadhyay, A. Fern, and M. Burnett, “Explaining reinforcement learning to mere mortals: An empirical study,” in Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, IJCAI-19 . International Joint Conferences on Artificial Intelligence Organization, 7 2019, pp. 1328–1334. [Online]. Available: https://doi.org/10.24963/ijcai.2019/184
2019
Cited alongside, same era.
Later among the works it cites.
B. Freed, R. James, G. Sartoretti, and H. Choset, “Sparse discrete communication learning for multi-agent cooperation through backpropagation,” in 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2020, pp. 7993–7998
2020
Later among the works it cites.
D. Hughes, A. Agarwal, Y. Guo, and K. Sycara, “Inferring non-stationary human preferences for human-agent teams,” in 2020 29th IEEE International Conference on Robot and Human Interactive Communication (RO-MAN) . IEEE, 2020, pp. 1178–1185
2020
Later among the works it cites.
E. M. van Zoelen, A. Cremers, F. P. Dignum, J. van Diggelen, and M. M. Peeters, “Learning to communicate proactively in human-agent teaming,” in International Conference on Practical Applications of Agents and Multi-Agent Systems . Springer, 2020, pp. 238–249
2020
Later among the works it cites.
S. Gronauer and K. Diepold, “Multi-agent deep reinforcement learning: a survey,” Artificial Intelligence Review , pp. 1–49, 2021
2021
Later among the works it cites.
K. Zhang, Z. Yang, and T. Başar, “Multi-agent reinforcement learning: A selective overview of theories and algorithms,” Handbook of Reinforcement Learning and Control , pp. 321–384, 2021
2021
Later among the works it cites.
H. C. Siu, J. Peña, E. Chen, Y. Zhou, V. Lopez, K. Palko, K. Chang, and R. Allen, “Evaluation of human-ai teams for learned and rule-based agents in hanabi,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Later among the works it cites.
T. Kliegr, Štěpán Bahník, and J. Fürnkranz, “A review of possible effects of cognitive biases on interpretation of rule-based machine learning models,” Artificial Intelligence , vol. 295, p. 103458, 2021. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0004370221000096
2021
Later among the works it cites.
2021
Later among the works it cites.
M. Tucker, H. Li, S. Agrawal, D. Hughes, K. Sycara, M. Lewis, and J. A. Shah, “Emergent discrete communication in semantic spaces,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Later among the works it cites.
S. Agrawal, “Learning to imitate, adapt and communicate,” Master’s thesis, Carnegie Mellon University, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
H. Li, T. Ni, S. Agrawal, F. Jia, S. Raja, Y. Gui, D. Hughes, M. Lewis, and K. Sycara, “Individualized mutual adaptation in human-agent teams,” IEEE Transactions on Human-Machine Systems , vol. 51, no. 6, pp. 706–714, 2021
2021
Later among the works it cites.
E. Seraj, Z. Wang, R. Paleja, D. Martin, M. Sklar, A. Patel, and M. Gombolay, “Learning efficient diverse communication for cooperative heterogeneous teaming,” in Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems , 2022, pp. 1173–1182
2022
Closest in time.