Fetching the paper…
Reading the bibliography…
In multi-agent reinforcement learning (MARL), effective communication improves agent performance, particularly under partial observability.
Lewis, D.: Convention: A philosophical study. Harvard University Press (1969)
1969
Earlier work this paper cites.
Seyfarth, R.M., Cheney, D.L., Marler, P.: Monkey responses to three different alarm calls: evidence of predator classification and semantic communication. Science 210
1980
Earlier work this paper cites.
Farrell, J., Rabin, M.: Cheap talk. Journal of Economic perspectives 10
1996
Earlier work this paper cites.
Cangelosi, A., Parisi, D.: The emergence of a’language’in an evolving population of neural networks. Connection Science 10
1998
Earlier work this paper cites.
Nowak, M.A., Krakauer, D.C.: The evolution of language. Proceedings of the National Academy of Sciences 96
1999
Earlier work this paper cites.
Rao, R.P., Ballard, D.H.: Predictive coding in the visual cortex: a functional interpretation of some extra-classical receptive-field effects. Nature neuroscience 2
1999
Earlier work this paper cites.
Hansen, E.A., Bernstein, D.S., Zilberstein, S.: Dynamic programming for partially observable stochastic games. In: AAAI. vol. 4, pp. 709–715 (2004)
2004
Earlier work this paper cites.
Friston, K., Kilner, J., Harrison, L.: A free energy principle for the brain. Journal of physiology-Paris 100
2006
Earlier work this paper cites.
Busoniu, L., Babuska, R., De Schutter, B.: A comprehensive survey of multiagent reinforcement learning. IEEE Transactions on Systems, Man, and Cybernetics, Part C (Applications and Reviews) 38
2008
Earlier work this paper cites.
Tomasello, M.: Why we cooperate. MIT press (2009)
2009
Earlier work this paper cites.
Mirolli, M., Parisi, D.: Producer Biases and Kin Selection in the Evolution of Communication, pp. 135–159. Springer Berlin Heidelberg, Berlin, Heidelberg (2010)
2010
Earlier work this paper cites.
Skyrms, B.: Signals: Evolution, learning, and information (2010)
2010
Earlier work this paper cites.
Tomasello, M.: Origins of human communication. MIT press (2010)
2010
Earlier work this paper cites.
2013
Earlier work this paper cites.
Kingma, D.P., Welling, M.: Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 (2013)
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
Foerster, J., Assael, I.A., De Freitas, N., Whiteson, S.: Learning to communicate with deep multi-agent reinforcement learning. Advances in neural information processing systems 29
2016
Cited alongside, same era.
2016
Cited alongside, same era.
Schulman, J., Moritz, P., Levine, S., Jordan, M., Abbeel, P.: High-dimensional continuous control using generalized advantage estimation. In: International Conference on Learning Representations (ICLR) (2016)
2016
Cited alongside, same era.
Sukhbaatar, S., Fergus, R., et al.: Learning multiagent communication with backpropagation. Advances in neural information processing systems 29
2016
Cited alongside, same era.
Lin, T., Huh, J., Stauffer, C., Lim, S.N., Isola, P.: Learning to ground multi-agent communication with autoencoders. Advances in Neural Information Processing Systems 34
2021
Later among the works it cites.
Gronauer, S., Diepold, K.: Multi-agent deep reinforcement learning: a survey. Artificial Intelligence Review 55
2022
Later among the works it cites.
Yu, C., Velu, A., Vinitsky, E., Gao, J., Wang, Y., Bayen, A., Wu, Y.: The surprising effectiveness of ppo in cooperative multi-agent games. Advances in neural information processing systems 35
2022
Later among the works it cites.
Ebara, H., Nakamura, T., Taniguchi, A., Taniguchi, T.: Multi-agent reinforcement learning with emergent communication using discrete and indifferentiable message. In: 2023 15th international congress on advanced applied informatics winter (IIAI-AAI-Winter). pp. 366–371. IEEE (2023)
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Higgins, I., Matthey, L., Pal, A., Burgess, C., Glorot, X., Botvinick, M., Mohamed, S., Lerchner, A.: beta-vae: Learning basic visual concepts with a constrained variational framework. In: International conference on learning representations (2017)
2017
Cited alongside, same era.
Lazaridou, A., Peysakhovich, A., Baroni, M.: Multi-agent cooperation and the emergence of (natural) language. In: International Conference on Learning Representations (2017)
2017
Cited alongside, same era.
Lowe, R., Wu, Y.I., Tamar, A., Harb, J., Pieter Abbeel, O., Mordatch, I.: Multi-agent actor-critic for mixed cooperative-competitive environments. Advances in neural information processing systems 30
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Lazaridou, A., Hermann, K.M., Tuyls, K., Clark, S.: Emergence of linguistic communication from referential games with symbolic and pixel input. In: International Conference on Learning Representations (2018)
2018
Cited alongside, same era.
Sunehag, P., Lever, G., Gruslys, A., Czarnecki, W.M., Zambaldi, V., Jaderberg, M., Lanctot, M., Sonnerat, N., Leibo, J.Z., Tuyls, K., et al.: Value-decomposition networks for cooperative multi-agent learning based on team reward. In: Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems. pp. 2085–2087 (2018)
2018
Cited alongside, same era.
Hagiwara, Y., Kobayashi, H., Taniguchi, A., Taniguchi, T.: Symbol emergence as an interpersonal multimodal categorization. Frontiers in Robotics and AI 6
2019
Cited alongside, same era.
Lowe, R., Foerster, J., Boureau, Y.L., Pineau, J., Dauphin, Y.: On the pitfalls of measuring emergent communication. In: Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems. pp. 693–701 (2019)
2019
Cited alongside, same era.
2023
Later among the works it cites.
Taniguchi, T., Yoshida, Y., Matsui, Y., Le Hoang, N., Taniguchi, A., Hagiwara, Y.: Emergent communication through metropolis-hastings naming game with deep generative models. Advanced Robotics 37
2023
Later among the works it cites.
Wong, A., Bäck, T., Kononova, A.V., Plaat, A.: Deep multiagent reinforcement learning: challenges and directions. Artificial Intelligence Review 56
2023
Later among the works it cites.
Albrecht, S.V., Christianos, F., Schäfer, L.: Multi-agent reinforcement learning: Foundations and modern approaches. MIT Press (2024)
2024
Later among the works it cites.
Hoang, N.L., Taniguchi, T., Hagiwara, Y., Taniguchi, A.: Emergent communication of multimodal deep generative models based on metropolis-hastings naming game. Frontiers in Robotics and AI 10
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Taniguchi, T.: Collective predictive coding hypothesis: Symbol emergence as decentralized bayesian inference. Frontiers in Robotics and AI 11
2024
Later among the works it cites.
Ueda, R., Taniguchi, T.: Lewis’s signaling game as beta-vae for natural word lengths and segments. In: The Twelfth International Conference on Learning Representations (2024)
2024
Later among the works it cites.
Zhu, C., Dastani, M., Wang, S.: A survey of multi-agent deep reinforcement learning with communication. Autonomous Agents and Multi-Agent Systems 38
2024
Later among the works it cites.