Fetching the paper…
Reading the bibliography…
Various applications for inter-machine communications are on the rise.
S. Arimoto, “An algorithm for computing the capacity of arbitrary discrete memoryless channels,”
1972
Earlier work this paper cites.
H. Witsenhausen, “Indirect rate distortion problems,”
1980
Earlier work this paper cites.
Y. Linde, A. Buzo, and R. Gray, “An algorithm for vector quantizer design,”
1980
Earlier work this paper cites.
G. E. Monahan, “State of the art—a survey of partially observable markov decision processes: theory, models, and algorithms,”
1982
Earlier work this paper cites.
S. Lloyd, “Least squares quantization in pcm,”
1982
Earlier work this paper cites.
P. Ioannou and J. Sun, “Theory and design of robust direct and indirect adaptive-control schemes,”
1988
Earlier work this paper cites.
D. P. Bertsekas and D. A. Castanon, “Adaptive aggregation methods for infinite horizon dynamic programming,”
1989
Earlier work this paper cites.
G. Rubino, “On weak lumpability in markov chains,”
1989
Earlier work this paper cites.
T. Jaakkola, M. I. Jordan, and S. P. Singh, “Convergence of stochastic iterative dynamic programming algorithms,” in
1994
Earlier work this paper cites.
C. Boutilier, “Multiagent systems: Challenges and opportunities for decision-theoretic planning,”
1999
Earlier work this paper cites.
M. Lauer and M. A. Riedmiller, “An algorithm for distributed reinforcement learning in cooperative multi-agent systems,” in
2000
Earlier work this paper cites.
P. Xuan, V. Lesser, and S. Zilberstein, “Communication decisions in multi-agent cooperation: Model and experiments,” in
2001
Earlier work this paper cites.
D. V. Pynadath and M. Tambe, “The communicative multiagent team decision problem: Analyzing teamwork theories and models,”
2002
Earlier work this paper cites.
G. N. Nair and R. J. Evans, “Exponential stabilisability of finite-dimensional linear systems with limited data rates,”
2003
Earlier work this paper cites.
——, “Stabilizability of stochastic linear systems with finite feedback data rates,”
2004
Earlier work this paper cites.
F. Fischer, M. Rovatsos, and G. Weiss, “Hierarchical reinforcement learning in communication-mediated multiagent coordination,” in
2004
Earlier work this paper cites.
F. A. Oliehoek, M. T. Spaan, N. Vlassis
2007
Earlier work this paper cites.
T. Kasai, H. Tenmoto, and A. Kamiya, “Learning of communication codes in multi-agent reinforcement learning problem,” in
2008
Earlier work this paper cites.
F. A. Oliehoek, M. T. Spaan, and N. Vlassis, “Optimal and approximate q-value functions for decentralized pomdps,”
2008
Earlier work this paper cites.
C. Amato, J. S. Dibangoye, and S. Zilberstein, “Incremental policy generation for finite-horizon dec-pomdps,” in
2009
Cited alongside, same era.
F. Wu, S. Zilberstein, and X. Chen, “Online planning for multi-agent systems with bounded communication,”
2011
Cited alongside, same era.
M. G. Azar, R. Munos, M. Ghavamzadaeh, and H. J. Kappen, “Speedy q-learning,” 2011
2011
Cited alongside, same era.
C. Zhang and V. Lesser, “Coordinating multi-agent reinforcement learning with limited communication,” in
2013
Cited alongside, same era.
S. Yüksel, “Jointly optimal lqg quantization and control policies for multi-dimensional systems,”
2013
Cited alongside, same era.
J. Foerster, Y. Assael, N. de Freitas, and S. Whiteson, “Learning to communicate with deep multi-agent reinforcement learning,” in
L. Chaari, M. Fourati, and J. Rezgui, “Heterogeneous lorawan & leo satellites networks concepts, architectures and future directions,” in
2019
Later among the works it cites.
V. Kostina and B. Hassibi, “Rate-cost tradeoffs in control,”
2019
Later among the works it cites.
A. Mostaani, O. Simeone, S. Chatzinotas, and B. Ottersten, “Learning-based physical layer communications for multiagent collaboration,” in
2019
Later among the works it cites.
D. Kim, S. Moon, D. Hostallero, W. J. Kang, T. Lee, K. Son, and Y. Yi, “Learning to schedule communication in multi-agent reinforcement learning,” in
2019
Later among the works it cites.
R. Lowe, J. Foerster, Y.-L. Boureau, J. Pineau, and Y. Dauphin, “On the pitfalls of measuring emergent communication,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
D. Abel, D. Hershkowitz, and M. Littman, “Near optimal behavior via approximate state abstraction,” in
2016
Cited alongside, same era.
S. Sukhbaatar, R. Fergus
2016
Cited alongside, same era.
F. Heylighen, “Stigmergy as a universal coordination mechanism i: Definition and components,”
2016
Cited alongside, same era.
F. A. Oliehoek, C. Amato
2016
Cited alongside, same era.
A. Barel, R. Manor, and A. M. Bruckstein, “Come together: Multi-agent geometric consensus,”
2017
Cited alongside, same era.
R. S. Sutton and A. G. Barto,
2017
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
N. Shlezinger and Y. C. Eldar, “Task-based quantization with application to mimo receivers,”
2020
Closest in time.
D. Lee, N. He, P. Kamalaruban, and V. Cevher, “Optimization for reinforcement learning: From a single agent to cooperative agents,”
2020
Closest in time.
A. Mostaani, T. X. Vu, S. Chatzinotas, and B. Ottersten, “State aggregation for multiagent communication over rate-limited channels,” in
2020
Closest in time.
2021
Closest in time.
E. C. Strinati and S. Barbarossa, “6g networks: Beyond shannon towards semantic and goal-oriented communications,”
2021
Closest in time.
2021
Closest in time.
N. Shlezinger and Y. C. Eldar, “Deep task-based quantization,”
2021
Closest in time.
D. Gunduz, Z. Qin, I. E. Aguerri, H. S. Dhillon, Z. Yang, A. Yener, K. Kit Wong, and C.-B. Chae, “Beyond transmitting bits: Context, semantics, and task-oriented communications,”
2022
Closest in time.
H. Xie, Z. Qin, X. Tao, and K. B. Letaief, “Task-oriented multi-user semantic communications,”
2022
Closest in time.
M. M. Azari, S. Solanki, S. Chatzinotas, O. Kodheli, H. Sallouha, A. Colpaert, J. F. M. Montoya, S. Pollin, A. Haqiqatnejad, A. Mostaani
2022
Closest in time.
P. A. Stavrou and M. Kountouris, “A rate distortion approach to goal-oriented communication,” 2022
2022
Closest in time.