Fetching the paper…
Reading the bibliography…
We study a decentralized cooperative stochastic multi-armed bandit problem with $K$ arms on a network of $N$ agents.
Asymptotically efficient adaptive allocation rules
Tze Leung Lai and Herbert Robbins · 1985
Earlier work this paper cites.
Matrix analysis
Roger A Horn and Charles R Johnson · 1990
Earlier work this paper cites.
Finite-time analysis of the multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, and Paul Fischer · 2002
Earlier work this paper cites.
Fast linear iterations for distributed averaging
Lin Xiao and Stephen Boyd · 2004
Earlier work this paper cites.
Randomized gossip algorithms
Stephen Boyd, Arpita Ghosh, Balaji Prabhakar, and Devavrat Shah · 2006
Earlier work this paper cites.
Competitive collaborative learning
Baruch Awerbuch and Robert Kleinberg · 2008
Earlier work this paper cites.
Enforcing consensus while monitoring the environment in wireless sensor networks
Paolo Braca, Stefano Marano, and Vincenzo Matta · 2008
Earlier work this paper cites.
Minimax policies for adversarial and stochastic bandits
Jean-Yves Audibert and Sébastien Bubeck · 2009
Earlier work this paper cites.
Distributed subgradient methods for multi-agent optimization
Angelia Nedic and Asuman Ozdaglar · 2009
Earlier work this paper cites.
Gossip algorithms
Devavrat Shah et al · 2009
Earlier work this paper cites.
Gossip algorithms for distributed signal processing
Alexandros G Dimakis, Soummya Kar, José MF Moura, Michael G Rabbat, and Anna Scaglione · 2010
Earlier work this paper cites.
Learning multiuser channel allocations in cognitive radio networks: A combinatorial multi-armed bandit formulation
Yi Gai, Bhaskar Krishnamachari, and Rahul Jain · 2010
Earlier work this paper cites.
Distributed learning in multi-armed bandit with multiple players
Keqin Liu and Qing Zhao · 2010
Earlier work this paper cites.
Distributed algorithms for learning and cognitive medium access with logarithmic regret
Animashree Anandkumar, Nithin Michael, Ao Kevin Tang, and Ananthram Swami · 2011
Cited alongside, same era.
Iterative solution of large linear systems
W Auzinger and J Melenk · 2011
Cited alongside, same era.
Bandit problems in networks: Asymptotically efficient distributed allocation rules
Soummya Kar, H Vincent Poor, and Shuguang Cui · 2011
Cited alongside, same era.
Regret analysis of stochastic and nonstochastic multi-armed bandit problems
Sébastien Bubeck and Nicolo Cesa-Bianchi · 2012
Cited alongside, same era.
Dual averaging for distributed optimization: Convergence analysis and network scaling
John C Duchi, Alekh Agarwal, and Martin J Wainwright · 2012
Cited alongside, same era.
Dcops and bandits: Exploration and exploitation in decentralised coordination
Chebyshev acceleration of iterative refinement
Mario Arioli and J Scott · 2014
Later among the works it cites.
Decentralized learning for multiplayer multiarmed bandits
Dileep Kalathil, Naumaan Nayyar, and Rahul Jain · 2014
Later among the works it cites.
Distributed multi-agent online learning based on global feedback
Jie Xu, Cem Tekin, Simpson Zhang, and Mihaela Van Der Schaar · 2015
Later among the works it cites.
Delay and cooperation in nonstochastic bandits
Nicolo Cesa-Bianchi, Claudio Gentile, Yishay Mansour, and Alberto Minora · 2016
Later among the works it cites.
Distributed clustering of linear bandits in peer to peer networks
Nathan Korda, Balázs Szörényi, and Li Shuai · 2016
Later among the works it cites.
Distributed cooperative decision-making in multiarmed bandits: frequentist and bayesian algorithms
Peter Landgren, Vaibhav Srivastava, and Naomi Ehrich Leonard · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ruben Stranders, Long Tran-Thanh, Francesco M Delle Fave, Alex Rogers, and Nicholas R Jennings · 2012
Cited alongside, same era.
Online learning in decentralized multi-user spectrum access with synchronized explorations
Cem Tekin and Mingyan Liu · 2012
Cited alongside, same era.
Long-term information collection with energy harvesting wireless sensors: a multi-armed bandit based approach
Long Tran-Thanh, Alex Rogers, and Nicholas R Jennings · 2012
Cited alongside, same era.
Multi-armed bandits in the presence of side observations in social networks
Swapna Buccapatnam, Atilla Eryilmaz, and Ness B Shroff · 2013
Cited alongside, same era.
Distributed exploration in multi-armed bandits
Eshcar Hillel, Zohar S Karnin, Tomer Koren, Ronny Lempel, and Oren Somekh · 2013
Cited alongside, same era.
A classical introduction to modern number theory , volume 84
Kenneth Ireland and Michael Rosen · 2013
Cited alongside, same era.
Online learning under delayed feedback
Pooria Joulani, Andras Gyorgy, and Csaba Szepesvári · 2013
Cited alongside, same era.
Later among the works it cites.
On distributed cooperative decision-making in multiarmed bandits
Peter Landgren, Vaibhav Srivastava, and Naomi Ehrich Leonard · 2016
Later among the works it cites.
On regret-optimal learning in decentralized multi-player multi-armed bandits
Naumaan Nayyar, Dileep Kalathil, and Rahul Jain · 2016
Later among the works it cites.
Batched bandit problems
Vianney Perchet, Philippe Rigollet, Sylvain Chassang, Erik Snowberg, et al · 2016
Later among the works it cites.
Coordinated versus decentralized exploration in multi-agent multi-armed bandits
Mithun Chakraborty, Kai Yee Phoebe Chua, Sanmay Das, and Brendan Juba · 2017
Later among the works it cites.
Optimal algorithms for smooth and strongly convex distributed optimization in networks
Kevin Scaman, Francis Bach, Sébastien Bubeck, Yin Tat Lee, and Laurent Massoulié · 2017
Later among the works it cites.
Multi-armed bandits in multi-agent networks
Shahin Shahrampour, Alexander Rakhlin, and Ali Jadbabaie · 2017
Later among the works it cites.