Fetching the paper…
Reading the bibliography…
In light of the COVID-19 pandemic, it is an open challenge and critical practical problem to find a optimal way to dynamically prescribe the best policies that balance both the governmental resources and epidemic control in different countries and regions.
“On the likelihood that one unknown probability exceeds another in view of the evidence of two samples,”
William R Thompson, · 1933
Earlier work this paper cites.
“Bandit processes and dynamic allocation indices,”
J. C. Gittins, · 1979
Earlier work this paper cites.
“Asymptotically efficient adaptive allocation rules,”
T. L. Lai and Herbert Robbins, · 1985
Earlier work this paper cites.
“Asymptotically efficient adaptive allocation rules,”
T. L. Lai and Herbert Robbins, · 1985
Earlier work this paper cites.
“Finite-time analysis of the multiarmed bandit problem,”
Peter Auer, Nicolò Cesa-Bianchi, and Paul Fischer, · 2002
Earlier work this paper cites.
“The nonstochastic multiarmed bandit problem,”
Peter Auer, Nicolo Cesa-Bianchi, Yoav Freund, and Robert E Schapire, · 2002
Earlier work this paper cites.
“Epoch-greedy algorithm for multi-armed bandits with side information,”
John Langford and Tong Zhang, · 2007
Earlier work this paper cites.
“Explore/exploit schemes for web content optimization,”
Deepak Agarwal, Bee-Chung Chen, and Pradheep Elango, · 2009
Earlier work this paper cites.
“A contextual-bandit approach to personalized news article recommendation,”
Lihong Li, Wei Chu, John Langford, and Robert E. Schapire, · 2010
Earlier work this paper cites.
“Improved algorithms for linear stochastic bandits,”
Yasin Abbasi-Yadkori, Dávid Pál, and Csaba Szepesvári, · 2011
Earlier work this paper cites.
“Unbiased offline evaluation of contextual-bandit-based news article recommendation algorithms.,”
Lihong Li, Wei Chu, John Langford, and Xuanhui Wang, · 2011
Earlier work this paper cites.
“Efficient learning with partially observed attributes,”
Nicolò Cesa-Bianchi, Shai Shalev-Shwartz, and Ohad Shamir, · 2011
Cited alongside, same era.
“Thompson sampling for contextual bandits with linear payoffs,”
Shipra Agrawal and Navin Goyal, · 2013
Cited alongside, same era.
“A neural networks committee for the contextual bandit problem,”
Robin Allesiardo, Raphaël Féraud, and Djallel Bouneffouf, · 2014
Cited alongside, same era.
“Contextual combinatorial bandit and its application on diversified online recommendation,”
Lijing Qin, Shouyuan Chen, and Xiaoyan Zhu, · 2014
Cited alongside, same era.
“Multi-armed bandit problem with known trend,”
Djallel Bouneffouf and Raphaël Féraud, · 2016
Cited alongside, same era.
“Model-informed risk assessment for zika virus outbreaks in the asia-pacific regions,”
Yue Teng, Dehua Bi, Guigang Xie, Yuan Jin, Yong Huang, Baihan Lin, Xiaoping An, Yigang Tong, and Dan Feng, · 2017
“Incorporating behavioral constraints in online AI systems,”
Avinash Balakrishnan, Djallel Bouneffouf, Nicholas Mattei, and Francesca Rossi, · 2019
Later among the works it cites.
“Using multi-armed bandits to learn ethical priorities for online AI systems,”
Avinash Balakrishnan, Djallel Bouneffouf, Nicholas Mattei, and Francesca Rossi, · 2019
Later among the works it cites.
“Online learning in iterated prisoner’s dilemma to mimic human behavior,”
Baihan Lin, Djallel Bouneffouf, and Guillermo Cecchi, · 2020
Later among the works it cites.
“Speaker diarization as a fully online learning problem in minivox,”
Baihan Lin and Xinxin Zhang, · 2020
Later among the works it cites.
“VoiceID on the fly: A speaker recognition system that learns from scratch,”
Baihan Lin and Xinxin Zhang, · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Dynamic forecasting of zika epidemics using google trends,”
Yue Teng, Dehua Bi, Guigang Xie, Yuan Jin, Yong Huang, Baihan Lin, Xiaoping An, Dan Feng, and Yigang Tong, · 2017
Cited alongside, same era.
“Context attentive bandits: Contextual bandit with restricted context,”
Djallel Bouneffouf, Irina Rish, Guillermo A. Cecchi, and Raphaël Féraud, · 2017
Cited alongside, same era.
“Contextual bandit with adaptive feature extraction,”
Baihan Lin, Djallel Bouneffouf, Guillermo Cecchi, and Irina Rish, · 2018
Cited alongside, same era.
“A survey on practical applications of multi-armed and contextual bandits,”
Djallel Bouneffouf and Irina Rish, · 2019
Cited alongside, same era.
“Unified models of human behavioral agents in bandits, contextual bandits and RL,”
Baihan Lin, Guillermo Cecchi, Djallel Bouneffouf, Jenna Reinen, and Irina Rish, · 2020
Later among the works it cites.
“Online semi-supervised learning in contextual bandits with episodic reward,”
Baihan Lin, · 2020
Later among the works it cites.
“Models of human behavioral agents in bandits, contextual bandits and rl,”
Baihan Lin, Guillermo Cecchi, Djallel Bouneffouf, Jenna Reinen, and Irina Rish, · 2021
Closest in time.
“Speaker diarization as a fully online bandit learning problem in minivox,”
Baihan Lin and Xinxin Zhang, · 2021
Closest in time.
“Evolutionary multi-armed bandits with genetic thompson sampling,”
Baihan Lin, · 2022
Closest in time.