Fetching the paper…
Reading the bibliography…
In this paper, we consider the multi-armed bandit problem with high-dimensional features.
Regression shrinkage and selection via the lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Variable selection via nonconcave penalized likelihood and its oracle properties
Jianqing Fan and Runze Li · 2001
Earlier work this paper cites.
Reinforcement learning with immediate rewards and linear hypotheses
Naoki Abe, Alan W. Biermann, and Philip M. Long · 2003
Earlier work this paper cites.
Using confidence bounds for exploitation-exploration trade-offs
Peter Auer · 2003
Earlier work this paper cites.
Decoding by linear programming
Emmanuel Candes and Terence Tao · 2005
Earlier work this paper cites.
Fast learning rates for plug-in classifiers
Jean-Yves Audibert, Alexandre B Tsybakov, et al · 2007
Earlier work this paper cites.
Stochastic linear optimization under bandit feedback
Varsha Dani, Thomas P. Hayes, and Sham M. Kakade · 2008
Earlier work this paper cites.
Estimation of the warfarin dose with clinical and pharmacogenetic data
International Warfarin Pharmacogenetics Consortium · 2009
Earlier work this paper cites.
Woodroofe’s one-armed bandit problem revisited
Alexander Goldenshluger and Assaf Zeevi · 2009
Cited alongside, same era.
Linearly parameterized bandits
Paat Rusmevichientong and John N. Tsitsiklis · 2010
Cited alongside, same era.
Nearly unbiased variable selection under minimax concave penalty
Cun-Hui Zhang et al · 2010
Cited alongside, same era.
Improved algorithms for linear stochastic bandits
Yasin Abbasi-Yadkori, Dávid Pál, and Casaba Szepesvári · 2011
Cited alongside, same era.
Statistics for high-dimensional data: methods, theory and applications
Peter Bühlmann and Sara Van De Geer · 2011
Cited alongside, same era.
Contextual bandits with linear payoff functions
Wei Chu, Lihong Li, Lev Reyzin, and Robert E Schapire · 2011
Cited alongside, same era.
A linear response bandit problem
Alexander Goldenshluger and Assaf Zeevi · 2013
Later among the works it cites.
Reconstruction from anisotropic random measurements
Mark Rudelson and Shuheng Zhou · 2013
Later among the works it cites.
The lower tail of random quadratic forms, with applications to ordinary least squares and restricted eigenvalue properties
Roberto Imbuzeiro Oliveira · 2016
Later among the works it cites.
Minimax concave penalized multi-armed bandit model with high-dimensional covariates
Xue Wang, Mingcheng Wei, and Tao Yao · 2018
Later among the works it cites.
Online decision making with high-dimensional covariates
Hamsa Bastani and Mohsen Bayati · 2019
Later among the works it cites.
Doubly-robust lasso bandit
Gi-Soo Kim and Myunghee Cho Paik · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Online-to-confidence-set conversions and application to sparse stochastic bandits
Yasin Abbasi-Yadkori, Dávid Pál, and Casaba Szepesvári · 2012
Cited alongside, same era.
Introduction to the non-asymptotic analysis of random matrices
Roman Vershynin · 2012
Cited alongside, same era.
Finite-time analysis of the multiarmed bandit problem
Peter Auer, Nicolò Cesa-Bianchi, and Paul Fischer
Cited in the paper.
High-Dimensional Statistics: A Non-Asymptotic Viewpoint
Martin J. Wainwright · 2019
Later among the works it cites.