Fetching the paper…
Reading the bibliography…
We study the problem of learning shared structure \emph{across} a sequence of dynamic pricing experiments for related products.
Meta-learning for contextual bandit exploration
Sharaf, Amr, Hal Daumé III. 2019 · 1901
Earlier work this paper cites.
A short note on concentration inequalities for random vectors with subgaussian norm
Jin, Chi, Praneeth Netrapalli, Michael I. Jordan. 2019 · 1902
Earlier work this paper cites.
Data-pooling in stochastic optimization
Gupta, Vishal, Nathan Kallus. 2020 · 1906
Earlier work this paper cites.
On the likelihood that one unknown probability exceeds another in view of the evidence of two samples
Thompson, William R. 1933 · 1933
Earlier work this paper cites.
Asymptotically efficient adaptive allocation rules
Lai, Tze Leung, Herbert Robbins. 1985 · 1985
Earlier work this paper cites.
Matrix Variate Distributions
Gupta, A. K., D. K. Nagar. 1999 · 1999
Earlier work this paper cites.
Eligibility traces for off-policy policy evaluation
Precup, Doina, Richard Sutton, Satinder Singh. 2000 · 2000
Earlier work this paper cites.
Marginal mean models for dynamic regimes
Murphy, Susan, Mark van der Laan, James Robins, CPPRG. 2001 · 2001
Earlier work this paper cites.
Using confidence bounds for exploitation-exploration trade-offs
Auer, Peter. 2002 · 2002
Earlier work this paper cites.
The value of knowing a demand curve: Bounds on regret for online posted-price auctions
Kleinberg, Robert, Tom Leighton. 2003 · 2003
Earlier work this paper cites.
Matrix Analysis for Scientists and Engineers
Laub, Alan. 2004 · 2004
Earlier work this paper cites.
Bayesian Statistics and Marketing
Rossi, Peter E., Greg M. Allenby, Robert McCulloch. 2005 · 2005
Earlier work this paper cites.
Bayesian clinical trials
Berry, Donald A. 2006 · 2006
Earlier work this paper cites.
Pattern Recognition and Machine Learning
Bishop, Christopher M. 2006 · 2006
Earlier work this paper cites.
Prediction, Learning, and Games
Cesa-Bianchi, Nicolò, Gábor Lugosi. 2006 · 2006
Earlier work this paper cites.
On worst-case regret of linear thompson sampling
Hamidi, Nima, Mohsen Bayati. 2020 · 2006
Earlier work this paper cites.
Multi-armed bandit, dynamic environments and meta-bandits
Hartland, Cédric, Sylvain Gelly, Nicolas Baskiotis, Olivier Teytaud, Michéle Sebag. 2006 · 2006
Earlier work this paper cites.
Constructing informative priors using transfer learning
Raina, Rajat, Andrew Y Ng, Daphne Koller. 2006 · 2006
Earlier work this paper cites.
Stochastic linear optimization under bandit feedback
Dani, Varsha, Thomas Hayes, Sham Kakade. 2008 · 2008
Earlier work this paper cites.
Dynamic pricing for nonperishable products with demand learning
Araman, Victor F, René Caldentey. 2009 · 2009
Earlier work this paper cites.
Dynamic pricing without knowing the demand function: Risk bounds and near-optimal algorithms
Besbes, Omar, Assaf Zeevi. 2009 · 2009
Earlier work this paper cites.
Introduction to Algorithms
Cormen, Thomas H., Charles E. Leiserson, Ronald L. Rivest, Clifford Stein. 2009 · 2009
Earlier work this paper cites.
Dynamic pricing with a prior on market response
Farias, Vivek F, Benjamin Van Roy. 2010 · 2010
Earlier work this paper cites.
A contextual-bandit approach to personalized news article recommendation
Li, Lihong, Wei Chu, John Langford, Robert Schapire. 2010 · 2010
Earlier work this paper cites.
Linearly parameterized bandits
Rusmevichientong, Paat, John N Tsitsiklis. 2010 · 2010
Cited alongside, same era.
Improved algorithms for linear stochastic bandits
Abbasi-Yadkori, Yasin, David Pál, Csaba. Szepesvári. 2011 · 2011
Cited alongside, same era.
An empirical evaluation of thompson sampling
Chapelle, Olivier, Lihong Li. 2011 · 2011
Cited alongside, same era.
User-friendly tail bounds for matrix martingales
Tropp, Joel. 2011 · 2011
Cited alongside, same era.
Dynamic pricing under a general parametric choice model
Broder, Josef, Paat Rusmevichientong. 2012 · 2012
Cited alongside, same era.
Bayesian dynamic pricing policies: Learning and earning under a binary prior distribution
Harrison, J Michael, N Bora Keskin, Assaf Zeevi. 2012 · 2012
Cited alongside, same era.
Feature-based dynamic pricing
Cohen, Maxime, Ilan Lobel, Renato Paes Leme. 2016 · 2016
Later among the works it cites.
Dynamic pricing with demand covariates
Qiang, Sheng, Mohsen Bayati. 2016 · 2016
Later among the works it cites.
Linear thompson sampling revisited
Abeille, Marc, Alessandro Lazaric. 2017 · 2017
Later among the works it cites.
Personalized dynamic pricing with machine learning
Ban, Gah-Yi, N Bora Keskin. 2017 · 2017
Later among the works it cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, Chelsea, Pieter Abbeel, Sergey Levine. 2017 · 2017
Later among the works it cites.
Competition-based dynamic pricing in online retailing: A methodology validated with field experiments
Fisher, Marshall, Santiago Gallino, Jun Li. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Meta-learning of exploration/exploitation strategies: The multi-armed bandit case
Maes, Francis, Louis Wehenkel, Damien Ernst. 2012 · 2012
Cited alongside, same era.
Thompson sampling for contextual bandits with linear payoffs
Agrawal, Shipra, Navin Goyal. 2013 · 2013
Cited alongside, same era.
Prior-free and prior-dependent regret bounds for thompson sampling
Bubeck, Sébastien, Che-Yu Liu. 2013 · 2013
Cited alongside, same era.
Simultaneously learning and optimizing using controlled variance pricing
den Boer, Arnoud V, Bert Zwart. 2013 · 2013
Cited alongside, same era.
Bandits with concave rewards and convex knapsacks
Agrawal, Shipra, Nikhil R Devanur. 2014 · 2014
Cited alongside, same era.
Stochastic multi-armed-bandit problem with non-stationary rewards
Besbes, Omar, Yonatan Gur, Assaf Zeevi. 2014 · 2014
Cited alongside, same era.
Lecture notes on advanced statistical theory
Rinaldo, Alessandro. 2017 · 2017
Later among the works it cites.
How does dynamic pricing affect customer behavior on retailing platforms? evidence from a large randomized experiment on alibaba
Zhang, Dennis J, Hengchen Dai, Lingxiu Dong, Fangfang Qi, Nannan Zhang, Xiaofei Liu, Zhongyi Liu. 2017 · 2017
Later among the works it cites.
Online network revenue management using thompson sampling
Ferreira, Kris, David Simchi-Levi, He Wang. 2018 · 2018
Later among the works it cites.
Probabilistic model-agnostic meta-learning
Finn, Chelsea, Kelvin Xu, Sergey Levine. 2018 · 2018
Later among the works it cites.
High Dimensional Statistics
Rigollet, R., J. Hütter. 2018 · 2018
Later among the works it cites.
A tutorial on thompson sampling
Russo, Daniel J, Benjamin Van Roy, Abbas Kazerouni, Ian Osband, Zheng Wen, et al. 2018 · 2018
Later among the works it cites.
Regret bounds for meta bayesian optimization with an unknown gaussian process prior
Wang, Zi, Beomjoon Kim, Leslie Pack Kaelbling. 2018 · 2018
Later among the works it cites.
Bayesian model-agnostic meta-learning
Yoon, Jaesik, Taesup Kim, Ousmane Dia, Sungwoong Kim, Yoshua Bengio, Sungjin Ahn. 2018 · 2018
Later among the works it cites.
Learning to route efficiently with end-to-end feedback: The value of networked structure
Zhu, Ruihao, Eytan Modiano. 2018 · 2018
Later among the works it cites.
Adaptive clinical trial designs with surrogates: When should we bother?
Anderer, Arielle, Hamsa Bastani, John Silberholz. 2019 · 2019
Closest in time.
Near optimal ab testing
Bhat, Nikhil, Vivek F Farias, Ciamac C Moallemi, Deeksha Sinha. 2019 · 2019
Closest in time.
Dynamic pricing in high-dimensions
Javanmard, Adel, Hamid Nazerzadeh. 2019 · 2019
Closest in time.
High-Dimensional Statistics: A Non-Asymptotic Viewpoint
Wainwright, Martin. 2019 · 2019
Closest in time.
Designing and evaluating dynamic pricing policies for major league baseball tickets
Xu, Joseph, Peter Fader, Senthil K Veeraraghavan. 2019 · 2019
Closest in time.
Predicting with proxies: Transfer learning in high dimension
Bastani, Hamsa. 2020 · 2020
Closest in time.
Mostly exploration-free algorithms for contextual bandits
Bastani, Hamsa, Mohsen Bayati, Khashayar Khosravi. 2020 · 2020
Closest in time.
The value of personalized pricing
Elmachtoub, Adam N., Vishal Gupta, Michael Hamilton. 2020 · 2020
Closest in time.