Fetching the paper…
Reading the bibliography…
We consider a firm that sells products over $T$ periods without knowing the demand function.
Adaptive design and stochastic approximation
T L_ Lai and Herbert Robbins · 1979
Earlier work this paper cites.
A one-armed bandit problem with a concomitant variable
Michael Woodroofe · 1979
Earlier work this paper cites.
Probability and measure
P. Billingsley · 1979
Earlier work this paper cites.
Iterated least squares in multiperiod control
TL Lai and Herbert Robbins · 1982
Earlier work this paper cites.
Asymptotically efficient adaptive allocation rules
Tze Leung Lai and Herbert Robbins · 1985
Earlier work this paper cites.
One-armed bandit problems with covariates
Jyotirmoy Sarkar · 1991
Earlier work this paper cites.
On consistency of bayes estimates in a certainty equivalence adaptive system
Kani Chen and Inchi Hu · 1998
Earlier work this paper cites.
The nonstochastic multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, Yoav Freund, and Robert E Schapire · 2002
Earlier work this paper cites.
Using confidence bounds for exploitation-exploration trade-offs
Peter Auer · 2003
Earlier work this paper cites.
Optimal pricing mechanisms with unknown demand
Ilya Segal · 2003
Earlier work this paper cites.
The value of knowing a demand curve: Bounds on regret for online posted-price auctions
Robert Kleinberg and Tom Leighton · 2003
Earlier work this paper cites.
Learning and pricing in an internet environment with binomial demands
Alexandre X Carvalho and Martin L Puterman · 2005
Earlier work this paper cites.
A practical inventory control policy using operational statistics
Liwan H. Liyanage and J.George Shanthikumar · 2005
Earlier work this paper cites.
The epoch-greedy algorithm for contextual multi-armed bandits
John Langford and Tong Zhang · 2007
Earlier work this paper cites.
Performance limitations in bandit problems with side observations
Alexander Goldenshluger and Assaf Zeevi · 2007
Earlier work this paper cites.
Dynamic pricing for nonperishable products with demand learning
Victor F. Araman and René Caldentey · 2009
Cited alongside, same era.
Dynamic pricing without knowing the demand function: Risk bounds and near-optimal algorithms
Omar Besbes and Assaf Zeevi · 2009
Cited alongside, same era.
Nonparametric bandits with covariates
Philippe Rigollet and Assaf Zeevi · 2010
Cited alongside, same era.
User-friendly tail bounds for matrix martingales
Joel A Tropp · 2011
Cited alongside, same era.
On the minimax complexity of pricing in a changing environment
Omar Besbes and Assaf Zeevi · 2011
Cited alongside, same era.
Simultaneously learning and optimizing using controlled variance pricing
Arnoud V. den Boer and Bert Zwart · 2014
Later among the works it cites.
Dynamic pricing with an unknown demand model: Asymptotically optimal semi-myopic policies
N Bora Keskin and Assaf Zeevi · 2014
Later among the works it cites.
Choosing a good toolkit: An essay in behavioral economics
David M. Kreps and Alejandro Francetich · 2014
Later among the works it cites.
Contextual-bandit approach to personalized news article recommendation, August 25 2014
Lihong Li, Wei Chu, John Langford, and Robert Schapire · 2014
Later among the works it cites.
Resourceful contextual bandits
Ashwinkumar Badanidiyuru, John Langford, and Aleksandrs Slivkins · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Miroslav Dudik, Daniel Hsu, Satyen Kale, Nikos Karampatziakis, John Langford, Lev Reyzin, and Tong Zhang · 2011
Cited alongside, same era.
Contextual bandits with linear payoff functions
Wei Chu, Lihong Li, Lev Reyzin, and Robert E Schapire · 2011
Cited alongside, same era.
Pac-bayesian analysis of contextual bandits
Yevgeny Seldin, Peter Auer, John S Shawe-taylor, Ronald Ortner, and François Laviolette · 2011
Cited alongside, same era.
A note on performance limitations in bandit problems with side information
Alexander Goldenshluger and Assaf Zeevi · 2011
Cited alongside, same era.
Bayesian dynamic pricing policies: Learning and earning under a binary prior distribution
J Michael Harrison, N Bora Keskin, and Assaf Zeevi · 2012
Cited alongside, same era.
Dynamic pricing under a general parametric choice model
Josef Broder and Paat Rusmevichientong · 2012
Cited alongside, same era.
Chasing demand: Learning and earning in a changing environment
N Bora Keskin and Assaf Zeevi · 2013
Cited alongside, same era.
Cynthia Rudin and Gah-Yi Vahn · 2014
Later among the works it cites.
Repeated contextual auctions with strategic buyers
Kareem Amin, Afshin Rostamizadeh, and Umar Syed · 2014
Later among the works it cites.
Dynamic pricing with multiple products and partially specified demand distribution
Arnoud V. den Boer · 2014
Later among the works it cites.
Online network revenue management using thompson sampling
Kris Johnson, David Simchi-Levi, and He Wang · 2015
Later among the works it cites.
Nonparametric algorithms for joint pricing and inventorycontrol with lost-sales and censored demand, 2015
Boxiao Chen, Xiuli Chao, and Cong Shi · 2015
Later among the works it cites.
Online decision-making with high-dimensional covariates
Hamsa Bastani and Mohsen Bayati · 2015
Later among the works it cites.
The data-driven newsvendor problem: New bounds and insights
Retsef Levi, Georgia Perakis, and Joline Uichanco · 2015
Later among the works it cites.
From predictive to prescriptive analytics
Dimitris Bertsimas and Nathan Kallus · 2015
Later among the works it cites.
Feature-based dynamic pricing
Maxime C. Cohen, Ilan Lobel, and Renato Paes Leme · 2016
Closest in time.