Fetching the paper…
Reading the bibliography…
Personalized web services strive to adapt their services (advertisements, news articles, etc) to individual users by making use of both content and user information.
On the likelihood that one unknown probability exceeds another in view of the evidence of two samples
W. R. Thompson · 1933
Earlier work this paper cites.
Some aspects of the sequential design of experiments
H. Robbins · 1952
Earlier work this paper cites.
Bandit processes and dynamic allocation indices
J. Gittins · 1979
Earlier work this paper cites.
Bandit Problems: Sequential Allocation of Experiments
D. A. Berry and B. Fristedt · 1985
Earlier work this paper cites.
Asymptotically efficient adaptive allocation rules
T. L. Lai and H. Robbins · 1985
Earlier work this paper cites.
Sample mean based index policies with
R. Agrawal · 1995
Earlier work this paper cites.
Text-learning and related intelligent agents: A survey
D. Mladenic · 1999
Earlier work this paper cites.
Recommender systems in e-commerce
J. B. Schafer, J. Konstan, and J. Riedi · 1999
Earlier work this paper cites.
Eligibility traces for off-policy policy evaluation
D. Precup, R. S. Sutton, and S. P. Singh · 2000
Earlier work this paper cites.
Using confidence bounds for exploitation-exploration trade-offs
P. Auer · 2002
Cited alongside, same era.
Finite-time analysis of the multiarmed bandit problem
P. Auer, N. Cesa-Bianchi, and P. Fischer · 2002
Cited alongside, same era.
The nonstochastic multiarmed bandit problem
P. Auer, N. Cesa-Bianchi, Y. Freund, and R. E. Schapire · 2002
Cited alongside, same era.
Reinforcement learning with immediate rewards and linear hypotheses
N. Abe, A. W. Biermann, and P. M. Long · 2003
Cited alongside, same era.
Information Theory, Inference, and Learning Algorithms
D. J. C. MacKay · 2003
Cited alongside, same era.
Hybrid systems for personalized recommendations
R. Burke · 2005
Cited alongside, same era.
Google news personalization: scalable online collaborative filtering
A. Das, M. Datar, A. Garg, and S. Rajaram · 2007
Later among the works it cites.
Efficient bandit algorithms for online multiclass prediction
S. M. Kakade, S. Shalev-Shwartz, and A. Tewari · 2008
Later among the works it cites.
The epoch-greedy algorithm for contextual multi-armed bandits
J. Langford and T. Zhang · 2008
Later among the works it cites.
Simulation studies of multi-armed bandits with covariates
N. G. Pavlidis, D. K. Tasoulis, and D. J. Hand · 2008
Later among the works it cites.
Explore/exploit schemes for web content optimization
D. Agarwal, B.-C. Chen, and P. Elango · 2009
Later among the works it cites.
Online models for content optimization
D. Agarwal, B.-C. Chen, P. Elango, N. Motgi, S.-T. Park, R. Ramakrishnan, S. Roy, and J. Zachariah · 2009
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Naïve filterbots for robust cold-start recommendations
S.-T. Park, D. Pennock, O. Madani, N. Good, and D. DeCoste · 2006
Cited alongside, same era.
Just-in-time contextual advertising
A. Anagnostopoulos, A. Z. Broder, E. Gabrilovich, V. Josifovski, and L. Riedel · 2007
Cited alongside, same era.
The Adaptive Web — Methods and Strategies of Web Personalization
P. Brusilovsky, A. Kobsa, and W. Nejdl, editors · 2007
Cited alongside, same era.
Personalized recommendation on dynamic content using predictive bilinear models
W. Chu and S.-T. Park · 2009
Later among the works it cites.
A case study of behavior-driven conjoint analysis on Yahoo!: Front Page Today Module
W. Chu, S.-T. Park, T. Beaupre, N. Motgi, A. Phadke, S. Chakraborty, and J. Zachariah · 2009
Later among the works it cites.
Exploring compact reinforcement-learning representations with linear regression
T. J. Walsh, I. Szita, C. Diuk, and M. L. Littman · 2009
Later among the works it cites.