Fetching the paper…
Reading the bibliography…
Optimizing an interactive system against a predefined online metric is particularly challenging, when the metric is computed from user feedback such as clicks and payments.
Probability inequalities for sums of bounded random variables
Wassily Hoeffding · 1963
Earlier work this paper cites.
Bayesian inference for causal effects: The role of randomization
Donald B. Rubin · 1978
Earlier work this paper cites.
Pattern-recognizing stochastic learning automata
Andrew G. Barto and P. Anandan · 1985
Earlier work this paper cites.
Statistics and causal inference
Paul W. Holland · 1986
Earlier work this paper cites.
Modern Information Retrieval
Ricardo Baeza-Yates and Berthier Ribeiro-Neto · 1999
Earlier work this paper cites.
Online bagging and boosting
Nikunj C. Oza and Stuart Russell · 2001
Earlier work this paper cites.
Why batch and user evaluations do not give the same results
Andrew H. Turpin and William Hersh · 2001
Earlier work this paper cites.
Cumulated gain-based evaluation of IR techniques
Kalervo Järvelin and Jaana Kekäläinen · 2002
Earlier work this paper cites.
Personalized search
James Pitokow, Hinrich Schütze, Todd Cass, Rob Cooley, Don Turnbull, Andy Edmonds, Eytan Adar, and Thomas Breuel · 2002
Earlier work this paper cites.
Learning to rank using gradient descent
Christopher J. C. Burges, Tal Shaked, Erin Renshaw, Ari Lazier, Matt Deeds, Nicole Hamilton, and Gregory N. Hullender · 2005
Earlier work this paper cites.
More bang for their bucks: Assessing new features for online advertisers
Diane Lambert and Daryl Pregibon · 2007
Earlier work this paper cites.
Some(what) grand challenges for information retrieval
Nicholas J. Belkin · 2008
Cited alongside, same era.
Exploration scavenging
John Langford, Alexander L. Strehl, and Jennifer Wortman · 2008
Cited alongside, same era.
On the history of evaluation in IR
Stephen Robertson · 2008
Cited alongside, same era.
A general boosting method and its application to learning ranking functions for web search
Zhaohui Zheng, Hongyuan Zha, Tong Zhang, Olivier Chapelle, Keke Chen, and Gordon Sun · 2008
Cited alongside, same era.
The offset tree for learning with partial labels
Alina Beygelzimer and John Langford · 2009
Cited alongside, same era.
A dynamic Bayesian network click model for Web search ranking
Olivier Chapelle and Ya Zhang · 2009
Cited alongside, same era.
A large scale ranker-based system for search query spelling correction
Jianfeng Gao, Xiaolong Li, Daniel Micol, Chris Quirk, and Xu Sun · 2010
Later among the works it cites.
A contextual-bandit approach to personalized news article recommendation
Lihong Li, Wei Chu, John Langford, and Robert E. Schapire · 2010
Later among the works it cites.
Potential for personalization
Jaime Teevan, Susan T. Dumais, and Eric Horvitz · 2010
Later among the works it cites.
Doubly robust policy evaluation and learning
Miroslav Dudík, John Langford, and Lihong Li · 2011
Later among the works it cites.
Unbiased offline evaluation of contextual-bandit-based news article recommendation algorithms
Lihong Li, Wei Chu, John Langford, and Xuanhui Wang · 2011
Later among the works it cites.
Lambdamerge: merging the results of query reformulations
Daniel Sheldon, Milad Shokouhi, Martin Szummer, and Nick Craswell · 2011
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Integration of news content into Web results
Fernando Diaz · 2009
Cited alongside, same era.
Click chain model in Web search
Fan Guo, Chao Liu, Anitha Kannan, Tom Minka, Michael J. Taylor, Yi-Min Wang, and Christos Faloutsos · 2009
Cited alongside, same era.
Controlled experiments on the web: Survey and practical guide
Ron Kohavi, Roger Longbotham, Dan Sommerfield, and Randal M. Henne · 2009
Cited alongside, same era.
Evaluating online ad campaigns in a pipeline: Causal models at scale
David Chan, Rong Ge, Ori Gershony, Tim Hesterberg, and Diane Lambert · 2010
Cited alongside, same era.
Towards recency ranking in Web search
Anlei Dong, Yi Chang, Zhaohui Zheng, Gilad Mishne, Jing Bai, Ruiqiang Zhang, Karolina Buchner, Ciya Liao, and Fernando Diaz · 2010
Cited alongside, same era.
The Text REtrieval Conference
TREC
Cited in the paper.
Later among the works it cites.
Learning from logged implicit exploration data
Alexander L. Strehl, John Langford, Lihong Li, and Sham M. Kakade · 2011
Later among the works it cites.
Large scale validation and analysis of interleaved search evaluation
Olivier Chapelle, Thorsten Joachims, Filip Radlinski, and Yisong Yue · 2012
Later among the works it cites.
An online learning framework for refining recency search results with user click feedback
Taesup Moon, Wei Chu, Lihong Li, Zhaohui Zheng, and Yi Chang · 2012
Later among the works it cites.
Counterfactual reasoning and learning systems: The example of computational advertising
Léon Bottou, Jonas Peters, Joaquin Quiñonero-Candela, Denis Xavier Charles, D. Max Chickering, Elon Portugaly, Dipankar Ray, Patrice Simard, and Ed Snelson · 2013
Later among the works it cites.
Automatic ad format selection via contextual bandits
Liang Tang, Romer Rosales, Ajit Singh, and Deepak Agarwal · 2013
Later among the works it cites.