Fetching the paper…
Reading the bibliography…
Online learning to rank (OLTR) aims to learn a ranker directly from implicit feedback derived from users' interactions, such as clicks.
Machine learning 8
Williams, R.J.: Simple statistical gradient-following algorithms for connectionist reinforcement learning · 1992
Earlier work this paper cites.
In: Proceedings of the Sixteenth International Conference on Machine Learning, vol. 99, pp. 278–287 (1999)
Ng, A.Y., Harada, D., Russell, S.: Policy invariance under reward transformations: Theory and application to reward shaping · 1999
Earlier work this paper cites.
In: Advances in neural information processing systems, pp. 1057–1063 (2000)
Sutton, R.S., McAllester, D.A., Singh, S.P., Mansour, Y.: Policy gradient methods for reinforcement learning with function approximation · 2000
Earlier work this paper cites.
In: Proceedings of the eighth ACM SIGKDD international conference on Knowledge discovery and data mining, pp. 133–142 (2002)
Joachims, T.: Optimizing search engines using clickthrough data · 2002
Earlier work this paper cites.
ACM Transactions on Information Systems (TOIS) 23
Adomavicius, G., Sankaranarayanan, R., Sen, S., Tuzhilin, A.: Incorporating contextual information in recommender systems using a multidimensional approach · 2005
Earlier work this paper cites.
In: Proceedings of the SIGCHI conference on Human factors in computing systems, pp. 417–420 (2007)
Guan, Z., Cutrell, E.: An eye tracking study of the effect of target rank on web search · 2007
Earlier work this paper cites.
Journal of computer-mediated communication 12
Pan, B., Hembrooke, H., Joachims, T., Lorigo, L., Gay, G., Granka, L.: In google we trust: Users’ decisions on rank, position, and relevance · 2007
Earlier work this paper cites.
In: Proceedings of the 26th Annual International Conference on Machine Learning, pp. 1201–1208 (2009)
Yue, Y., Joachims, T.: Interactively optimizing information retrieval systems as a dueling bandits problem · 2009
Earlier work this paper cites.
Journal of the American Society for Information Science and Technology 61
Al-Maskari, A., Sanderson, M.: A review of factors influencing user satisfaction in information retrieval · 2010
Earlier work this paper cites.
Information Retrieval 13
Qin, T., Liu, T.Y., Xu, J., Li, H.: Letor: A benchmark collection for research on learning to rank for information retrieval · 2010
Earlier work this paper cites.
Now Publishers Inc (2010)
Sanderson, M.: Test collection based evaluation of information retrieval systems · 2010
Earlier work this paper cites.
In: Proceedings of the learning to rank challenge, pp. 1–24 (2011)
Chapelle, O., Chang, Y.: Yahoo! learning to rank challenge overview · 2011
Earlier work this paper cites.
In: European Conference on Information Retrieval, pp. 251–263. Springer (2011)
Hofmann, K., Whiteson, S., De Rijke, M.: Balancing exploration and exploitation in learning to rank online · 2011
Earlier work this paper cites.
In: Proceedings of the 20th ACM international conference on Information and knowledge management, pp. 249–258 (2011)
Hofmann, K., Whiteson, S., De Rijke, M.: A probabilistic method for inferring preferences from clicks · 2011
Earlier work this paper cites.
In: NIPS 2011 Workshop on Bayesian Optimization, Experimental Design, and Bandits, Granada, vol. 12, p. 2011 (2011)
Hofmann, K., Whiteson, S., de Rijke, M., et al.: Contextual bandits for information retrieval · 2011
Earlier work this paper cites.
Synthesis lectures on human language technologies 4
Li, H.: Learning to rank for information retrieval and natural language processing · 2011
Earlier work this paper cites.
Springer Science & Business Media (2011)
Liu, T.Y.: Learning to rank for information retrieval · 2011
Earlier work this paper cites.
In: Proceedings of the sixth ACM international conference on Web search and data mining, pp. 183–192. ACM (2013)
Hofmann, K., Schuth, A., Whiteson, S., de Rijke, M.: Reusing historical interaction data for faster online learning to rank for ir · 2013
Earlier work this paper cites.
arXiv preprint arXiv:1306.2597 (2013)
Qin, T., Liu, T.Y.: Introducing letor 4.0 datasets · 2013
Earlier work this paper cites.
In: Proceedings of the 23rd ACM International Conference on Conference on Information and Knowledge Management, pp. 589–598 (2014)
Lefortier, D., Serdyukov, P., De Rijke, M.: Online exploration for detecting shifts in fresh intent · 2014
Earlier work this paper cites.
In: Proceedings of the 38th International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 955–958. ACM (2015)
Schuth, A., Bruintjes, R.J., Buüttner, F., van Doorn, J., Groenland, C., Oosterhuis, H., Tran, C.N., Veeling, B., van der Velde, J., Wechsler, R., et al.: Probabilistic multileave for online retrieval evaluation · 2015
Earlier work this paper cites.
ACM Transactions on Information Systems (TOIS) 35
Dato, D., Lucchese, C., Nardini, F.M., Orlando, S., Perego, R., Tonellotto, N., Venturini, R.: Fast ranking with additive ensembles of oblivious and non-oblivious regression trees · 2016
Earlier work this paper cites.
Foundations and trends in information retrieval 10
Hofmann, K., Li, L., Radlinski, F.: Online evaluation for information retrieval · 2016
Cited alongside, same era.
In: European Conference on Information Retrieval, pp. 661–668. Springer (2016)
Oosterhuis, H., Schuth, A., de Rijke, M.: Probabilistic multileave gradient descent · 2016
Cited alongside, same era.
In: Proceedings of the Ninth ACM International Conference on Web Search and Data Mining, pp. 457–466 (2016)
Schuth, A., Oosterhuis, H., Whiteson, S., de Rijke, M.: Multileave gradient descent for fast online learning to rank · 2016
Cited alongside, same era.
In: Proceedings of the 39th International ACM SIGIR conference on Research and Development in Information Retrieval, pp. 115–124 (2016)
Wang, X., Bendersky, M., Metzler, D., Najork, M.: Learning to rank with selection bias in personal search · 2016
Cited alongside, same era.
In: Proceedings of the 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 687–696 (2017)
Agarwal, A., Basu, S., Schnabel, T., Joachims, T.: Effective evaluation using logged bandit feedback from multiple loggers · 2017
In: The World Wide Web Conference, pp. 2830–2836 (2019)
Hu, Z., Wang, Y., Peng, Q., Li, H.: Unbiased lambdamart: an unbiased pairwise learning-to-rank algorithm · 2019
Later among the works it cites.
In: Proceedings of the 42nd International ACM SIGIR Conference on Research and Development in Information Retrieval, SIGIR’19, pp. 15–24. Association for Computing Machinery (2019)
Jagerman, R., Oosterhuis, H., de Rijke, M.: To model or to intervene: A comparison of counterfactual and online learning to rank from user interactions · 2019
Later among the works it cites.
In: European Conference on Information Retrieval, pp. 382–396. Springer (2019)
Oosterhuis, H., de Rijke, M.: Optimizing ranking models in an online setting · 2019
Later among the works it cites.
In: Proceedings of the 42Nd International ACM SIGIR Conference on Research and Development in Information Retrieval, SIGIR’19 (2019)
Wang, H., Kim, S., McCord-Snook, E., Wu, Q., Wang, H.: Variance reduction in gradient exploration for online learning to rank · 2019
Later among the works it cites.
ACM Transactions on Information Systems (TOIS) 38
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
In: Proceedings of the Tenth ACM International Conference on Web Search and Data Mining, pp. 781–789 (2017)
Joachims, T., Swaminathan, A., Schnabel, T.: Unbiased learning-to-rank with biased feedback · 2017
Cited alongside, same era.
In: Proceedings of the 40th International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 135–144 (2017)
Maxwell, D., Azzopardi, L., Moshfeghi, Y.: A study of snippet length and informativeness: Behaviour, performance and user experience · 2017
Cited alongside, same era.
arXiv preprint arXiv:1704.03073 (2017)
Popov, I., Heess, N., Lillicrap, T., Hafner, R., Barth-Maron, G., Vecerik, M., Lampe, T., Tassa, Y., Erez, T., Riedmiller, M.: Data-efficient deep reinforcement learning for dexterous manipulation · 2017
Cited alongside, same era.
In: Proceedings of the 40th International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 945–948 (2017)
Wei, Z., Xu, J., Lan, Y., Guo, J., Cheng, X.: Reinforcement learning to rank with markov decision process · 2017
Cited alongside, same era.
In: Proceedings of the 40th International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 535–544 (2017)
Xia, L., Xu, J., Lan, Y., Guo, J., Zeng, W., Cheng, X.: Adapting markov decision process for search result diversification · 2017
Cited alongside, same era.
arXiv preprint arXiv:1805.00065 (2018)
Agarwal, A., Zaitsev, I., Joachims, T.: Counterfactual learning-to-rank for additive metrics and deep models · 2018
Cited alongside, same era.
In: The 41st International ACM SIGIR Conference on Research & Development in Information Retrieval, pp. 385–394 (2018)
Ai, Q., Bi, K., Luo, C., Guo, J., Croft, W.B.: Unbiased learning to rank with unbiased propensity estimation · 2018
Cited alongside, same era.
Jagerman, R., Markov, I., Rijke, M.D.: Safe exploration for optimizing contextual bandits · 2020
Later among the works it cites.
In: Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 469–478 (2020)
Jagerman, R., de Rijke, M.: Accelerated convergence for counterfactual learning to rank · 2020
Later among the works it cites.
In: Proceedings of the 42nd International ACM SIGIR Conference on Research and Development in Information Retrieval (2020)
Jun, X., Zeng, W., Long, X., Yanyan, L., Dawei, Y., Xueqi, C., Ji-Rong, W.: Reinforcement learning to rank with pairwise policy gradient · 2020
Later among the works it cites.
ACM Transactions on Information Systems (TOIS) 38
Li, C., Markov, I., Rijke, M.D., Zoghi, M.: Mergedts: A method for effective large-scale online ranker evaluation · 2020
Later among the works it cites.
In: Proceedings of the 42nd International ACM SIGIR Conference on Research and Development in Information Retrieval (2020)
Oosterhuis, H., de Rijke, M.: Policy-aware unbiased learning to rank for top-k rankings · 2020
Later among the works it cites.
In: Proceedings of the 2020 ACM SIGIR on International Conference on Theory of Information Retrieval, pp. 137–144 (2020)
Oosterhuis, H., de Rijke, M.: Taking the counterfactual online: Efficient and unbiased online evaluation for ranking · 2020
Later among the works it cites.
In: Proceedings of The Web Conference 2020, pp. 1863–1873 (2020)
Ovaisi, Z., Ahsan, R., Zhang, Y., Vasilaky, K., Zheleva, E.: Correcting for selection bias in learning-to-rank systems · 2020
Later among the works it cites.
arXiv preprint arXiv:2005.11938 (2020)
Vardasbi, A., de Rijke, M., Markov, I.: Cascade model-based propensity estimation for counterfactual learning to rank · 2020
Later among the works it cites.
In: Proceedings of The Web Conference 2020, pp. 2298–2308 (2020)
Yao, J., Dou, Z., Xu, J., Wen, J.R.: Rlper: A reinforcement learning model for personalized search · 2020
Later among the works it cites.
In: European Conference on Information Retrieval, pp. 415–430. Springer (2020)
Zhuang, S., Zuccon, G.: Counterfactual online learning to rank · 2020
Later among the works it cites.
ACM Transactions on Information Systems (TOIS) 39
Ai, Q., Yang, T., Wang, H., Mao, J.: Unbiased learning to rank: Online or offline? · 2021
Later among the works it cites.
In: Proceedings of the 14th ACM International Conference on Web Search and Data Mining, pp. 463–471 (2021)
Oosterhuis, H., de Rijke, M.: Unifying online and counterfactual learning to rank: A novel counterfactual estimator that effectively utilizes online interventions · 2021
Later among the works it cites.
In: Proceedings of the 14th ACM International Conference on Web Search and Data Mining, pp. 481–489 (2021)
Wang, N., Qin, Z., Wang, X., Wang, H.: Non-clicks mean irrelevant? propensity ratio scoring as a correction · 2021
Later among the works it cites.
In: Proceedings of the 2021 ACM SIGIR International Conference on Theory of Information Retrieval, pp. 3–12 (2021)
Wang, S., Liu, B., Zhuang, S., Zuccon, G.: Effective and privacy-preserving federated online learning to rank · 2021
Later among the works it cites.
In: The 43rd European Conference On Information Retrieval (ECIR) (2021)
Wang, S., Zhuang, S., Zuccon, G.: Federated online learning to rank with evolution strategies: A reproducibility study · 2021
Later among the works it cites.
In: Proceedings of the AAAI Conference on Artificial Intelligence, vol. 35, pp. 750–758 (2021)
Zhao, X., Gu, C., Zhang, H., Yang, X., Liu, X., Liu, H., Tang, J.: Dear: Deep reinforcement learning for online advertising impression in recommender systems · 2021
Later among the works it cites.
In: Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval (2021)
Zhuang, S., Zuccon, G.: How do online learning to rank methods adapt to changes of intent? · 2021
Later among the works it cites.