Fetching the paper…
Reading the bibliography…
Click-through rate (CTR) prediction, which aims to predict the probability of a user clicking on an ad or an item, is critical to many online applications such as online advertising and recommender systems.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Predicting clicks: estimating the click-through rate for new ads. In Proceedings of the 16th international conference on World Wide Web . ACM, 521–530
Matthew Richardson, Ewa Dominowska, and Robert Ragno. 2007 · 2007
Earlier work this paper cites.
Web-scale Bayesian Click-through Rate Prediction for Sponsored Search Advertising in Microsoft’s Bing Search Engine. In Proceedings of the 27th International Conference on International Conference on Machine Learning . 13–20
Thore Graepel, Joaquin Quiñonero Candela, Thomas Borchert, and Ralf Herbrich. 2010 · 2010
Earlier work this paper cites.
Factorization machines. In Data Mining (ICDM), 2010 IEEE 10th International Conference on . IEEE, 995–1000
Steffen Rendle. 2010 · 2010
Earlier work this paper cites.
Factorizing personalized markov chains for next-basket recommendation. In Proceedings of the 19th international conference on World wide web . ACM, 811–820
Steffen Rendle, Christoph Freudenthaler, and Lars Schmidt-Thieme. 2010 · 2010
Earlier work this paper cites.
Unsupervised learning of hierarchical representations with convolutional deep belief networks
Honglak Lee, Roger Grosse, Rajesh Ranganath, and Andrew Y Ng. 2011 · 2011
Earlier work this paper cites.
Fast context-aware recommendations with factorization machines. In Proceedings of the 34th international ACM SIGIR conference on Research and development in Information Retrieval . ACM, 635–644
Steffen Rendle, Zeno Gantner, Christoph Freudenthaler, and Lars Schmidt-Thieme. 2011 · 2011
Earlier work this paper cites.
Representation learning: A review and new perspectives
Yoshua Bengio, Aaron Courville, and Pascal Vincent. 2013 · 2013
Earlier work this paper cites.
Ad Click Prediction: A View from the Trenches. In Proceedings of the 19th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 1222–1230
H. Brendan McMahan, Gary Holt, D. Sculley, Michael Young, Dietmar Ebner, Julian Grady, Lan Nie, Todd Phillips, et al · 2013
Earlier work this paper cites.
Gradient boosting factorization machines. In Proceedings of the 8th ACM Conference on Recommender systems . ACM, 265–272
Chen Cheng, Fen Xia, Tong Zhang, Irwin King, and Michael R Lyu. 2014 · 2014
Earlier work this paper cites.
Practical lessons from predicting clicks on ads at facebook. In Proceedings of the Eighth International Workshop on Data Mining for Online Advertising . ACM, 1–9
Xinran He, Junfeng Pan, Ou Jin, Tianbing Xu, Bo Liu, Tao Xu, Yanxin Shi, Antoine Atallah, Ralf Herbrich, Stuart Bowers, et al · 2014
Earlier work this paper cites.
Predicting response in mobile advertising with hierarchical importance-aware factorization machine. In Proceedings of the 7th ACM international conference on Web search and data mining . ACM, 123–132
Richard J Oentaryo, Ee-Peng Lim, Jia-Wei Low, David Lo, and Michael Finegold. 2014 · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate. In International Conference on Learning Representations
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization. In International Conference on Learning Representations
Diederick P Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
A Neural Attention Model for Abstractive Sentence Summarization. In Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 379–389
Alexander M. Rush, Sumit Chopra, and Jason Weston. 2015 · 2015
Cited alongside, same era.
End-to-end memory networks. In Advances in neural information processing systems . 2440–2448
Sainbayar Sukhbaatar, Jason Weston, Rob Fergus, et al · 2015
Cited alongside, same era.
TensorFlow: A System for Large-Scale Machine Learning.. In OSDI , Vol. 16. 265–283
Martín Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, et al · 2016
Cited alongside, same era.
Wide & deep learning for recommender systems. In Proceedings of the 1st Workshop on Deep Learning for Recommender Systems . ACM, 7–10
Heng-Tze Cheng, Levent Koc, Jeremiah Harmsen, Tal Shaked, Tushar Chandra, Hrishi Aradhye, Glen Anderson, Greg Corrado, Wei Chai, Mustafa Ispir, et al · 2016
Cited alongside, same era.
A structured self-attentive sentence embedding. In International Conference on Learning Representations
Zhouhan Lin, Minwei Feng, Cicero Nogueira dos Santos, Mo Yu, Bing Xiang, Bowen Zhou, and Yoshua Bengio. 2017 · 2017
Later among the works it cites.
Attention is all you need. In Advances in Neural Information Processing Systems . 6000–6010
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Deep & Cross Network for Ad Click Predictions. In Proceedings of the ADKDD’17 . ACM, 12:1–12:7
Ruoxi Wang, Bin Fu, Gang Fu, and Mingliang Wang. 2017 · 2017
Later among the works it cites.
Attentional factorization machines: learning the weight of feature interactions via attention networks. In Proceedings of the 26th International Joint Conference on Artificial Intelligence . AAAI Press, 3119–3125
Jun Xiao, Hao Ye, Xiangnan He, Hanwang Zhang, Fei Wu, and Tat-Seng Chua. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Paul Covington, Jay Adams, and Emre Sargin. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Field-aware factorization machines for CTR prediction. In Proceedings of the 10th ACM Conference on Recommender Systems . ACM, 43–50
Yuchin Juan, Yong Zhuang, Wei-Sheng Chin, and Chih-Jen Lin. 2016 · 2016
Cited alongside, same era.
Key-Value Memory Networks for Directly Reading Documents. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 1400–1409
Alexander Miller, Adam Fisch, Jesse Dodge, Amir-Hossein Karimi, Antoine Bordes, and Jason Weston. 2016 · 2016
Cited alongside, same era.
Alexander Novikov, Mikhail Trofimov, and Ivan Oseledets. 2016 · 2016
Cited alongside, same era.
Product-based neural networks for user response prediction. In Data Mining (ICDM), 2016 IEEE 16th International Conference on . IEEE, 1149–1154
Yanru Qu, Han Cai, Kan Ren, Weinan Zhang, Yong Yu, Ying Wen, and Jun Wang. 2016 · 2016
Cited alongside, same era.
Predicting ad click-through rates via feature-based fully coupled interaction tensor factorization
Lili Shan, Lei Lin, Chengjie Sun, and Xiaolong Wang. 2016b · 2016
Cited alongside, same era.
Deep learning over multi-field categorical data. In European conference on information retrieval . Springer, 45–57
Weinan Zhang, Tianming Du, and Jun Wang. 2016 · 2016
Cited alongside, same era.
GB-CENT: Gradient Boosted Categorical Embedding and Numerical Trees. In Proceedings of the 26th International Conference on World Wide Web . International World Wide Web Conferences Steering Committee, 1311–1319
Qian Zhao, Yue Shi, and Liangjie Hong. 2017 · 2017
Later among the works it cites.
Deep embedding forest: Forest-based serving with deep embedding features. In Proceedings of the 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 1703–1711
Jie Zhu, Ying Shan, JC Mao, Dong Yu, Holakou Rahmanian, and Yi Zhang. 2017 · 2017
Later among the works it cites.
Latent Cross: Making Use of Context in Recurrent Recommender Systems. In Proceedings of the Eleventh ACM International Conference on Web Search and Data Mining . ACM, 46–54
Alex Beutel, Paul Covington, Sagar Jain, Can Xu, Jia Li, Vince Gatto, and Ed H Chi. 2018 · 2018
Closest in time.
NAIS: Neural attentive item similarity model for recommendation
Xiangnan He, Zhankui He, Jingkuan Song, Zhenguang Liu, Yu-Gang Jiang, and Tat-Seng Chua. 2018 · 2018
Closest in time.
xDeepFM: Combining Explicit and Implicit Feature Interactions for Recommender Systems. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 1754–1763
Jianxun Lian, Xiaohuan Zhou, Fuzheng Zhang, Zhongxia Chen, Xing Xie, and Guangzhong Sun. 2018 · 2018
Closest in time.
Graph Attention Networks. In International Conference on Learning Representations
Petar Velickovic, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Lio, and Yoshua Bengio. 2018 · 2018
Closest in time.
TEM: Tree-enhanced Embedding Model for Explainable Recommendation. In Proceedings of the 2018 World Wide Web Conference on World Wide Web . International World Wide Web Conferences Steering Committee, 1543–1552
Xiang Wang, Xiangnan He, Fuli Feng, Liqiang Nie, and Tat-Seng Chua. 2018 · 2018
Closest in time.
Deep Interest Network for Click-Through Rate Prediction. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 1059–1068
Guorui Zhou, Xiaoqiang Zhu, Chenru Song, Ying Fan, Han Zhu, Xiao Ma, Yanghui Yan, Junqi Jin, Han Li, and Kun Gai. 2018 · 2018
Closest in time.
Session-based Social Recommendation via Dynamic Graph Attention Networks. In Proceedings of the Twelfth ACM International Conference on Web Search and Data Mining . ACM, 555–563
Weiping Song, Zhiping Xiao, Yifan Wang, Laurent Charlin, Ming Zhang, and Jian Tang. 2019 · 2019
Closest in time.