Fetching the paper…
Reading the bibliography…
Learning feature interaction is the critical backbone to building recommender systems.
Embedding-based news recommendation for millions of users
Okura, S., Tagami, Y., Ono, S., and Tajima, A · 1942
Earlier work this paper cites.
Multitask learning
Caruana, R · 1997
Earlier work this paper cites.
Recommender systems
Resnick, P., and Varian, H. R · 1997
Earlier work this paper cites.
Recommender systems in e-commerce
Schafer, J. B., Konstan, J., and Riedl, J · 1999
Earlier work this paper cites.
The youtube video recommendation system
Davidson, J., Liebald, B., Liu, J., Nandy, P., Van Vleet, T., Gargi, U., Gupta, S., He, Y., Lambert, M., Livingston, B., et al · 2010
Earlier work this paper cites.
Factorization machines
Rendle, S · 2010
Earlier work this paper cites.
Practical lessons from predicting clicks on ads at facebook
He, X., Pan, J., Jin, O., Xu, T., Liu, B., Xu, T., Shi, Y., Atallah, A., Herbrich, R., Bowers, S., et al · 2014
Earlier work this paper cites.
Deep learning
LeCun, Y., Bengio, Y., and Hinton, G · 2015
Earlier work this paper cites.
Recommender system application developments: a survey
Lu, J., Wu, D., Mao, M., Wang, W., and Zhang, G · 2015
Earlier work this paper cites.
Wide & deep learning for recommender systems
Cheng, H.-T., Koc, L., Harmsen, J., Shaked, T., Chandra, T., Aradhye, H., Anderson, G., Corrado, G., Chai, W., Ispir, M., et al · 2016
Earlier work this paper cites.
Deep neural networks for youtube recommendations
Covington, P., Adams, J., and Sargin, E · 2016
Earlier work this paper cites.
Gaussian error linear units (gelus)
Hendrycks, D., and Gimpel, K · 2016
Earlier work this paper cites.
Field-aware factorization machines for ctr prediction
Juan, Y., Zhuang, Y., Chin, W.-S., and Lin, C.-J · 2016
Earlier work this paper cites.
Difacto: Distributed factorization machines
Li, M., Liu, Z., Smola, A. J., and Wang, Y.-X · 2016
Earlier work this paper cites.
Product-based neural networks for user response prediction
Qu, Y., Cai, H., Ren, K., Zhang, W., Yu, Y., Wen, Y., and Wang, J · 2016
Earlier work this paper cites.
Neural architecture search: A survey
Elsken, T., Metzen, J. H., and Hutter, F · 2017
Earlier work this paper cites.
Deepfm: a factorization-machine based neural network for ctr prediction
Guo, H., Tang, R., Ye, Y., Li, Z., and He, X · 2017
Earlier work this paper cites.
Neural collaborative filtering
He, X., Liao, L., Zhang, H., Nie, L., Hu, X., and Chua, T.-S · 2017
Cited alongside, same era.
In-datacenter performance analysis of a tensor processing unit
Jouppi, N. P., Young, C., Patil, N., Patterson, D., Agrawal, G., Bajwa, R., Bates, S., Bhatia, S., Boden, N., Borchers, A., et al · 2017
Cited alongside, same era.
Detecting statistical interactions from neural network weights
Tsang, M., Cheng, D., and Liu, Y · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Cited alongside, same era.
Deep & cross network for ad click predictions
Wang, R., Fu, B., Fu, G., and Wang, M · 2017
Cited alongside, same era.
Latent cross: Making use of context in recurrent recommender systems
Neural input search for large scale recommendation models
Joglekar, M. R., Li, C., Chen, M., Xu, T., Wang, X., Adams, J. K., Khaitan, P., Liu, J., and Le, Q. V · 2020
Later among the works it cites.
Interpretable click-through rate prediction through hierarchical attention
Li, Z., Cheng, W., Chen, Y., Chen, H., and Wang, W · 2020
Later among the works it cites.
Glu variants improve transformer
Shazeer, N · 2020
Later among the works it cites.
Introduction to tensorflow 2.0
Singh, P., Manure, A., Singh, P., and Manure, A · 2020
Later among the works it cites.
Towards automated neural interaction discovery for click-through rate prediction
Song, Q., Cheng, D., Zhou, H., Yang, J., Tian, Y., and Hu, X · 2020
Later among the works it cites.
Mixed negative sampling for learning two-tower neural networks in recommendations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Beutel, A., Covington, P., Jain, S., Xu, C., Li, J., Gatto, V., and Chi, E. H · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2018
Cited alongside, same era.
Motivation for and evaluation of the first tensor processing unit
Jouppi, N., Young, C., Patil, N., and Patterson, D · 2018
Cited alongside, same era.
Self-attentive sequential recommendation
Kang, W.-C., and McAuley, J · 2018
Cited alongside, same era.
xdeepfm: Combining explicit and implicit feature interactions for recommender systems
Lian, J., Zhou, X., Zhang, F., Chen, Z., Xie, X., and Sun, G · 2018
Cited alongside, same era.
Product-based neural networks for user response prediction over multi-field categorical data
Qu, Y., Fang, B., Zhang, W., Tang, R., Niu, M., Guo, H., Yu, Y., and He, X · 2018
Cited alongside, same era.
Deep learning recommendation model for personalization and recommendation systems
Naumov, M., Mudigere, D., Shi, H.-J. M., Huang, J., Sundaraman, N., Park, J., Wang, X., Gupta, U., Wu, C.-J., Azzolini, A. G., et al · 2019
Cited alongside, same era.
Yang, J., Yi, X., Zhiyuan Cheng, D., Hong, L., Li, Y., Xiaoming Wang, S., Xu, T., and Chi, E. H · 2020
Later among the works it cites.
Tensorflow model garden
Yu, H., Chen, C., Du, X., Li, Y., Rashwan, A., Hou, L., Jin, P., Yang, F., Liu, F., Kim, J., et al · 2020
Later among the works it cites.
Feature transformation for neural ranking models
Zhuang, H., Wang, X., Bendersky, M., and Najork, M · 2020
Later among the works it cites.
Perceiver: General perception with iterative attention
Jaegle, A., Gimeno, F., Brock, A., Vinyals, O., Zisserman, A., and Carreira, J · 2021
Later among the works it cites.
Ammus: A survey of transformer-based pretrained models in natural language processing
Kalyan, K. S., Rajasekharan, A., and Sangeetha, S · 2021
Later among the works it cites.
Attentive capsule network for click-through rate and conversion rate prediction in online advertising
Li, D., Hu, B., Chen, Q., Wang, X., Qi, Q., Wang, L., and Liu, H · 2021
Later among the works it cites.
A graph placement methodology for fast chip design
Mirhoseini, A., Goldie, A., Yazgan, M., Jiang, J. W., Songhori, E., Wang, S., Lee, Y.-J., Johnson, E., Pathak, O., Nazi, A., et al · 2021
Later among the works it cites.
Dcn v2: Improved deep & cross network and practical lessons for web-scale learning to rank systems
Wang, R., Shivanna, R., Cheng, D., Jain, S., Lin, D., Hong, L., and Chi, E · 2021
Later among the works it cites.
A survey on vision transformer
Han, K., Wang, Y., Chen, H., Chen, X., Guo, J., Liu, Z., Tang, Y., Xiao, A., Xu, C., Xu, Y., et al · 2022
Later among the works it cites.
Transformers in vision: A survey
Khan, S., Naseer, M., Hayat, M., Zamir, S. W., Khan, F. S., and Shah, M · 2022
Later among the works it cites.
Metacvr: Conversion rate prediction via meta learning in small-scale recommendation scenarios
Pan, X., Li, M., Zhang, J., Yu, K., Wen, H., Wang, L., Mao, C., and Cao, B · 2022
Later among the works it cites.
Efficient transformers: A survey
Tay, Y., Dehghani, M., Bahri, D., and Metzler, D · 2022
Later among the works it cites.