Fetching the paper…
Reading the bibliography…
Learning effective feature crosses is the key behind building recommender systems.
Modeling task relationships in multi-task learning with multi-gate mixture-of-experts. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining
Jiaqi Ma, Zhe Zhao, Xinyang Yi, Jilin Chen, Lichan Hong, and Ed H Chi. 2018 · 1939
Earlier work this paper cites.
Learning internal representations by error propagation
David E Rumelhart, Geoffrey E Hinton, and Ronald J Williams. 1985 · 1985
Earlier work this paper cites.
Backpropagation applied to handwritten zip code recognition
Yann LeCun, Bernhard Boser, John S Denker, Donnie Henderson, Richard E Howard, Wayne Hubbard, and Lawrence D Jackel. 1989 · 1989
Earlier work this paper cites.
Adaptive mixtures of local experts
Robert A Jacobs, Michael I Jordan, Steven J Nowlan, and Geoffrey E Hinton. 1991 · 1991
Earlier work this paper cites.
Matrix Computations Johns Hopkins University Press
Gene H Golub and Charles F Van Loan. 1996 · 1996
Earlier work this paper cites.
Neural networks for optimal approximation of smooth and analytic functions
Hrushikesh N Mhaskar. 1996 · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Face recognition: A convolutional neural-network approach
Steve Lawrence, C Lee Giles, Ah Chung Tsoi, and Andrew D Back. 1997 · 1997
Earlier work this paper cites.
Recommender systems
Paul Resnick and Hal R Varian. 1997 · 1997
Earlier work this paper cites.
Recommender systems in e-commerce. In Proceedings of the 1st ACM conference on Electronic commerce
J Ben Schafer, Joseph Konstan, and John Riedl. 1999 · 1999
Earlier work this paper cites.
Evaluating collaborative filtering recommender systems
Jonathan L Herlocker, Joseph A Konstan, Loren G Terveen, and John T Riedl. 2004 · 2004
Earlier work this paper cites.
On the Nyström method for approximating a Gram matrix for improved kernel-based learning
Petros Drineas and Michael W Mahoney. 2005 · 2005
Earlier work this paper cites.
Learning to rank: from pairwise approach to listwise approach. In Proceedings of the 24th international conference on Machine learning
Zhe Cao, Tao Qin, Tie-Yan Liu, Ming-Feng Tsai, and Hang Li. 2007 · 2007
Earlier work this paper cites.
Computational advertising and recommender systems. In Proceedings of the 2008 ACM conference on Recommender systems
Andrei Z Broder. 2008 · 2008
Earlier work this paper cites.
Factorization machines. In 2010 IEEE International Conference on Data Mining
Steffen Rendle. 2010 · 2010
Earlier work this paper cites.
Finding structure with randomness: Probabilistic algorithms for constructing approximate matrix decompositions
Nathan Halko, Per-Gunnar Martinsson, and Joel A Tropp. 2011 · 2011
Earlier work this paper cites.
Learning to rank for information retrieval
Tie-Yan Liu. 2011 · 2011
Earlier work this paper cites.
Extensions of recurrent neural network language model. In 2011 IEEE international conference on acoustics, speech and signal processing (ICASSP)
Tomáš Mikolov, Stefan Kombrink, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur. 2011 · 2011
Earlier work this paper cites.
Feature engineering in context-dependent deep neural networks for conversational speech transcription. In 2011 IEEE Workshop on Automatic Speech Recognition & Understanding
Frank Seide, Gang Li, Xie Chen, and Dong Yu. 2011 · 2011
Cited alongside, same era.
Factorization Machines with libFM
Steffen Rendle. 2012 · 2012
Cited alongside, same era.
Counterfactual reasoning and learning systems: The example of computational advertising
Léon Bottou, Jonas Peters, Joaquin Quiñonero-Candela, Denis X Charles, D Max Chickering, Elon Portugaly, Dipankar Ray, Patrice Simard, and Ed Snelson. 2013 · 2013
Cited alongside, same era.
Learning factored representations in a deep mixture of experts
David Eigen, Marc’Aurelio Ranzato, and Ilya Sutskever. 2013 · 2013
Cited alongside, same era.
Speeding up convolutional neural networks with low rank expansions
Max Jaderberg, Andrea Vedaldi, and Andrew Zisserman. 2014 · 2014
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer
Noam Shazeer, Azalia Mirhoseini, Krzysztof Maziarz, Andy Davis, Quoc Le, Geoffrey Hinton, and Jeff Dean. 2017 · 2017
Later among the works it cites.
Attention is all you need. In Advances in neural information processing systems
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Deep & Cross Network for Ad Click Predictions
Ruoxi Wang, Bin Fu, Gang Fu, and Mingliang Wang. 2017 · 2017
Later among the works it cites.
On compressing deep models by low rank and sparse decomposition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
Xiyu Yu, Tongliang Liu, Xinchao Wang, and Dacheng Tao. 2017 · 2017
Later among the works it cites.
Latent cross: Making use of context in recurrent recommender systems. In Proceedings of the Eleventh ACM International Conference on Web Search and Data Mining
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Learning polynomials with neural networks
Gregory Valiant. 2014 · 2014
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification. In Proceedings of the IEEE international conference on computer vision
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015 · 2015
Cited alongside, same era.
Deep learning in neural networks: An overview
Jürgen Schmidhuber. 2015 · 2015
Cited alongside, same era.
Wide & Deep Learning for Recommender Systems
Heng-Tze Cheng, Levent Koc, Jeremiah Harmsen, Tal Shaked, Tushar Chandra, Hrishi Aradhye, Glen Anderson, Greg Corrado, Wei Chai, Mustafa Ispir, et al · 2016
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole. 2016 · 2016
Cited alongside, same era.
Product-based neural networks for user response prediction. In 2016 IEEE 16th International Conference on Data Mining (ICDM)
Yanru Qu, Han Cai, Kan Ren, Weinan Zhang, Yong Yu, Ying Wen, and Jun Wang. 2016 · 2016
Cited alongside, same era.
Alex Beutel, Paul Covington, Sagar Jain, Can Xu, Jia Li, Vince Gatto, and Ed H Chi. 2018 · 2018
Later among the works it cites.
Adaptive mixture of low-rank factorizations for compact neural modeling
Ting Chen, Ji Lin, Tian Lin, Song Han, Chong Wang, and Denny Zhou. 2018 · 2018
Later among the works it cites.
xdeepfm: Combining explicit and implicit feature interactions for recommender systems. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining
Jianxun Lian, Xiaohuan Zhou, Fuzheng Zhang, Zhongxia Chen, Xing Xie, and Guangzhong Sun. 2018 · 2018
Later among the works it cites.
Adaptive Factorization Network: Learning Adaptive-Order Feature Interactions
Weiyu Cheng, Yanyan Shen, and Linpeng Huang. 2019 · 2019
Later among the works it cites.
Are we really making much progress? A worrying analysis of recent neural recommendation approaches. In Proceedings of the 13th ACM Conference on Recommender Systems
Maurizio Ferrari Dacrema, Paolo Cremonesi, and Dietmar Jannach. 2019 · 2019
Later among the works it cites.
A multiscale neural network based on hierarchical nested bases
Yuwei Fan, Jordi Feliu-Faba, Lin Lin, Lexing Ying, and Leonardo Zepeda-Núnez. 2019 · 2019
Later among the works it cites.
Snr: Sub-network routing for flexible parameter sharing in multi-task learning. In Proceedings of the AAAI Conference on Artificial Intelligence
Jiaqi Ma, Zhe Zhao, Jilin Chen, Ang Li, Lichan Hong, and Ed H Chi. 2019 · 2019
Later among the works it cites.
Deep learning recommendation model for personalization and recommendation systems
Maxim Naumov, Dheevatsa Mudigere, Hao-Jun Michael Shi, Jianyu Huang, Narayanan Sundaraman, Jongsoo Park, Xiaodong Wang, Udit Gupta, Carole-Jean Wu, Alisson G Azzolini, et al · 2019
Later among the works it cites.
Autoint: Automatic feature interaction learning via self-attentive neural networks. In Proceedings of the 28th ACM International Conference on Information and Knowledge Management
Weiping Song, Chence Shi, Zhiping Xiao, Zhijian Duan, Yewen Xu, Ming Zhang, and Jian Tang. 2019 · 2019
Later among the works it cites.
Block Basis Factorization for Scalable Kernel Evaluation
Ruoxi Wang, Yingzhou Li, Michael W Mahoney, and Eric Darve. 2019 · 2019
Later among the works it cites.
Interpretable Click-Through Rate Prediction through Hierarchical Attention. In Proceedings of the 13th International Conference on Web Search and Data Mining
Zeyu Li, Wei Cheng, Yang Chen, Haifeng Chen, and Wei Wang. 2020 · 2020
Closest in time.
A metric learning reality check
Kevin Musgrave, Serge Belongie, and Ser-Nam Lim. 2020 · 2020
Closest in time.
Neural Collaborative Filtering vs. Matrix Factorization Revisited
Steffen Rendle, Walid Krichene, Li Zhang, and John Anderson. 2020 · 2020
Closest in time.