Fetching the paper…
Reading the bibliography…
Embedding learning of categorical features (e.g.
A mathematical theory of communication
Claude E. Shannon. 1948 · 1948
Earlier work this paper cites.
A note on the generation of random normal deviates
George EP Box. 1958 · 1958
Earlier work this paper cites.
Space/Time Trade-offs in Hash Coding with Allowable Errors
Burton H. Bloom. 1970 · 1970
Earlier work this paper cites.
Universal Classes of Hash Functions (Extended Abstract). In STOC
Larry Carter and Mark N. Wegman. 1977 · 1977
Earlier work this paper cites.
Locality-sensitive hashing scheme based on p-stable distributions. In SoCG . ACM, 253–262
Mayur Datar, Nicole Immorlica, Piotr Indyk, and Vahab S. Mirrokni. 2004 · 2004
Earlier work this paper cites.
Implicit Neural Representations with Periodic Activation Functions
Vincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell, and Gordon Wetzstein. 2020 · 2006
Earlier work this paper cites.
BPR: Bayesian Personalized Ranking from Implicit Feedback. In UAI
Steffen Rendle, Christoph Freudenthaler, Zeno Gantner, and Lars Schmidt-Thieme. 2009 · 2009
Earlier work this paper cites.
Feature hashing for large scale multitask learning. In ICML
Kilian Q. Weinberger, Anirban Dasgupta, John Langford, Alexander J. Smola, and Josh Attenberg. 2009 · 2009
Earlier work this paper cites.
Mish: A Self Regularized Non-Monotonic Neural Activation Function. In BMVC
Diganta Misra. 2010 · 2010
Earlier work this paper cites.
Distributed Representations of Words and Phrases and their Compositionality. In NIPS
Tomas Mikolov, Ilya Sutskever, Kai Chen, Gregory S. Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Generative Adversarial Nets. In NIPS
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron C. Courville, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. In ICML (JMLR Workshop and Conference Proceedings, Vol. 37) . JMLR.org
Sergey Ioffe and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
Deep Neural Networks for YouTube Recommendations. In RecSys . ACM, 191–198
Paul Covington, Jay Adams, and Emre Sargin. 2016 · 2016
Cited alongside, same era.
The MovieLens Datasets: History and Context
F. Maxwell Harper and Joseph A. Konstan. 2016 · 2016
Cited alongside, same era.
Deep Residual Learning for Image Recognition. In CVPR . IEEE Computer Society, 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Densely Connected Convolutional Networks
Gao Huang, Zhuang Liu, and Kilian Q. Weinberger. 2016 · 2016
Cited alongside, same era.
Neural Machine Translation of Rare Words with Subword Units. In ACL
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
DeepFM: A Factorization-Machine based Neural Network for CTR Prediction. In IJCAI . ijcai.org
The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks. In ICLR . OpenReview.net
Jonathan Frankle and Michael Carbin. 2019 · 2019
Later among the works it cites.
Candidate Generation with Binary Codes for Large-Scale Top-N Recommendation. In CIKM . ACM
Wang-Cheng Kang and Julian John McAuley. 2019 · 2019
Later among the works it cites.
PRADO: Projection Attention Networks for Document Classification On-Device. In EMNLP-IJCNLP . Association for Computational Linguistics, 5011–5020
Karthik Krishnamoorthi, Sujith Ravi, and Zornitsa Kozareva. 2019 · 2019
Later among the works it cites.
Justifying Recommendations using Distantly-Labeled Reviews and Fine-Grained Aspects. In EMNLP-IJCNLP . Association for Computational Linguistics
Jianmo Ni, Jiacheng Li, and Julian J. McAuley. 2019 · 2019
Later among the works it cites.
Efficient On-Device Models using Neural Projections. In ICML
Sujith Ravi. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Huifeng Guo, Ruiming Tang, Yunming Ye, Zhenguo Li, and Xiuqiang He. 2017 · 2017
Cited alongside, same era.
Why Deep Neural Networks for Function Approximation?. In ICLR . OpenReview.net
Shiyu Liang and R. Srikant. 2017 · 2017
Cited alongside, same era.
The Expressive Power of Neural Networks: A View from the Width. In NIPS
Zhou Lu, Hongming Pu, Feicheng Wang, Zhiqiang Hu, and Liwei Wang. 2017 · 2017
Cited alongside, same era.
Getting Deep Recommenders Fit: Bloom Embeddings for Sparse Binary Input/Output Networks. In RecSys . ACM
Joan Serrà and Alexandros Karatzoglou. 2017 · 2017
Cited alongside, same era.
Hash Embeddings for Efficient Word Representations. In NIPS
Dan Svenstrup, Jonas Meinertz Hansen, and Ole Winther. 2017 · 2017
Cited alongside, same era.
Understanding Deep Neural Networks with Rectified Linear Units. In ICLR
Raman Arora, Amitabh Basu, Poorya Mianjy, and Anirbit Mukherjee. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In NAACL-HLT . Association for Computational Linguistics
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Sampling-bias-corrected neural modeling for large corpus item recommendations. In RecSys . ACM
Xinyang Yi, Ji Yang, Lichan Hong, Derek Zhiyuan Cheng, Lukasz Heldt, Aditee Kumthekar, Zhe Zhao, Li Wei, and Ed H. Chi. 2019 · 2019
Later among the works it cites.
Neural Input Search for Large Scale Recommendation Models. In SIGKDD . ACM
Manas R. Joglekar, Cong Li, Mei Chen, Taibai Xu, Xiaoming Wang, Jay K. Adams, Pranav Khaitan, Jiahui Liu, and Quoc V. Le. 2020 · 2020
Closest in time.
ALBERT: A Lite BERT for Self-supervised Learning of Language Representations. In ICLR . OpenReview.net
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2020 · 2020
Closest in time.
Learnable Embedding Sizes for Recommender Systems. In ICLR
Siyi Liu, Chen Gao, Yihong Chen, Depeng Jin, and Yong Li. 2020 · 2020
Closest in time.
Compositional Embeddings Using Complementary Partitions for Memory-Efficient Recommendation Systems. In SIGKDD
Hao-Jun Michael Shi, Dheevatsa Mudigere, Maxim Naumov, and Jiyan Yang. 2020 · 2020
Closest in time.
Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional Domains. In NeurIPS
Matthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil, Nithin Raghavan, Utkarsh Singhal, Ravi Ramamoorthi, Jonathan T. Barron, and Ren Ng. 2020 · 2020
Closest in time.
Model Size Reduction Using Frequency Based Double Hashing for Recommender Systems. In RecSys
Caojin Zhang, Yicun Liu, Yuanpu Xie, Sofia Ira Ktena, Alykhan Tejani, Akshay Gupta, Pranay Kumar Myana, Deepak Dilipkumar, Suvadip Paul, Ikuhiro Ihara, Prasang Upadhyaya, Ferenc Huszar, and Wenzhe Shi. 2020 · 2020
Closest in time.