Fetching the paper…
Reading the bibliography…
Embedding models have been an effective learning paradigm for high-dimensional data.
Time, Clocks, and the Ordering of Events in a Distributed System
Leslie Lamport. 1978 · 1978
Earlier work this paper cites.
MPI: A Message-Passing Interface Standard
Message P Forum. 1994 · 1994
Earlier work this paper cites.
Impact of data characteristics on recommender systems performance
Gediminas Adomavicius and Jingjing Zhang. 2012 · 2012
Earlier work this paper cites.
Distributed GraphLab: A Framework for Machine Learning in the Cloud
Yucheng Low, Joseph Gonzalez, Aapo Kyrola, Danny Bickson, Carlos Guestrin, and Joseph M. Hellerstein. 2012 · 2012
Earlier work this paper cites.
More Effective Distributed ML via a Stale Synchronous Parallel Parameter Server. In NeurIPS . 1223–1231
Qirong Ho, James Cipar, Henggang Cui, Seunghak Lee, Jin Kyu Kim, Phillip B. Gibbons, Garth A. Gibson, Gregory R. Ganger, and Eric P. Xing. 2013 · 2013
Earlier work this paper cites.
Efficient Estimation of Word Representations in Vector Space. In ICLR Workshop
Tomás Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Criteo Kaggle Ad
2014 · 2014
Earlier work this paper cites.
Scaling Distributed Machine Learning with the Parameter Server. In OSDI . 583–598
Mu Li, David G. Andersen, Jun Woo Park, Alexander J. Smola, Amr Ahmed, Vanja Josifovski, James Long, Eugene J. Shekita, and Bor-Yiing Su. 2014 · 2014
Earlier work this paper cites.
DeepWalk: online learning of social representations. In SIGKDD . 701–710
Bryan Perozzi, Rami Al-Rfou, and Steven Skiena. 2014 · 2014
Earlier work this paper cites.
Document Embedding with Paragraph Vectors
Andrew M. Dai, Christopher Olah, and Quoc V. Le. 2015 · 2015
Earlier work this paper cites.
Asynchronous Parallel Stochastic Gradient for Nonconvex Optimization. In NeurIPS . 2737–2745
Xiangru Lian, Yijun Huang, Yuncheng Li, and Ji Liu. 2015 · 2015
Earlier work this paper cites.
GeoSoCa: Exploiting Geographical, Social and Categorical Correlations for Point-of-Interest Recommendations. In SIGIR . 443–452
Jia-Dong Zhang and Chi-Yin Chow. 2015 · 2015
Earlier work this paper cites.
TensorFlow: A System for Large-Scale Machine Learning. In OSDI . 265–283
Martín Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, Michael Isard, Manjunath Kudlur, Josh Levenberg, Rajat Monga, Sherry Moore, Derek Gordon Murray, Benoit Steiner, Paul A. Tucker, Vijay Vasudevan, Pete Warden, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng. 2016 · 2016
Earlier work this paper cites.
Wide & Deep Learning for Recommender Systems. In DLRS@RecSys . 7–10
Heng-Tze Cheng, Levent Koc, Jeremiah Harmsen, Tal Shaked, Tushar Chandra, Hrishi Aradhye, Glen Anderson, Greg Corrado, Wei Chai, Mustafa Ispir, Rohan Anil, Zakaria Haque, Lichan Hong, Vihan Jain, Xiaobing Liu, and Hemal Shah. 2016 · 2016
Earlier work this paper cites.
Deep Neural Networks for YouTube Recommendations. In RecSys . 191–198
Paul Covington, Jay Adams, and Emre Sargin. 2016 · 2016
Earlier work this paper cites.
Mini-batch stochastic approximation methods for nonconvex stochastic composite optimization
Saeed Ghadimi, Guanghui Lan, and Hongchao Zhang. 2016 · 2016
Earlier work this paper cites.
STRADS: a distributed framework for scheduled model parallel machine learning. In EuroSys . ACM, 5:1–5:16
Jin Kyu Kim, Qirong Ho, Seunghak Lee, Xun Zheng, Wei Dai, Garth A. Gibson, and Eric P. Xing. 2016 · 2016
Earlier work this paper cites.
Strategies and principles of distributed machine learning on big data
Eric P Xing, Qirong Ho, Pengtao Xie, and Dai Wei. 2016 · 2016
Cited alongside, same era.
DeepFM: A Factorization-Machine based Neural Network for CTR Prediction. In IJCAI . 1725–1731
Huifeng Guo, Ruiming Tang, Yunming Ye, Zhenguo Li, and Xiuqiang He. 2017 · 2017
Cited alongside, same era.
Inductive Representation Learning on Large Graphs. In NeurIPS . 1024–1034
William L. Hamilton, Zhitao Ying, and Jure Leskovec. 2017 · 2017
Cited alongside, same era.
Heterogeneity-aware Distributed Parameter Servers. In SIGMOD . 463–478
Jiawei Jiang, Bin Cui, Ce Zhang, and Lele Yu. 2017 · 2017
Cited alongside, same era.
Deep & Cross Network for Ad Click Predictions. In ADKDD . 12:1–12:7
Ruoxi Wang, Bin Fu, Gang Fu, and Mingliang Wang. 2017 · 2017
Cited alongside, same era.
LDA*: A Robust and Large-scale Topic Modeling System
MLPerf Benchmark
2020 · 2020
Later among the works it cites.
Deep Learning for User Interest and Response Prediction in Online Display Advertising
Zhabiz Gharibshah, Xingquan Zhu, Arthur Hainline, and Michael Conway. 2020 · 2020
Later among the works it cites.
Bi-Labeled LDA: Inferring Interest Tags for Non-famous Users in Social Network
Jun He, Hongyan Liu, Yiqing Zheng, Shu Tang, Wei He, and Xiaoyong Du. 2020 · 2020
Later among the works it cites.
Open Graph Benchmark: Datasets for Machine Learning on Graphs. In NeurIPS
Weihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong, Hongyu Ren, Bowen Liu, Michele Catasta, and Jure Leskovec. 2020 · 2020
Later among the works it cites.
Learning Multi-granular Quantized Embeddings for Large-Vocab Categorical Features in Recommender Systems. In WWW . ACM / IW3C2, 562–566
Wang-Cheng Kang, Derek Zhiyuan Cheng, Ting Chen, Xinyang Yi, Dong Lin, Lichan Hong, and Ed H. Chi. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lele Yu, Bin Cui, Ce Zhang, and Yingxia Shao. 2017 · 2017
Cited alongside, same era.
Horovod: fast and easy distributed deep learning in TensorFlow
Alexander Sergeev and Mike Del Balso. 2018 · 2018
Cited alongside, same era.
Categorical-attributes-based item classification for recommender systems. In RecSys . ACM, 320–328
Qian Zhao, Jilin Chen, Minmin Chen, Sagar Jain, Alex Beutel, Francois Belletti, and Ed H. Chi. 2018 · 2018
Cited alongside, same era.
Deep Interest Network for Click-Through Rate Prediction. In SIGKDD . 1059–1068
Guorui Zhou, Xiaoqiang Zhu, Chengru Song, Ying Fan, Han Zhu, Xiao Ma, Yanghui Yan, Junqi Jin, Han Li, and Kun Gai. 2018 · 2018
Cited alongside, same era.
Asynchronous Training of Word Embeddings for Large Text Corpora. In WSDM . 168–176
Avishek Anand, Megha Khosla, Jaspreet Singh, Jan-Hendrik Zab, and Zijian Zhang. 2019 · 2019
Cited alongside, same era.
Cluster-GCN: An Efficient Algorithm for Training Deep and Large Graph Convolutional Networks. In SIGKDD . 257–266
Wei-Lin Chiang, Xuanqing Liu, Si Si, Yang Li, Samy Bengio, and Cho-Jui Hsieh. 2019 · 2019
Cited alongside, same era.
Parallax: Sparsity-aware Data Parallel Training of Deep Neural Networks. In EuroSys . 43:1–43:15
Soojeong Kim, Gyeong-In Yu, Hojin Park, Sungwoo Cho, Eunji Jeong, Hyeonmin Ha, Sanha Lee, Joo Seong Jeong, and Byung-Gon Chun. 2019 · 2019
Cited alongside, same era.
PyTorch Distributed: Experiences on Accelerating Data Parallel Training
Shen Li, Yanli Zhao, Rohan Varma, Omkar Salpekar, Pieter Noordhuis, Teng Li, Adam Paszke, Jeff Smith, Brian Vaughan, Pritam Damania, and Soumith Chintala. 2020 · 2020
Later among the works it cites.
CuWide: Towards Efficient Flow-based Training for Sparse Wide Models on GPUs
X. Miao, L. Ma, Z. Yang, Y. Shao, B. Cui, L. Yu, and J. Jiang. 2020 · 2020
Later among the works it cites.
Dynamic network embedding via incremental skip-gram with negative sampling
Hao Peng, Jianxin Li, Hao Yan, Qiran Gong, Senzhang Wang, Lin Liu, Lihong Wang, and Xiang Ren. 2020 · 2020
Later among the works it cites.
Kraken: memory-efficient continual learning for large-scale real-time recommendations. In SC . 1–17
Minhui Xie, Kai Ren, Youyou Lu, Guangxu Yang, Qingxing Xu, Bihai Wu, Jiazhen Lin, Hongbo Ao, Wanhong Xu, and Jiwu Shu. 2020 · 2020
Later among the works it cites.
General-Purpose User Embeddings based on Mobile App Usage. In SIGKDD . 2831–2840
Junqi Zhang, Bing Bai, Ye Lin, Jian Liang, Kun Bai, and Fei Wang. 2020 · 2020
Later among the works it cites.
Distributed Hierarchical GPU Parameter Server for Massive Scale Deep Learning Ads Systems. In MLSys
Weijie Zhao, Deping Xie, Ronglai Jia, Yulei Qian, Ruiquan Ding, Mingming Sun, and Ping Li. 2020 · 2020
Later among the works it cites.
HET Appendix
2021 · 2021
Closest in time.
NVIDIA collective communications library (NCCL)
2021 · 2021
Closest in time.
NVIDIA HugeCTR
2021 · 2021
Closest in time.
Syntax-guided text generation via graph neural network
Qipeng Guo, Xipeng Qiu, Xiangyang Xue, and Zheng Zhang. 2021 · 2021
Closest in time.
Lasagne: A Multi-Layer Graph Convolutional Network Framework via Node-aware Deep Architecture
Xupeng Miao, Wentao Zhang, Yingxia Shao, Bin Cui, Lei Chen, Ce Zhang, and Jiawei Jiang. 2021c · 2021
Closest in time.