Fetching the paper…
Reading the bibliography…
The widespread application of deep learning has changed the landscape of computation in the data center.
1906
Earlier work this paper cites.
1912
Earlier work this paper cites.
D. M. Tullsen, S. J. Eggers, and H. M. Levy, “Simultaneous multithreading: Maximizing on-chip parallelism,” in Proceedings of the ACM SIGARCH Computer Architecture News , vol. 23, no. 2. ACM, 1995, pp. 392–403
1995
Earlier work this paper cites.
D. M. Tullsen, S. J. Eggers, J. S. Emer, H. M. Levy, J. L. Lo, and R. L. Stamm, “Exploiting choice: Instruction fetch and issue on an implementable simultaneous multithreading processor,” ACM SIGARCH Computer Architecture News , vol. 24, no. 2, pp. 191–202, 1996
1996
Earlier work this paper cites.
E. Ebrahimi, O. Mutlu, C. J. Lee, and Y. N. Patt, “Coordinated control of multiple prefetchers in multi-core systems,” in Proceedings of the 42Nd Annual IEEE/ACM International Symposium on Microarchitecture , ser. MICRO 42, 2009, pp. 316–326
2009
Earlier work this paper cites.
A. Jaleel, E. Borch, M. Bhandaru, S. C. Steely Jr, and J. Emer, “Achieving non-inclusive cache performance with inclusive caches: Temporal locality aware (tla) cache management policies,” in Proceedings of the 43rd Annual IEEE/ACM International Symposium on Microarchitecture . IEEE Computer Society, 2010, pp. 151–162
2010
Earlier work this paper cites.
C.-J. Wu, A. Jaleel, M. Martonosi, S. C. Steely, Jr., and J. Emer, “PACMan: Prefetch-aware cache management for high performance caching,” in Proceedings of the 44th Annual IEEE/ACM International Symposium on Microarchitecture , ser. MICRO-44, 2011
2011
Earlier work this paper cites.
C.-J. Wu and M. Martonosi, “Characterization and dynamic mitigation of intra-application cache interference,” in Proceedings of the IEEE International Symposium on Performance Analysis of Systems and Software , ser. ISPASS ’11, 2011
2011
Earlier work this paper cites.
N. Beckmann and D. Sanchez, “Jigsaw: Scalable software-defined caches,” in Proceedings of the 22nd international conference on Parallel architectures and compilation techniques . IEEE, 2013, pp. 213–224
2013
Earlier work this paper cites.
J. Dean and L. A. Barroso, “The tail at scale,” Communications of the ACM , vol. 56, no. 2, pp. 74–80, 2013
2013
Earlier work this paper cites.
J. Mars and L. Tang, “Whare-map: Heterogeneity in homogeneous warehouse-scale computers,” in Proceedings of the ACM SIGARCH Computer Architecture News , vol. 41, no. 3. ACM, 2013, pp. 619–630
2013
Earlier work this paper cites.
“Crieto dataset.” [Online]. Available: https://labs.criteo.com/2014/02/download-dataset/
2014
Earlier work this paper cites.
Y. Chen, T. Luo, S. Liu, S. Zhang, L. He, J. Wang, L. Li, T. Chen, Z. Xu, N. Sun et al. , “Dadiannao: A machine-learning supercomputer,” in Proceedings of the 47th Annual IEEE/ACM International Symposium on Microarchitecture . IEEE Computer Society, 2014, pp. 609–622
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
C. A. Gomez-Uribe and N. Hunt, “The netflix recommender system: Algorithms, business value, and innovation,” ACM Trans. Manage. Inf. Syst. , vol. 6, no. 4, pp. 13:1–13:19, Dec. 2015. [Online]. Available: http://doi.acm.org/10.1145/2843948
2015
Earlier work this paper cites.
S. Gupta, A. Agrawal, K. Gopalakrishnan, and P. Narayanan, “Deep learning with limited numerical precision,” in Proceedings of the International Conference on Machine Learning , 2015, pp. 1737–1746
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
A. Jaleel, J. Nuzman, A. Moga, S. C. Steely, and J. Emer, “High performing cache hierarchies for server workloads: Relaxing inclusion to capture the latency benefits of exclusive caches,” in Proceedings of the IEEE 21st International Symposium on High Performance Computer Architecture (HPCA) . IEEE, 2015, pp. 343–353
2015
Cited alongside, same era.
S. Kanev, J. P. Darago, K. Hazelwood, P. Ranganathan, T. Moseley, G.-Y. Wei, and D. Brooks, “Profiling a warehouse-scale computer,” in Proceedings of the ACM SIGARCH Computer Architecture News , vol. 43, no. 3. ACM, 2015, pp. 158–169
2015
Cited alongside, same era.
M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard et al. , “Tensorflow: A system for large-scale machine learning,” in Proceedings of the 12th USENIX Symposium on Operating Systems Design and Implementation ( O S D I OSDI 16) , 2016, pp. 265–283
2016
Cited alongside, same era.
2017
Later among the works it cites.
N. P. Jouppi, C. Young, N. Patil, D. Patterson, G. Agrawal, R. Bajwa, S. Bates, S. Bhatia, N. Boden, A. Borchers et al. , “In-datacenter performance analysis of a tensor processing unit,” in Proceedings of the ACM/IEEE 44th Annual International Symposium on Computer Architecture (ISCA) . IEEE, 2017, pp. 1–12
2017
Later among the works it cites.
R. Wang, B. Fu, G. Fu, and M. Wang, “Deep & cross network for ad click predictions,” in Proceedings of the ADKDD’17 . ACM, 2017, p. 12
2017
Later among the works it cites.
M. Chui, J. Manyika, M. Miremadi, N. Henke, R. Chung, P. Nel, and S. Malhotra, “Notes from the AI frontier insights from hundreds of use cases,” 2018
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Adolf, S. Rama, B. Reagen, G.-Y. Wei, and D. Brooks, “Fathom: Reference workloads for modern deep learning methods,” in Proceedings of the IEEE International Symposium on Workload Characterization (IISWC) . IEEE, 2016, pp. 1–10
2016
Cited alongside, same era.
D. Amodei, S. Ananthanarayanan, R. Anubhai, J. Bai, E. Battenberg, C. Case, J. Casper, B. Catanzaro, Q. Cheng, G. Chen et al. , “Deep speech 2: End-to-end speech recognition in english and mandarin,” in Proceedings of the International conference on machine learning , 2016, pp. 173–182
2016
Cited alongside, same era.
H.-T. Cheng, L. Koc, J. Harmsen, T. Shaked, T. Chandra, H. Aradhye, G. Anderson, G. Corrado, W. Chai, M. Ispir et al. , “Wide & deep learning for recommender systems,” in Proceedings of the 1st Workshop on Deep Learning for Recommender Systems . ACM, 2016, pp. 7–10
2016
Cited alongside, same era.
2016
Cited alongside, same era.
P. Covington, J. Adams, and E. Sargin, “Deep neural networks for youtube recommendations,” in Proceedings of the 10th ACM conference on recommender systems . ACM, 2016, pp. 191–198
2016
Cited alongside, same era.
S. Han, X. Liu, H. Mao, J. Pu, A. Pedram, M. A. Horowitz, and W. J. Dally, “EIE: Efficient inference engine on compressed deep neural network,” in Proceedings of the ACM/IEEE 43rd Annual International Symposium on Computer Architecture (ISCA) . IEEE, 2016, pp. 243–254
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Cited alongside, same era.
P. Judd, J. Albericio, T. Hetherington, T. M. Aamodt, and A. Moshovos, “Stripes: Bit-serial deep neural network computing,” in Proceedings of the 49th Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) . IEEE, 2016, pp. 1–12
2016
Cited alongside, same era.
B. Reagen, P. Whatmough, R. Adolf, S. Rama, H. Lee, S. K. Lee, J. M. Hernández-Lobato, G.-Y. Wei, and D. Brooks, “Minerva: Enabling low-power, highly-accurate deep neural network accelerators,” in Proceedings of the ACM/IEEE 43rd Annual International Symposium on Computer Architecture (ISCA) . IEEE, 2016, pp. 267–278
2016
Cited alongside, same era.
E. Chung, J. Fowers, K. Ovtcharov, M. Papamichael, A. Caulfield, T. Massengill, M. Liu, D. Lo, S. Alkalay, M. Haselman et al. , “Serving DNNs in real time at datacenter scale with project brainwave,” IEEE Micro , vol. 38, no. 2, pp. 8–20, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
K. Hazelwood, S. Bird, D. Brooks, S. Chintala, U. Diril, D. Dzhulgakov, M. Fawzy, B. Jia, Y. Jia, A. Kalro et al. , “Applied machine learning at Facebook: a datacenter infrastructure perspective,” in Proceedings of the IEEE International Symposium on High Performance Computer Architecture (HPCA) . IEEE, 2018, pp. 620–629
2018
Later among the works it cites.
2018
Later among the works it cites.
X. Xie, J. Lian, Z. Liu, X. Wang, F. Wu, H. Wang, and Z. Chen, “Personalized recommendation systems: Five hot research topics you must know,” 2018. [Online]. Available: https://www.microsoft.com/en-us/research/lab/microsoft-research-asia/articles/personalized-recommendation-systems/
2018
Later among the works it cites.
L. Yang, E. Bagdasaryan, J. Gruenstein, C.-K. Hsieh, and D. Estrin, “OpenRec: A modular framework for extensible and adaptable recommendation algorithms,” in Proceedings of the 11th ACM International Conference on Web Search and Data Mining . ACM, 2018, pp. 664–672
2018
Later among the works it cites.
Y. You, Z. Zhang, C.-J. Hsieh, J. Demmel, and K. Keutzer, “Imagenet training in minutes,” in Proceedings of the 47th International Conference on Parallel Processing . ACM, 2018, p. 1
2018
Later among the works it cites.
H. Zhu, M. Akrout, B. Zheng, A. Pelegris, A. Jayarajan, A. Phanishayee, B. Schroeder, and G. Pekhimenko, “Benchmarking and analyzing deep neural network training,” in Proceedings of the IEEE International Symposium on Workload Characterization (IISWC) . IEEE, 2018, pp. 88–100
2018
Later among the works it cites.
Y. Kwon, Y. Lee, and M. Rhu, “TensorDIMM: A practical near-memory processing architecture for embeddings and tensor operations in deep learning,” in Proceedings of the 52nd Annual IEEE/ACM International Symposium on Microarchitecture . ACM, 2019, pp. 740–753
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
M. Naumov and D. Mudigere, “Deep learning recommendation model for personalization and recommendation systems,” 2019. [Online]. Available: https://ai.facebook.com/blog/dlrm-an-advanced-open-source-deep-learning-recommendation-model/
2019
Closest in time.
C. Underwood, “Use cases of recommendation systems in business – current applications and methods,” 2019. [Online]. Available: https://emerj.com/ai-sector-overviews/use-cases-recommendation-systems/
2019
Closest in time.
C.-J. Wu, D. Brooks, K. Chen, D. Chen, S. Choudhury, M. Dukhan, K. Hazelwood, E. Isaac, Y. Jia, B. Jia, T. Leyvand, H. Lu, Y. Lu, L. Qiao, B. Reagen, J. Spisak, F. Sun, A. Tulloch, P. Vajda, X. Wang, Y. Wang, B. Wasti, Y. Wu, R. Xian, S. Yoo, and P. Zhang, “Machine learning at Facebook: Understanding inference at the edge,” in Proceedings of the IEEE International Symposium on High Performance Computer Architecture (HPCA) , Feb 2019, pp. 331–344
2019
Closest in time.