Fetching the paper…
Reading the bibliography…
Many recent deep learning platforms rely on third-party libraries (such as cuBLAS) to utilize the computing power of modern hardware accelerators (such as GPUs).
J. R. Rice, “The algorithm selection problem,” Advances in computers , vol. 15, pp. 65–118, 1976
1976
Earlier work this paper cites.
J. R. Quinlan, “Induction of decision trees,” Machine learning , vol. 1, no. 1, pp. 81–106, 1986
1986
Earlier work this paper cites.
R. Hecht-Nielsen, “Theory of the backpropagation neural network,” in Neural Networks, 1989. IJCNN., International Joint Conference on . IEEE, 1989, pp. 593–605
1989
Earlier work this paper cites.
C. Cortes and V. Vapnik, “Support-vector networks,” Machine learning , vol. 20, no. 3, pp. 273–297, 1995
1995
Earlier work this paper cites.
J. A. Suykens and J. Vandewalle, “Least squares support vector machine classifiers,” Neural processing letters , vol. 9, no. 3, pp. 293–300, 1999
1999
Earlier work this paper cites.
J. H. Friedman, “Greedy function approximation: a gradient boosting machine,” Annals of statistics , pp. 1189–1232, 2001
2001
Earlier work this paper cites.
R. Caruana and A. Niculescu-Mizil, “An empirical comparison of supervised learning algorithms,” in Proceedings of the 23rd international conference on Machine learning . ACM, 2006, pp. 161–168
2006
Earlier work this paper cites.
V. Volkov and J. W. Demmel, “Benchmarking gpus to tune dense linear algebra,” in High Performance Computing, Networking, Storage and Analysis, 2008. SC 2008. International Conference for . IEEE, 2008, pp. 1–11
2008
Earlier work this paper cites.
Y. Li, J. Dongarra, and S. Tomov, “A note on auto-tuning gemm for gpus,” in International Conference on Computational Science . Springer, 2009, pp. 884–892
2009
Earlier work this paper cites.
G. Ruetsch and P. Micikevicius, “Optimizing matrix transpose in cuda,” Nvidia CUDA SDK Application Note , vol. 18, 2009
2009
Earlier work this paper cites.
M. N. Anyanwu and S. G. Shiva, “Comparative analysis of serial decision tree classification algorithms,” International Journal of Computer Science and Security , vol. 3, no. 3, pp. 230–240, 2009
2009
Cited alongside, same era.
R. Nath, S. Tomov, and J. Dongarra, “An improved magma gemm for fermi graphics processing units,” The International Journal of High Performance Computing Applications , vol. 24, no. 4, pp. 511–515, 2010
2010
Cited alongside, same era.
G. Tan, L. Li, S. Triechle, E. Phillips, Y. Bao, and N. Sun, “Fast implementation of dgemm on fermi gpu,” in Proceedings of 2011 International Conference for High Performance Computing, Networking, Storage and Analysis . ACM, 2011, p. 35
2011
Cited alongside, same era.
W.-Y. Loh, “Classification and regression trees,” Wiley Interdisciplinary Reviews: Data Mining and Knowledge Discovery , vol. 1, no. 1, pp. 14–23, 2011
2011
Cited alongside, same era.
——, C4. 5: programs for machine learning . Elsevier, 2014
2014
Later among the works it cites.
Y. LeCun et al. , “Lenet-5, convolutional neural networks,” URL: http://yann. lecun. com/exdb/lenet , 2015
2015
Later among the works it cites.
O. Spillinger, D. Eliahu, A. Fox, and J. Demmel, “Matrix multiplication algorithm selection with support vector machines,” EECS Department, University of California, Berkeley , 2015
2015
Later among the works it cites.
N. Sedaghati, T. Mu, L.-N. Pouchet, S. Parthasarathy, and P. Sadayappan, “Automatic selection of sparse matrix representation on gpus,” in Proceedings of the 29th ACM on International Conference on Supercomputing . ACM, 2015, pp. 99–108
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2011
Cited alongside, same era.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Cited alongside, same era.
J. Kurzak, S. Tomov, and J. Dongarra, “Autotuning gemm kernels for the fermi gpu,” IEEE Transactions on Parallel and Distributed Systems , vol. 23, no. 11, pp. 2045–2057, 2012
2012
Cited alongside, same era.
T. W. Hungerford, Abstract algebra: an introduction . Cengage Learning, 2012
2012
Cited alongside, same era.
J. Lai and A. Seznec, “Performance upper bound analysis and optimization of sgemm on fermi and kepler gpus,” in Code Generation and Optimization (CGO), 2013 IEEE/ACM International Symposium on . IEEE, 2013, pp. 1–10
2013
Cited alongside, same era.
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell, “Caffe: Convolutional architecture for fast feature embedding,” in Proceedings of the 22nd ACM international conference on Multimedia , 2014, pp. 675–678
2014
Cited alongside, same era.
2016
Later among the works it cites.
A. Abdelfattah, A. Haidar, S. Tomov, and J. Dongarra, “Performance, design, and autotuning of batched gemm for gpus,” in International Conference on High Performance Computing . Springer, 2016, pp. 21–38
2016
Later among the works it cites.
A. Benatia, W. Ji, Y. Wang, and F. Shi, “Sparse matrix format selection with multiclass svm for spmv on gpu,” in Parallel Processing (ICPP), 2016 45th International Conference on . IEEE, 2016, pp. 496–505
2016
Later among the works it cites.
J. Gomez-Luna, I.-J. Sung, L.-W. Chang, J. M. González-Linares, N. Guil, and W.-M. W. Hwu, “In-place matrix transposition on gpus,” IEEE Transactions on Parallel and Distributed Systems , vol. 27, no. 3, pp. 776–788, 2016
2016
Later among the works it cites.
T. Chen and C. Guestrin, “Xgboost: Reliable large-scale tree boosting system,” in Proceedings of the 22nd SIGKDD Conference on Knowledge Discovery and Data Mining, San Francisco, CA, USA , 2016, pp. 13–17
2016
Later among the works it cites.
NVIDIA, “cublas — nvidia,” https://developer.nvidia.com/cublas , 2017, accessed: 2017-02-20
2017
Closest in time.