Fetching the paper…
Reading the bibliography…
Deep neural networks have been applied in many applications exhibiting extraordinary abilities in the field of computer vision.
The State of Sparsity in Deep Neural Networks
Gale, T., Elsen, E., Hooker, S., 2019 · 1902
Earlier work this paper cites.
Liu, Z., Mu, H., Zhang, X., Guo, Z., Yang, X., Cheng, T.K.T., Sun, J., 2019a · 1903
Earlier work this paper cites.
DeepShift: Towards Multiplication-Less Neural Networks
Elhoushi, M., Chen, Z., Shafiq, F., Tian, Y.H., Li, J.Y., 2019 · 1905
Earlier work this paper cites.
Once-for-All: Train One Network and Specialize it for Efficient Deployment
Cai, H., Gan, C., Wang, T., Zhang, Z., Han, S., 2019 · 1908
Earlier work this paper cites.
Exploiting Channel Similarity for Accelerating Deep Convolutional Neural Networks
Zhang, Y., Zhao, C., Ni, B., Zhang, J., Deng, H., 2019b · 1908
Earlier work this paper cites.
Chen, H., Wang, Y., Xu, C., Shi, B., Xu, C., Tian, Q., Xu, C., 2020 · 1912
Earlier work this paper cites.
The State of Knowledge Distillation for Classification
Ruffy, F., Chahal, K., 2019 · 1912
Earlier work this paper cites.
Zhu, F., Gong, R., Yu, F., Liu, X., Wang, Y., Li, Z., Yang, X., Yan, J., · 1912
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, N., Hinton, G., …, A.K.T.j.o.m., 2014, U., 2014 · 1958
Earlier work this paper cites.
Coincidence Approach To Stochastic Point Process
Macchi, O., 1975 · 1975
Earlier work this paper cites.
Neocognitron: A hierarchical neural network capable of visual pattern recognition
Fukushima, K., 1988 · 1988
Earlier work this paper cites.
Comparing biases for minimal network construction with back-propagation, in: Advances in Neural Information Processing Systems (NIPS), pp. 177–185
HANSON, S., 1989 · 1989
Earlier work this paper cites.
A set of level 3 basic linear algebra subprograms
Dongarra, J.J., Du Croz, J., Hammarling, S., Duff, I.S., 1990 · 1990
Earlier work this paper cites.
Weight discretization paradigm for optical neural networks
Fiesler, E., Choudry, A., Caulfield, H.J., 1990 · 1990
Earlier work this paper cites.
Optimal Brain Damage, in: Advances in Neural Information Processing Systems (NIPS), p. 598–605
LeCun, Y., Denker, J.S., Solla, S.A., 1990 · 1990
Earlier work this paper cites.
Training Feed Forward Nets with Binary Weights Via a Modified CHIR Algorithm
Saad, D., Marom, E., 1990 · 1990
Earlier work this paper cites.
Weight quantization in Boltzmann machines
Balzer, W., Takahashi, M., Ohta, J., Kyuma, K., 1991 · 1991
Earlier work this paper cites.
Optimal brain surgeon and general network pruning
Hassibi, B., Stork, D.G., Wolff, G.J., 1993 · 1993
Earlier work this paper cites.
Pruning Algorithms - A Survey
Reed, R., 1993 · 1993
Earlier work this paper cites.
Parsimonious network design and feature selection through node pruning, in: Proceedings of the 12th IAPR International Conference on Pattern Recognition, Vol. 3 - Conference C: Signal Processing (Cat. No.94CH3440-5), IEEE Comput. Soc. Press. pp. 622–624
Jianchang Mao, Mohiuddin, K., Jain, A., 1994 · 1994
Earlier work this paper cites.
Regression shrinkage and selection via the Lasso
Tishbirani, R., 1996 · 1996
Earlier work this paper cites.
Low Weight and Fan-In Neural Networks for Basic Arithmetic Operations, in: 15th IMACS World Congress, pp. 227–232
Cotofana, S., Vassiliadis, S., Logic, T., Addition, B., Addition, S., 1997 · 1997
Earlier work this paper cites.
Feature selection for classification
Dash, M., Liu, H., 1997 · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., Haffner, P., 1998 · 1998
Earlier work this paper cites.
MeliusNet: Can Binary Neural Networks Achieve MobileNet-level Accuracy?
Bethge, J., Bartz, C., Yang, H., Chen, Y., Meinel, C., 2020 · 2001
Earlier work this paper cites.
A new pruning heuristic based on variance analysis of sensitivity information
Engelbrecht, A.P., 2001 · 2001
Earlier work this paper cites.
MLIR: A Compiler Infrastructure for the End of Moore’s Law
Lattner, C., Amini, M., Bondhugula, U., Cohen, A., Davis, A., Pienaar, J., Riddle, R., Shpeisman, T., Vasilache, N., Zinenko, O., 2020 · 2002
Earlier work this paper cites.
The Deep Learning Compiler: A Comprehensive Survey
Li, M., Liu, Y.I., Liu, X., Sun, Q., You, X.I.N., Yang, H., Luan, Z., Gan, L., Yang, G., Qian, D., 2020a · 2002
Earlier work this paper cites.
What is the State of Neural Network Pruning?
Blalock, D., Ortiz, J.J.G., Frankle, J., Guttag, J., 2020 · 2003
Earlier work this paper cites.
Language Models are Few-Shot Learners
Brown, T.B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D.M., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., Amodei, D., 2020 · 2005
Earlier work this paper cites.
Model selection and estimation in regression with grouped variables
Yuan, M., Lin, Y., 2006 · 2005
Earlier work this paper cites.
Model compression, in: Proceedings of the 12th ACM SIGKDD international conference on Knowledge discovery and data mining - KDD ’06, ACM Press, New York, New York, USA. p. 535
Buciluǎ, C., Caruana, R., Niculescu-Mizil, A., 2006 · 2006
Earlier work this paper cites.
High Performance Convolutional Neural Networks for Document Processing, in: Tenth International Workshop on Frontiers in Handwriting Recognition
Chellapilla, K., Puri, S., Simard, P., 2006 · 2006
Earlier work this paper cites.
Knowledge Distillation: A Survey
Gou, J., Yu, B., Maybank, S.J., Tao, D., 2020 · 2006
Earlier work this paper cites.
An Overview of Neural Network Compression
Neill, J.O., 2020 · 2006
Earlier work this paper cites.
Towards Lower Bit Multiplication for Convolutional Neural Network Training
Zhong, K., Zhao, T., Ning, X., Zeng, S., Guo, K., Wang, Y., Yang, H., 2020 · 2006
Earlier work this paper cites.
Neon technology introduction
ARM, Reddy, V.G., 2008 · 2008
Earlier work this paper cites.
Solving local minima problem with large number of hidden nodes on two-layered feed-forward artificial neural networks
Choi, B., Lee, J.H., Kim, D.H., 2008 · 2008
Earlier work this paper cites.
IEEE Standard for Floating-Point Arithmetic
Society, I.C., Committee, M.S., 2008 · 2008
Earlier work this paper cites.
Learning Structured Sparsity in Deep Neural Networks, in: Advances in Neural Information Processing Systems (NIPS), IEEE. pp. 2074–2082
Wen, W., Wu, C., Wang, Y., Chen, Y., Li, H., 2016 · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Socher, R., Li-Jia Li, Kai Li, Li Fei-Fei, 2009 · 2009
Earlier work this paper cites.
Learning Multiple Layers of Features from Tiny Images
Krizhevsky, A., 2009 · 2009
Earlier work this paper cites.
Improving the speed of neural networks on CPUs URL: https://research.google/pubs/pub37631/
Vanhoucke, V., Senior, A., Mao, M.Z., 2011 · 2011
Earlier work this paper cites.
Neural networks for machine learning
Hinton, G., 2012 · 2012
Earlier work this paper cites.
Improving neural networks by preventing co-adaptation of feature detectors
Hinton, G.E., Srivastava, N., Krizhevsky, A., Sutskever, I., Salakhutdinov, R.R., 2012 · 2012
Earlier work this paper cites.
Tensor Completion for Estimating Missing Values in Visual Data
Liu, J., Musialski, P., Wonka, P., Ye, J., 2013 · 2012
Earlier work this paper cites.
Pruning algorithms of neural networks - A comparative study
Augasta, M.G., Kathirvalavakumar, T., 2013 · 2013
Earlier work this paper cites.
Estimating or Propagating Gradients Through Stochastic Neurons
Bengio, Y., 2013 · 2013
Earlier work this paper cites.
Davis, A., Arel, I., 2013 · 2013
Earlier work this paper cites.
Group sparse optimization by alternating direction method, in: Van De Ville, D., Goyal, V.K., Papadakis, M. (Eds.), Wavelets and Sparsity XV, p. 88580R
Deng, W., Yin, W., Zhang, Y., 2013 · 2013
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J., 2015 · 2013
Earlier work this paper cites.
Fast Training of Convolutional Networks through FFTs
Mathieu, M., Henaff, M., LeCun, Y., 2013 · 2013
Earlier work this paper cites.
Sermanet, P., Eigen, D., Zhang, X., Mathieu, M., Fergus, R., LeCun, Y., 2013 · 2013
Earlier work this paper cites.
Convolutional Neural Networks for Speech Recognition
Abdel-Hamid, O., Mohamed, A.r., Jiang, H., Deng, L., Penn, G., Yu, D., 2014 · 2014
Earlier work this paper cites.
Courbariaux, M., Bengio, Y., David, J.P., 2014 · 2014
Earlier work this paper cites.
Gong, Y., Liu, L., Yang, M., Bourdev, L., 2014 · 2014
Earlier work this paper cites.
Deep Speech: Scaling up end-to-end speech recognition
Hannun, A., Case, C., Casper, J., Catanzaro, B., Diamos, G., Elsen, E., Prenger, R., Satheesh, S., Sengupta, S., Coates, A., Ng, A.Y., 2014 · 2014
Earlier work this paper cites.
Labeled faces in the wild: Updates and new reporting procedures
Huang, G.B., Learned-miller, E., 2014 · 2014
Earlier work this paper cites.
Fixed-point feedforward deep neural network design using weights +1, 0, and -1, in: 2014 IEEE Workshop on Signal Processing Systems (SiPS), IEEE. pp. 1–6
Hwang, K., Sung, W., 2014 · 2014
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks, in: Advances in Neural Information Processing Systems (NIPS), pp. 1–9
Krizhevsky, A., Sutskever, I., Hinton, G.E., 2012 · 2014
Earlier work this paper cites.
Network in network, in: International Conference on Learning Representations(ICLR), pp. 1–10
Lin, M., Chen, Q., Yan, S., 2014 · 2014
Earlier work this paper cites.
NVIDIA GeForce GTX 980 Featuring Maxwell, The Most Advanced GPU Ever Made
NVIDIA Corporation, 2014 · 2014
Earlier work this paper cites.
Simonyan, K., Zisserman, A., 2014 · 2014
Earlier work this paper cites.
Expectation backpropagation: Parameter-free training of multilayer neural networks with continuous or discrete weights, in: Advances in Neural Information Processing Systems (NIPS), pp. 963–971
Soudry, D., Hubara, I., Meir, R., 2014 · 2014
Earlier work this paper cites.
Efficient Transfer Learning Method for Automatic Hyperparameter Tuning, in: Kaski, S., Corander, J. (Eds.), Proceedings of the Seventeenth International Conference on Artificial Intelligence and Statistics, PMLR, Reykjavik, Iceland. pp. 1077–1085
Yogatama, D., Mann, G., 2014 · 2014
Earlier work this paper cites.
ARM Architecture Reference Manual ARMv8, for ARMv8-A architecture profile
Arm, 2015 · 2015
Earlier work this paper cites.
Sparse Convolutional Neural Networks, in: 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), IEEE. pp. 806–814
Baoyuan Liu, Min Wang, Foroosh, H., Tappen, M., Penksy, M., 2015 · 2015
Earlier work this paper cites.
Conditional Computation in Neural Networks for faster models
Bengio, E., Bacon, P.L., Pineau, J., Precup, D., 2015 · 2015
Earlier work this paper cites.
Chen, W., Wilson, J., Tyree, S., Weinberger, K., Chen, Y., 2015 · 2015
Earlier work this paper cites.
Cheng, Z., Soudry, D., Mao, Z., Lan, Z., 2015 · 2015
Earlier work this paper cites.
Intel ® AVX-512 Instructions and Their Use in the Implementation of Math Functions
Cornea, M., 2015 · 2015
Earlier work this paper cites.
Courbariaux, M., Bengio, Y., David, J.P., 2015 · 2015
Earlier work this paper cites.
Dettmers, T., 2015 · 2015
Earlier work this paper cites.
HSA-enabled DSPs and accelerators
Glossner, J., Blinzer, P., Takala, J., 2016 · 2015
Earlier work this paper cites.
Deep learning with limited numerical precision, in: International Conference on Machine Learning (ICML), pp. 1737–1746
Gupta, S., Agrawal, A., Gopalakrishnan, K., Narayanan, P., 2015 · 2015
Earlier work this paper cites.
Han, S., Pool, J., Tran, J., Dally, W.J., 2015 · 2015
Earlier work this paper cites.
Ioffe, S., Szegedy, C., 2015 · 2015
Earlier work this paper cites.
Deep learning
Lecun, Y., Bengio, Y., Hinton, G., 2015 · 2015
Earlier work this paper cites.
Rounding Methods for Neural Networks with Low Resolution Synaptic Weights
Muller, L.K., Indiveri, G., 2015 · 2015
Earlier work this paper cites.
NVIDIA Tesla P100
NVIDIA Corporation, 2015 · 2015
Earlier work this paper cites.
Channel-level acceleration of deep face representations
Polyak, A., Wolf, L., 2015 · 2015
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M., Berg, A.C., Fei-Fei, L., 2015 · 2015
Earlier work this paper cites.
Srinivas, S., Babu, R.V., 2015 · 2015
Earlier work this paper cites.
Going deeper with convolutions, in: Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition, IEEE. pp. 1–9
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V., Rabinovich, A., 2015 · 2015
Cited alongside, same era.
TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G.S., Davis, A., Dean, J., Devin, M., Ghemawat, S., Goodfellow, I., Harp, A., Irving, G., Isard, M., Jia, Y., Jozefowicz, R., Kaiser, L., Kudlur, M., Levenberg, J., Mane, D., Monga, R., Moore, S., Murray, D., Olah, C., Schuster, M., Shlens, J., Steiner, B., Sutskever, I., Talwar, K., Tucker, P., Vanhoucke, V., Vasudevan, V., Viegas, F., Vinyals, O., Warden, P., Wattenberg, M., Wicke, M., Yu, Y., Zheng, X., 2016 · 2016
Cited alongside, same era.
DianNao family: Energy-Efficient Hardware Accelerators for Machine Learning
Chen, Y., Chen, T., Xu, Z., Sun, N., Temam, O., 2016 · 2016
Cited alongside, same era.
Courbariaux, M., Hubara, I., Soudry, D., El-Yaniv, R., Bengio, Y., 2016 · 2016
He, Y., Kang, G., Dong, X., Fu, Y., Yang, Y., 2018 · 2018
Later among the works it cites.
From hashing to CNNs: Training binary weight networks via hashing, in: AAAI Conference on Artificial Intelligence, pp. 3247–3254
Hu, Q., Wang, P., Cheng, J., 2018 · 2018
Later among the works it cites.
Multi-scale dense networks for resource efficient image classification, in: International Conference on Learning Representations(ICLR)
Huang, G., Chen, D., Li, T., Wu, F., Van Der Maaten, L., Weinberger, K., 2018 · 2018
Later among the works it cites.
Data-Driven Sparse Structure Selection for Deep Neural Networks, in: Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics). volume 11220 LNCS, pp. 317–334
Huang, Z., Wang, N., 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Convolutional networks for fast, energy-efficient neuromorphic computing
Esser, S.K., Merolla, P.A., Arthur, J.V., Cassidy, A.S., Appuswamy, R., Andreopoulos, A., Berg, D.J., McKinstry, J.L., Melano, T., Barch, D.R., di Nolfo, C., Datta, P., Amir, A., Taba, B., Flickner, M.D., Modha, D.S., 2016 · 2016
Cited alongside, same era.
Adaptive Computation Time for Recurrent Neural Networks
Graves, A., 2016 · 2016
Cited alongside, same era.
Greff, K., Srivastava, R.K., Schmidhuber, J., 2016 · 2016
Cited alongside, same era.
Dynamic Network Surgery for Efficient DNNs, in: Advances in Neural Information Processing Systems (NIPS), pp. 1379–1387
Guo, Y., Yao, A., Chen, Y., 2016 · 2016
Cited alongside, same era.
Han, S., Liu, X., Mao, H., Pu, J., Pedram, A., Horowitz, M.A., Dally, W.J., 2016a · 2016
Cited alongside, same era.
Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures
Hu, H., Peng, R., Tai, Y.W., Tang, C.K., 2016 · 2016
Cited alongside, same era.
Iandola, F.N., Han, S., Moskewicz, M.W., Ashraf, K., Dally, W.J., Keutzer, K., 2016 · 2016
Cited alongside, same era.
Lavin, A., Gray, S., 2016 · 2016
Cited alongside, same era.
Quantization and Training of Neural Networks for Efficient Integer-Arithmetic-Only Inference, in: IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), IEEE. pp. 2704–2713
Jacob, B., Kligys, S., Chen, B., Zhu, M., Tang, M., Howard, A., Adam, H., Kalenichenko, D., 2018 · 2018
Later among the works it cites.
Learning to Quantize Deep Networks by Optimizing Quantization Intervals with Task Loss
Jung, S., Son, C., Lee, S., Son, J., Kwak, Y., Han, J.J., Hwang, S.J., Choi, C., 2018 · 2018
Later among the works it cites.
CMSIS NN Software Library
Keil, 2018 · 2018
Later among the works it cites.
Quantizing deep convolutional networks for efficient inference: A whitepaper
Krishnamoorthi, R., 2018 · 2018
Later among the works it cites.
Speeding-up convolutional neural networks: A survey
Lebedev, V., Lempitsky, V., 2018 · 2018
Later among the works it cites.
Survey of Deep Neural Network Model Compression
Lei, J., Gao, X., Song, J., Wang, X.L., Song, M.L., 2018 · 2018
Later among the works it cites.
Extremely Low Bit Neural Network: Squeeze the Last Bit Out with ADMM
Leng, C., Li, H., Zhu, S., Jin, R., 2018 · 2018
Later among the works it cites.
Bi-Real Net: Enhancing the performance of 1-bit CNNs with improved representational capability and advanced training algorithm
Liu, Z., Wu, B., Luo, W., Yang, X., Liu, W., Cheng, K.T., 2018 · 2018
Later among the works it cites.
WRPN: Wide reduced-precision networks, in: International Conference on Learning Representations(ICLR), pp. 1–11
Mishra, A., Nurvitadhi, E., Cook, J.J., Marr, D., 2018 · 2018
Later among the works it cites.
NVIDIA Turing GPU Architecture
NVIDIA Corporation, 2018b · 2018
Later among the works it cites.
Compression of convolutional neural networks: A short survey, in: 2018 17th International Symposium on INFOTEH-JAHORINA, INFOTEH 2018 - Proceedings, IEEE. pp. 1–6
Pilipović, R., Bulić, P., Risojević, V., 2018 · 2018
Later among the works it cites.
Inference of quantized neural networks on heterogeneous all-programmable devices, in: 2018 Design, Automation & Test in Europe Conference & Exhibition (DATE), IEEE. pp. 833–838
Preuser, T.B., Gambardella, G., Fraser, N., Blott, M., 2018 · 2018
Later among the works it cites.
Lower Numerical Precision Deep Learning Inference and Training
Rodriguez, A., Segal, E., Meiri, E., Fomenko, E., Kim, Y.J., Shen, H., 2018 · 2018
Later among the works it cites.
Glow: Graph lowering compiler techniques for neural networks
Rotem, N., Fix, J., Abdulrasool, S., Catron, G., Deng, S., Dzhabarov, R., Gibson, N., Hegeman, J., Lele, M., Levenstein, R., Montgomery, J., Maher, B., Nadathur, S., Olesen, J., Park, J., Rakhov, A., Smelyanskiy, M., Wang, M., 2018 · 2018
Later among the works it cites.
How does batch normalization help optimization?, in: Advances in Neural Information Processing Systems (NIPS), pp. 2483–2493
Santurkar, S., Tsipras, D., Ilyas, A., Madry, A., 2018 · 2018
Later among the works it cites.
Quantizing Convolutional Neural Networks for Low-Power High-Throughput Inference Engines
Settle, S.O., Bollavaram, M., D’Alberto, P., Delaye, E., Fernandez, O., Fraser, N., Ng, A., Sirasao, A., Wu, M., 2018 · 2018
Later among the works it cites.
A Quantization-Friendly Separable Convolution for MobileNets
Sheng, T., Feng, C., Zhuo, S., Zhang, X., Shen, L., Aleksic, M., 2018 · 2018
Later among the works it cites.
Toolflows for Mapping Convolutional Neural Networks on FPGAs
Venieris, S.I., Kouris, A., Bouganis, C.S., 2018 · 2018
Later among the works it cites.
Two-Step Quantization for Low-bit Neural Networks, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 4376–4384
Wang, P., Hu, Q., Zhang, Y., Zhang, C., Liu, Y., Cheng, J., 2018b · 2018
Later among the works it cites.
L1-Norm Batch Normalization for Efficient Training of Deep Neural Networks
Wu, S., Li, G., Deng, L., Liu, L., Wu, D., Xie, Y., Shi, L., 2019 · 2018
Later among the works it cites.
BlockDrop: Dynamic Inference Paths in Residual Networks, in: IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), IEEE. pp. 8817–8826
Wu, Z., Nagarajan, T., Kumar, A., Rennie, S., Davis, L.S., Grauman, K., Feris, R., 2018b · 2018
Later among the works it cites.
Accelerating DNNs with Xilinx Alveo Accelerator Cards (WP504)
Xilinx, Inc, 2018 · 2018
Later among the works it cites.
A Low-Power Arithmetic Element for Multi-Base Logarithmic Computation on Deep Neural Networks, in: International System on Chip Conference, IEEE. pp. 260–265
Xu, J., Huan, Y., Zheng, L.R., Zou, Z., 2019 · 2018
Later among the works it cites.
Quantization of Fully Convolutional Networks for Accurate Biomedical Image Segmentation, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 8300–8308
Xu, X., Lu, Q., Yang, L., Hu, S., Chen, D., Hu, Y., Shi, Y., 2018a · 2018
Later among the works it cites.
Rethinking the Smaller-Norm-Less-Informative Assumption in Channel Pruning of Convolution Layers
Ye, J., Lu, X., Lin, Z., Wang, J.Z., 2018 · 2018
Later among the works it cites.
Yu, J., Yang, L., Xu, N., Yang, J., Huang, T., 2018 · 2018
Later among the works it cites.
LQ-Nets: Learned quantization for highly accurate and compact deep neural networks, in: Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics), pp. 373–390
Zhang, D., Yang, J., Ye, D., Hua, G., 2018 · 2018
Later among the works it cites.
Recent Advances in Convolutional Neural Network Acceleration
Zhang, Q., Zhang, M., Chen, T., Sun, Z., Ma, Y., Yu, B., 2019a · 2018
Later among the works it cites.
PArallel Distributed Deep LEarning: Machine Learning Framework from Industrial Practice
Baidu, 2019 · 2019
Later among the works it cites.
Post training 4-bit quantization of convolutional networks for rapid-deployment, in: Advances in Neural Information Processing Systems (NIPS), pp. 7950–7958
Banner, R., Nahshan, Y., Soudry, D., 2019 · 2019
Later among the works it cites.
BinaryDenseNet: Developing an architecture for binary neural networks
Bethge, J., Yang, H., Bornstein, M., Meinel, C., 2019 · 2019
Later among the works it cites.
BUG1989/caffe-int8-convert-tools: Generate a quantization parameter file for ncnn framework int8 inference
BUG1989, 2019 · 2019
Later among the works it cites.
SeerNet : Predicting Convolutional Neural Network Feature-Map Sparsity through Low-Bit Quantization
Cao, S., Ma, L., Xiao, W., Zhang, C., Liu, Y., Zhang, L., Nie, L., Yang, Z., 2019 · 2019
Later among the works it cites.
Accelerating Convolutional Neural Networks with Dynamic Channel Pruning, in: 2019 Data Compression Conference (DCC), IEEE. pp. 563–563
Chiliang, Z., Tao, H., Yingda, G., Zuochang, Y., 2019 · 2019
Later among the works it cites.
QNNPACK: Open source library for optimized mobile deep learning - Facebook Engineering
Dukhan, M., Yiming, W., Hao, L., Lu, H., 2019 · 2019
Later among the works it cites.
Neural Architecture Search
Elsken, T., Metzen, J.H., Hutter, F., 2019 · 2019
Later among the works it cites.
Frankle, J., Carbin, M., 2019 · 2019
Later among the works it cites.
Gao, X., Zhao, Y., Dudziak, L., Mullins, R., Xu, C.Z., Dudziak, L., Mullins, R., Xu, C.Z., 2019 · 2019
Later among the works it cites.
Differentiable soft quantization: Bridging full-precision and low-bit neural networks, in: Proceedings of the IEEE International Conference on Computer Vision (ICCV), pp. 4851–4860
Gong, R., Liu, X., Jiang, S., Li, T., Hu, P., Lin, J., Yu, F., Yan, J., 2019 · 2019
Later among the works it cites.
Filter Pruning via Geometric Median for Deep Convolutional Neural Networks Acceleration
He, Y., Liu, P., Wang, Z., Hu, Z., Yang, Y., 2019 · 2019
Later among the works it cites.
AI benchmark: All about deep learning on smartphones in 2019
Ignatov, A., Timofte, R., Kulik, A., Yang, S., Wang, K., Baum, F., Wu, M., Xu, L., Van Gool, L., 2019 · 2019
Later among the works it cites.
Dissecting the graphcore IPU architecture via microbenchmarking
Jia, Z., Tillman, B., Maggioni, M., Scarpazza, D.P., 2019 · 2019
Later among the works it cites.
SnIP: Single-shot network pruning based on connection sensitivity, in: International Conference on Learning Representations(ICLR)
Lee, N., Ajanthan, T., Torr, P.H., 2019 · 2019
Later among the works it cites.
Improved Techniques for Training Adaptive Deep Networks, in: 2019 IEEE/CVF International Conference on Computer Vision (ICCV), IEEE. pp. 1891–1900
Li, H., Zhang, H., Qi, X., Ruigang, Y., Huang, G., 2019 · 2019
Later among the works it cites.
Learning low-precision neural networks without Straight-Through Estimator (STE), in: IJCAI International Joint Conference on Artificial Intelligence, International Joint Conferences on Artificial Intelligence Organization, California. pp. 3066–3072
Liu, Z.G., Mattina, M., 2019 · 2019
Later among the works it cites.
Habana Labs presentation
Medina, E., 2019 · 2019
Later among the works it cites.
PyTorch : An Imperative Style , High-Performance Deep Learning Library
Paszke, A., Gross, S., Bradbury, J., Lin, Z., Devito, Z., Massa, F., Steiner, B., Killeen, T., Yang, E., 2019 · 2019
Later among the works it cites.
Survey and Benchmarking of Machine Learning Accelerators, in: 2019 IEEE High Performance Extreme Computing Conference (HPEC), IEEE. pp. 1–9
Reuther, A., Michaleas, P., Jones, M., Gadepally, V., Samsi, S., Kepner, J., 2019 · 2019
Later among the works it cites.
Searching for accurate binary neural architectures
Shen, M., Han, K., Xu, C., Wang, Y., 2019 · 2019
Later among the works it cites.
A review of binarized neural networks
Simons, T., Lee, D.J., 2019 · 2019
Later among the works it cites.
Play and Prune: Adaptive Filter Pruning for Deep Model Compression, in: Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, International Joint Conferences on Artificial Intelligence Organization, California. pp. 3460–3466
Singh, P., Kumar Verma, V., Rai, P., Namboodiri, V.P., 2019 · 2019
Later among the works it cites.
Snapdragon Neural Processing Engine SDK
Technologies, Q., 2019 · 2019
Later among the works it cites.
NCNN is a high-performance neural network inference framework optimized for the mobile platform
Tencent, 2019 · 2019
Later among the works it cites.
Wang, K., Liu, Z., Lin, Y., Lin, J., Han, S., 2019a · 2019
Later among the works it cites.
Learning channel-wise interactions for binary convolutional neural networks, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 568–577
Wang, Z., Lu, J., Tao, C., Zhou, J., Tian, Q., 2019b · 2019
Later among the works it cites.
MACE is a deep learning inference framework optimized for mobile heterogeneous computing platforms
Xiaomi, 2019 · 2019
Later among the works it cites.
Quantization Networks, in: 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), IEEE. pp. 7300–7308
Yang, J., Shen, X., Xing, J., Tian, X., Li, H., Deng, B., Huang, J., Hua, X.s., 2019 · 2019
Later among the works it cites.
Training high-performance and large-scale deep neural networks with full 8-bit integers
Yang, Y., Deng, L., Wu, S., Yan, T., Xie, Y., Li, G., 2020 · 2019
Later among the works it cites.
Blended coarse gradient descent for full quantization of deep neural networks
Yin, P., Zhang, S., Lyu, J., Osher, S., Qi, Y., Xin, J., 2019 · 2019
Later among the works it cites.
Structured binary neural networks for accurate image classification and semantic segmentation
Zhuang, B., Shen, C., Tan, M., Liu, L., Reid, I., 2019 · 2019
Later among the works it cites.
FPGAs Enable the Next Generation of Communication and Networking Solutions
Achronix Semiconductor Corporation, 2020 · 2020
Later among the works it cites.
convnet-burden
Albanie, 2020 · 2020
Later among the works it cites.
Arm Cortex-M Processor Comparison Table
Arm, 2020 · 2020
Later among the works it cites.
MALI-G76 High-Performance GPU for Complex Graphics Features and Bene ts High Performance for Mixed Realities
Arm, Graphics, C., 2020 · 2020
Later among the works it cites.
A comprehensive survey on model compression and acceleration
Choudhary, T., Mishra, V., Goswami, A., Sarangapani, J., 2020 · 2020
Later among the works it cites.
Hanguang 800 NPU – The Ultimate AI Inference Solution for Data Centers, in: 2020 IEEE Hot Chips 32 Symposium (HCS), IEEE. pp. 1–29
Jiao, Y., Han, L., Long, X., 2020 · 2020
Later among the works it cites.
Xilinx Vitis Unified Software Platform, in: Proceedings of the 2020 ACM/SIGDA International Symposium on Field-Programmable Gate Arrays, ACM, New York, NY, USA. pp. 173–174
Kathail, V., 2020 · 2020
Later among the works it cites.
Group Sparsity: The Hinge Between Filter Pruning and Decomposition for Network Compression, in: 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), IEEE. pp. 8015–8024
Li, Y., Gu, S., Mayer, C., Van Gool, L., Timofte, R., 2020b · 2020
Later among the works it cites.
AutoPruner: An end-to-end trainable filter pruning method for efficient deep model inference
Luo, J.H., Wu, J., 2020 · 2020
Later among the works it cites.
Heterogeneous Edge CNN Hardware Accelerator, in: The 12th International Conference on Wireless Communications and Signal Processing, pp. 6–11
Moudgill, M., Glossner, J., Huang, W., Tian, C., Xu, C., Yang, N., Wang, L., Liang, T., Shi, S., Zhang, X., Iancu, D., Nacer, G., Li, K., 2020 · 2020
Later among the works it cites.
Baidu Kunlun An AI processor for diversified workloads, in: 2020 IEEE Hot Chips 32 Symposium (HCS), IEEE. pp. 1–18
Ouyang, J., Noh, M., Wang, Y., Qi, W., Ma, Y., Gu, C., Kim, S., Hong, K.i., Bae, W.K., Zhao, Z., Wang, J., Wu, P., Gong, X., Shi, J., Zhu, H., Du, X., 2020 · 2020
Later among the works it cites.
Binary neural networks: A survey
Qin, H., Gong, R., Liu, X., Bai, X., Song, J., Sebe, N., 2020a · 2020
Later among the works it cites.
Forward and Backward Information Retention for Accurate Binary Neural Networks, in: IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), IEEE. pp. 2247–2256
Qin, H., Gong, R., Liu, X., Shen, M., Wei, Z., Yu, F., Song, J., 2020b · 2020
Later among the works it cites.
Introducing the Intel® Vision Accelerator Design with Intel® Arria® 10 FPGA
Richard Chuang, Oliyide, O., Garrett, B., 2020 · 2020
Later among the works it cites.
Categorizing Malware via A Word2Vec-based Temporal Convolutional Network Scheme
Sun, J., Luo, X., Gao, H., Wang, W., Gao, Y., Yang, X., 2020 · 2020
Later among the works it cites.
Integer quantization for deep learning inference: Principles and empirical evaluation
Wu, H., Judd, P., Zhang, X., Isaev, M., Micikevicius, P., 2020 · 2020
Later among the works it cites.
Convolutional Neural Network Pruning: A Survey, in: 2020 39th Chinese Control Conference (CCC), IEEE. pp. 7458–7463
Xu, S., Huang, A., Chen, L., Zhang, B., 2020 · 2020
Later among the works it cites.
A dual-attention recurrent neural network method for deep cone thickener underflow concentration prediction
Yuan, Z., Hu, J., Wu, D., Ban, X., 2020 · 2020
Later among the works it cites.