Fetching the paper…
Reading the bibliography…
The computational demands of computer vision tasks based on state-of-the-art Convolutional Neural Network (CNN) image classification far exceed the energy budgets of mobile devices.
A Signed Binary Multiplication Technique
Booth, A · 1951
Earlier work this paper cites.
Static scheduling of synchronous data flow programs for digital signal processing
Lee, E. A. and Messerschmitt, D. G · 1987
Earlier work this paper cites.
Mapping multirate dataflow to complex rt level hardware models
Horstmannshoff, J., Grotker, T., and Meyr, H · 1997
Earlier work this paper cites.
Operator strength reduction
Cooper, K. D., Simpson, L. T., and Vick, C. A · 2001
Earlier work this paper cites.
Robust Real-time Object Detection
Viola, P. and Jones, M. J · 2004
Earlier work this paper cites.
Histograms of Oriented Gradients for Human Detection
Dalal, N. and Triggs, B · 2005
Earlier work this paper cites.
Automated flower classification over a large number of classes
Nilsback, M.-E. and Zisserman, A · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A. and Hinton, G · 2009
Earlier work this paper cites.
Datapath synthesis for standard-cell design
Zimmermann, R · 2009
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Netzer, Y., Wang, T., Coates, A., Bissacco, A., Wu, B., and Ng, A. Y · 2011
Earlier work this paper cites.
Man vs. computer: Benchmarking machine learning algorithms for traffic sign recognition
Stallkamp, J., Schlipsing, M., Salmen, J., and Igel, C · 2012
Earlier work this paper cites.
Fine-grained visual classification of aircraft
Maji, S., Rahtu, E., Kannala, J., Blaschko, M., and Vedaldi, A · 2013
Earlier work this paper cites.
Halide: A Language and Compiler for Optimizing Parallelism, Locality, and Recomputation in Image Processing Pipelines
Ragan-Kelley, J., Barnes, C., Adams, A., Paris, S., Durand, F., and Amarasinghe, S · 2013
Earlier work this paper cites.
Darkroom: Compiling High-Level Image Processing Code into Hardware Pipelines
Hegarty, J., Brunhaver, J., DeVito, Z., Ragan-Kelley, J., Cohen, N., Bell, S., Vasilyev, A., Horowitz, M., and Hanrahan, P · 2014
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition
Simonyan, K. and Zisserman, A · 2014
Earlier work this paper cites.
How transferable are features in deep neural networks?
Yosinski, J., Clune, J., Bengio, Y., and Lipson, H · 2014
Earlier work this paper cites.
Always-on Vision Processing Unit for Mobile Applications
Barry, B., Brick, C., Connor, F., Donohoe, D., Moloney, D., Richmond, R., O’Riordan, M., and Toma, V · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S. and Szegedy, C · 2015
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M., et al · 2015
Cited alongside, same era.
Simultaneous deep transfer across domains and tasks
Tzeng, E., Hoffman, J., Darrell, T., and Saenko, K · 2015
Cited alongside, same era.
Why GEMM is at the heard of deep learning, April 2015
Warden, P · 2015
Cited alongside, same era.
Cnvlutin: Ineffectual-neuron-free Deep Neural Network Computing
Albericio, J., Judd, P., Hetherington, T., Aamodt, T., Jerger, N. E., and Moshovos, A · 2016
Cited alongside, same era.
PRIME: A Novel Processing-in-Memory Architecture for Neural Network Computation in ReRAM-Based Main Memory
Chi, P., Li, S., Xu, C., Zhang, T., Zhao, J., Liu, Y., Wang, Y., and Xie, Y · 2016
Cited alongside, same era.
EIE: Efficient Inference Engine on Compressed Deep Neural Network
In-Datacenter Performance Analysis of a Tensor Processing Unit
Jouppi, N. P., Young, C., Patil, N., Patterson, D., Agrawal, G., Bajwa, R., Bates, S., Bhatia, S., Boden, N., Borchers, A., Boyle, R., Cantin, P., Chao, C., Clark, C., Coriell, J., Daley, M., Dau, M., Dean, J., Gelb, B., Ghaemmaghami, T. V., Gottipati, R., Gulland, W., Hagmann, R., Ho, R. C., Hogberg, D., Hu, J., Hundt, R., Hurt, D., Ibarz, J., Jaffey, A., Jaworski, A., Kaplan, A., Khaitan, H., Koch, A., Kumar, N., Lacy, S., Laudon, J., Law, J., Le, D., Leary, C., Liu, Z., Lucke, K., Lundin, A., MacKean, G., Maggiore, A., Mahony, M., Miller, K., Nagarajan, R., Narayanaswami, R., Ni, R., Nix, K., Norrie, T., Omernick, M., Penukonda, N., Phelps, A., Ross, J., Salek, A., Samadiani, E., Severn, C., Sizikov, G., Snelham, M., Souter, J., Steinberg, D., Swing, A., Tan, M., Thorson, G., Tian, B., Toma, H., Tuttle, E., Vasudevan, V., Walter, R., Wang, W., Wilcox, E., and Yoon, D. H · 2017
Later among the works it cites.
Applications of deep neural networks for ultra low power iot
Kodali, S., Hansen, P., Mulholland, N., Whatmough, P., Brooks, D., and Wei, G · 2017
Later among the works it cites.
SCNN: An Accelerator for Compressed-sparse Convolutional Neural Networks
Parashar, A., Rhu, M., Mukkara, A., Puglielli, A., Venkatesan, R., Khailany, B., Emer, J., Keckler, S. W., and Dally, W. J · 2017
Later among the works it cites.
A case for efficient accelerator design space exploration via bayesian optimization
Reagen, B., Hernández-Lobato, J. M., Adolf, R., Gelbart, M., Whatmough, P., Wei, G., and Brooks, D · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Han, S., Liu, X., Mao, H., Pu, J., Pedram, A., Horowitz, M., and Dally, W · 2016
Cited alongside, same era.
Rigel: Flexible Multi-Rate Image Processing Hardware
Hegarty, J., Daly, R., DeVito, Z., Ragan-Kelley, J., Horowitz, M., and Hanrahan, P · 2016
Cited alongside, same era.
Designing neural network hardware accelerators with decoupled objective evaluations
Hernández-Lobato, J. M., Gelbart, M. A., Reagen, B., Adolf, R., Hernández-Lobato, D., Whatmough, P. N., Brooks, D., Wei, G.-Y., and Adams, R. P · 2016
Cited alongside, same era.
Stripes: Bit-serial Deep Neural Network Computing
Judd, P., Albericio, J., Hetherington, T., Aamodt, T. M., and Moshovos, A · 2016
Cited alongside, same era.
Neurocube: A Programmable Digital Neuromorphic Architecture with High-Density 3D Memory
Kim, D., Kung, J., Chai, S., Yalamanchili, S., and Mukhopadhyay, S · 2016
Cited alongside, same era.
RedEye: Analog ConvNet Image Sensor Architecture for Continuous Mobile Vision
LiKamWa, R., Hou, Y., Gao, J., Polansky, M., and Zhong, L · 2016
Cited alongside, same era.
TABLA: A Unified Template-based Framework for Accelerating Statistical Machine Learning
Mahajan, D., Park, J., Amaro, E., Sharma, H., Yazdanbakhsh, A., Kim, J. K., and Esmaeilzadeh, H · 2016
Cited alongside, same era.
Later among the works it cites.
Learning multiple visual domains with residual adapters
Rebuffi, S., Bilen, H., and Vedaldi, A · 2017
Later among the works it cites.
Pipelayer: A pipelined reram-based accelerator for deep learning
Song, L., Qian, X., Li, H., and Chen, Y · 2017
Later among the works it cites.
Towards Closing the Energy Gap Between HOG and CNN Features for Embedded Vision
Suleiman, A., Chen, Y.-H., Emer, J., and Sze, V · 2017
Later among the works it cites.
Finn: A framework for fast, scalable binarized neural network inference
Umuroglu, Y., Fraser, N. J., Gambardella, G., Blott, M., Leong, P., Jahre, M., and Vissers, K · 2017
Later among the works it cites.
Scalpel: Customizing DNN Pruning to the Underlying Hardware Parallelism
Yu, J., Lukefahr, A., Palframan, D., Dasika, G., Das, R., and Mahlke, S · 2017
Later among the works it cites.
To prune, or not to prune: exploring the efficacy of pruning for model compression
Zhu, M. and Gupta, S · 2017
Later among the works it cites.
Eva2: Exploiting temporal redundancy in live computer vision
Buckler, M., Bedoukian, P., Jayasuriya, S., and Sampson, A · 2018
Later among the works it cites.
Fix your classifier: the marginal value of training the last weight layer
Hoffer, E., Hubara, I., and Soudry, D · 2018
Later among the works it cites.
Efficient parametrization of multi-domain deep neural networks
Rebuffi, S.-A., Bilen, H., and Vedaldi, A · 2018
Later among the works it cites.
Computation reuse in dnns by exploiting input similarity
Riera, M., Arnau, J., and Gonzalez, A · 2018
Later among the works it cites.
Scale-sim: Systolic CNN accelerator
Samajdar, A., Zhu, Y., Whatmough, P. N., Mattina, M., and Krishna, T · 2018
Later among the works it cites.
Toolflows for mapping convolutional neural networks on fpgas: A survey and future directions
Venieris, S. I., Kouris, A., and Bouganis, C.-S · 2018
Later among the works it cites.
Dnn engine: A 28-nm timing-error tolerant sparse deep neural network processor for iot applications
Whatmough, P. N., Lee, S. K., Brooks, D., and Wei, G · 2018
Later among the works it cites.
Euphrates: Algorithm-soc co-design for low-power mobile continuous vision
Zhu, Y., Samajdar, A., Mattina, M., and Whatmough, P · 2018
Later among the works it cites.