Fetching the paper…
Reading the bibliography…
Emerging edge computing platforms often contain machine learning (ML) accelerators that can accelerate inference for a wide range of neural network (NN) models.
K. Fukushima, “Neocognitron: A Self-Organizing Neural Network Model for a Mechanism of Pattern Recognition Unaffected by Shift in Position,” Biological Cybernetics , 1980
1980
Earlier work this paper cites.
Y. LeCun, “Une Procédure d’Apprentissage pour Réseau à Seuil Asymétrique,” in Cognitiva , 1985
1985
Earlier work this paper cites.
D. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning Representations by Back-Propagating Errors,” Nature , 1986
1986
Earlier work this paper cites.
Y. LeCun, B. Boser, J. Denker, D. Henderson, R. Howard, W. Hubbard, and L. Jackel, “Handwritten Digit Recognition with a Back-Propagation Network,” in NeurIPS , 1989
1989
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long Short-Term Memory,” NECO , 1997
1997
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-Based Learning Applied to Document Recognition,” Proc. IEEE , 1998
1998
Earlier work this paper cites.
F. Gers, J. Schmidhuber, and F. Cummins, “Learning to Forget: Continual Prediction with LSTM,” in ICANN , 1999
1999
Earlier work this paper cites.
N. Muralimanohar, R. Balasubramonian, and N. Jouppi, “Optimizing NUCA Organizations and Wiring Alternatives for Large Caches with CACTI 6.0,” in MICRO , 2007
2007
Earlier work this paper cites.
2012
Earlier work this paper cites.
J. W. Choi, D. Bedard, R. Fowler, and R. Vuduc, “A Roofline Model of Energy,” in IPDPS , 2013
2013
Earlier work this paper cites.
A. Graves, “Generating Sequences with Recurrent Neural Networks,” arXiv:1308.0850 [cs.NE], 2013
2013
Earlier work this paper cites.
A. Graves, N. Jaitly, and A.-r. Mohamed, “Hybrid Speech Recognition with Deep Bidirectional LSTM,” in ASRU , 2013
2013
Earlier work this paper cites.
T. Chen, Z. Du, N. Sun, J. Wang, C. Wu, Y. Chen, and O. Temam, “DianNao: A Small-Footprint High-Throughput Accelerator for Ubiquitous Machine-Learning,” in ASPLOS , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. Chung, C. Gulcehre, K. Cho, and Y. Bengio, “Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling,” in NeurIPS , 2014
2014
Earlier work this paper cites.
Hybrid Memory Cube Consortium, “HMC Specification 2.0,” 2014
2014
Earlier work this paper cites.
P. Pinheiro and R. Collobert, “Recurrent Convolutional Neural Networks for Scene Labeling,” in ICML , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to Sequence Learning with Neural Networks,” in NeurIPS , 2014
2014
Earlier work this paper cites.
M. D. Zeiler and R. Fergus, “Visualizing and Understanding Convolutional Networks,” in ECCV , 2014
2014
Earlier work this paper cites.
J. Donahue, L. A. Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell, “Long-Term Recurrent Convolutional Networks for Visual Recognition and Description,” in CVPR , 2015
2015
Earlier work this paper cites.
Z. Du, R. Fasthuber, T. Chen, P. Ienne, L. Li, T. Luo, X. Feng, Y. Chen, and O. Temam, “ShiDianNao: Shifting Vision Processing Closer to the Sensor,” in ISCA , 2015
2015
Earlier work this paper cites.
M. H. Ionica and D. Gregg, “The Movidius Myriad Architecture’s Potential for Scientific Computing,” IEEE Micro , 2015
2015
Earlier work this paper cites.
A. Karpathy and L. Fei-Fei, “Deep Visual-Semantic Alignments for Generating Image Descriptions,” in CVPR , 2015
2015
Earlier work this paper cites.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep Learning,” Nature , 2015
2015
Earlier work this paper cites.
H. Li, Z. Lin, X. Shen, J. Brandt, and G. Hua, “A Convolutional Neural Network Cascade for Face Detection,” in CVPR , 2015
2015
Earlier work this paper cites.
M. Liang and X. Hu, “Recurrent Convolutional Neural Network for Object Recognition,” in CVPR , 2015
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein et al. , “ImageNet Large Scale Visual Recognition Challenge,” IJCV , 2015
2015
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very Deep Convolutional Networks for Large-Scale Image Recognition,” in ICLR , 2015
2015
Earlier work this paper cites.
N. Srivastava, E. Mansimov, and R. Salakhudinov, “Unsupervised Learning of Video Representations Using LSTMs,” in ICML , 2015
2015
Earlier work this paper cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going Deeper with Convolutions,” in CVPR , 2015
2015
Earlier work this paper cites.
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan, “Show and Tell: A Neural Image Caption Generator,” in CVPR , 2015
2015
Cited alongside, same era.
S. Xingjian, Z. Chen, H. Wang, D.-Y. Yeung, W.-K. Wong, and W.-c. Woo, “Convolutional LSTM Network: A Machine Learning Approach for Precipitation Nowcasting,” in NeurIPS , 2015
2015
Cited alongside, same era.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio, “Show, Attend and Tell: Neural Image Caption Generation with Visual Attention,” in ICML , 2015
2015
Cited alongside, same era.
J. Albericio, P. Judd, T. Hetherington, T. Aamodt, N. Enright Jerger, and A. Moshovos, “Cnvlutin: Ineffectual-Neuron-Free Deep Neural Network Computing,” in ISCA , 2016
2016
Cited alongside, same era.
M. Alwani, H. Chen, M. Ferdman, and P. Milder, “Fused-Layer CNN Accelerators,” in MICRO , 2016
A. Boroumand, S. Ghose, Y. Kim, R. Ausavarungnirun, E. Shiu, R. Thakur, D. Kim, A. Kuusela, A. Knies, P. Ranganathan, and O. Mutlu, “Google Workloads for Consumer Devices: Mitigating Data Movement Bottlenecks,” in ASPLOS , 2018
2018
Later among the works it cites.
J. Fowers, K. Ovtcharov, M. Papamichael, T. Massengill, M. Liu, D. Lo, S. Alkalay, M. Haselman, L. Adams, M. Ghandi, S. Heil, P. Patel, A. Sapek, G. Weisz, L. Woods, S. Lanka, S. K. Reinhardt, A. M. Caulfield, E. S. Chung, and D. Burger, “A Configurable Cloud-Scale DNN Processor for Real-Time AI,” in ISCA , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
J. Gu, Z. Wang, J. Kuen, L. Ma, A. Shahroudy, B. Shuai, T. Liu, X. Wang, G. Wang, J. Cai, and T. Chen, “Recent Advances in Convolutional Neural Networks,” Pattern Recognition , 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep Residual Learning for Image Recognition,” in CVPR , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
A. Kannan, K. Kurach, S. Ravi, T. Kaufmann, A. Tomkins, B. Miklo, G. Corrado, L. Lukács, M. Ganea, P. Young, and V. Ramavajjala, “Smart Reply: Automated Response Suggestion for Email,” in KDD , 2016
2016
Cited alongside, same era.
H. Qin, J. Yan, X. Li, and X. Hu, “Joint Training of Cascaded CNN for Face Detection,” in CVPR , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
A. Boroumand, S. Ghose, M. Patel, H. Hassan, B. Lucia, K. Hsieh, K. T. Malladi, H. Zheng, and O. Mutlu, “LazyPIM: An Efficient Cache Coherence Mechanism for Processing-in-Memory,” IEEE CAL , 2017
2017
Cited alongside, same era.
2018
Later among the works it cites.
K. Hsieh, G. Ananthanarayanan, P. Bodik, S. Venkataraman, P. Bahl, M. Philipose, P. B. Gibbons, and O. Mutlu, “Focus: Querying Large Video Datasets with Low Latency and Low Cost,” in OSDI , 2018
2018
Later among the works it cites.
JEDEC Solid State Technology Assn., “JESD235B: High Bandwidth Memory (HBM) DRAM,” December 2018
2018
Later among the works it cites.
H. Kwon, A. Samajdar, and T. Krishna, “MAERI: Enabling Flexible Dataflow Mapping over DNN Accelerators via Reconfigurable Interconnects,” in ASPLOS , 2018
2018
Later among the works it cites.
N. Ma, X. Zhang, H.-T. Zheng, and J. Sun, “ShuffleNet V2: Practical Guidelines for Efficient CNN Architecture Design,” in ECCV , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
F. Silfa, G. Dot, J.-M. Arnau, and A. González, “E-PUR: An Energy-Efficient Processing Unit for Recurrent Neural Networks,” in PACT , 2018
2018
Later among the works it cites.
F. Tu, W. Wu, S. Yin, L. Liu, and S. Wei, “RANA: Towards Efficient Neural Acceleration with Refresh-Optimized Embedded DRAM,” in ISCA , 2018
2018
Later among the works it cites.
X. Xu, Y. Ding, S. X. Hu, M. T. Niemier, J. Cong, Y. Hu, and Y. Shi, “Scaling for Edge Inference of Deep Neural Networks,” Nature Electronics , 2018
2018
Later among the works it cites.
X. Zhang, X. Zhou, M. Lin, and J. Sun, “ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices,” in CVPR , 2018
2018
Later among the works it cites.
A. Boroumand, S. Ghose, M. Patel, H. Hassan, B. Lucia, R. Ausavarungnirun, K. Hsieh, N. Hajinazar, K. T. Malladi, H. Zheng, and O. Mutlu, “CoNDA: Efficient Cache Coherence Support for Near-data Accelerators,” in ISCA , 2019
2019
Later among the works it cites.
Y.-H. Chen, T. Yang, J. S. Emer, and V. Sze, “Eyeriss v2: A Flexible Accelerator for Emerging Deep Neural Networks on Mobile Devices,” JETCAS , 2019
2019
Later among the works it cites.
M. Gao, X. Yang, J. Pu, M. Horowitz, and C. Kozyrakis, “TANGRAM: Optimized Coarse-Grained Dataflow for Scalable NN Accelerators,” in ASPLOS , 2019
2019
Later among the works it cites.
S. Ghose, K. Hsieh, A. Boroumand, R. Ausavarungnirun, and O. Mutlu, “The Processing-in-Memory Paradigm: Mechanisms to Enable Adoption,” in Beyond-CMOS Technologies for Next Generation Computer Design , 2019
2019
Later among the works it cites.
S. Gudaparthi, S. Narayanan, R. Balasubramonian, E. Giacomin, H. Kambalasubramanyam, and P.-E. Gaillardon, “Wire-Aware Architecture and Dataflow for CNN Accelerators,” in MICRO , 2019
2019
Later among the works it cites.
U. Gupta, B. Reagen, L. Pentecost, M. Donato, T. Tambe, A. M. Rush, G. Wei, and D. Brooks, “MASR: A Modular Accelerator for Sparse RNNs,” in PACT , 2019
2019
Later among the works it cites.
Y. He, T. N. Sainath, R. Prabhavalkar, I. McGraw, R. Alvarez, D. Zhao, D. Rybach, A. Kannan, Y. Wu, R. Pang, Q. Liang, D. Bhatia, Y. Shangguan, B. Li, G. Pundak, K. C. Sim, T. Bagby, S. yiin Chang, K. Rao, and A. Gruenstein, “Streaming End-to-End Speech Recognition for Mobile Devices,” in ICASSP , 2019
2019
Later among the works it cites.
A. Khamparia, B. Pandey, S. Tiwari, D. Gupta, A. Khanna, and J. Rodrigues, “An Integrated Hybrid CNN–RNN Model for Visual Description and Generation of Captions,” CSSP , 2019
2019
Later among the works it cites.
H. Kwon, P. Chatarasi, M. Pellauer, A. Parashar, V. Sarkar, and T. Krishna, “Understanding Reuse, Performance, and Hardware Cost of DNN Dataflow: A Data-Centric Approach,” in MICRO , 2019
2019
Later among the works it cites.
J. Li, R. Zhao, H. Hu, and Y. Gong, “Improving RNN Transducer Modeling for End-to-End Speech Recognition,” in ASRU , 2019
2019
Later among the works it cites.
O. Mutlu, S. Ghose, J. Gomez-Luna, and R. Ausavarungnirun, “Processing Data Where It Makes Sense: Enabling In-Memory Computation,” MICPRO , 2019
2019
Later among the works it cites.
C. Rui, X. Wang, W. Zhang, X. Zhu, A. Li, and C. Yang, “A Hybrid CNN–LSTM Model for Typhoon Formation Forecasting,” GeoInformatica , 2019
2019
Later among the works it cites.
Y. S. Shao, J. Clemons, R. Venkatesan, B. Zimmer, M. Fojtik, N. Jiang, B. Keller, A. Klinefelter, N. Pinckney, P. Raina, S. G. Tell, Y. Zhang, W. J. Dally, J. Emer, C. T. Gray, B. Khailany, and S. W. Keckler, “Simba: Scaling Deep-Learning Inference with Multi-Chip-Module-Based Architecture,” in MICRO , 2019
2019
Later among the works it cites.
C. Wu, D. Brooks, K. Chen, D. Chen, S. Choudhury, M. Dukhan, K. Hazelwood, E. Isaac, Y. Jia, B. Jia, T. Leyvand, H. Lu, Y. Lu, L. Qiao, B. Reagen, J. Spisak, F. Sun, A. Tulloch, P. Vajda, X. Wang, Y. Wang, B. Wasti, Y. Wu, R. Xian, S. Yoo, and P. Zhang, “Machine Learning at Facebook: Understanding Inference at the Edge,” in HPCA , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
P. Gibson, J. Cano, J. Turner, E. J. Crowley, M. O’Boyle, and A. Storkey, “Optimizing Grouped Convolutions on Edge Devices,” in ASAP , 2020
2020
Later among the works it cites.
JEDEC Solid State Technology Assn., “JESD209-4C: Low Power Double Data Rate 4 (LPDDR4) Standard,” January 2020
2020
Later among the works it cites.
L. Ke, U. Gupta, B. Y. Cho, D. Brooks, V. Chandra, U. Diril, A. Firoozshahian, K. Hazelwood, B. Jia, H. S. Lee, M. Li, B. Maher, D. Mudigere, M. Naumov, M. Schatz, M. Smelyanskiy, X. Wang, B. Reagen, C. Wu, M. Hempstead, and X. Zhang, “RecNMP: Accelerating Personalized Recommendation with Near-Memory Processing,” in ISCA , 2020
2020
Later among the works it cites.
H. Kwon, L. Lai, M. Pellauer, T. Krishna, Y.-H. Chen, and V. Chandra, “Heterogeneous Dataflow Accelerators for Multi-DNN Workloads,” in HPCA , 2021
2021
Closest in time.