Fetching the paper…
Reading the bibliography…
DNNs are becoming less and less over-parametrised due to recent advances in efficient model design, through careful hand-crafted or NAS-based methods.
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015 · 2015
Earlier work this paper cites.
Distilling the Knowledge in a Neural Network. In
G. Hinton et al · 2015
Earlier work this paper cites.
Going deeper with convolutions. In
Christian Szegedy et al · 2015
Earlier work this paper cites.
Dropout as a bayesian approximation: Representing model uncertainty in deep learning. In
Yarin Gal and Zoubin Ghahramani. 2016 · 2016
Earlier work this paper cites.
Deep Compression: Compressing Deep Neural Network with Pruning, Trained Quantization and Huffman Coding
Song Han et al · 2016
Earlier work this paper cites.
Fengfu Li and Bin Liu. 2016 · 2016
Earlier work this paper cites.
Conditional deep learning for energy-efficient and enhanced pattern recognition. In
Priyadarshini Panda, Abhronil Sengupta, and Kaushik Roy. 2016 · 2016
Earlier work this paper cites.
Xnor-net: Imagenet classification using binary convolutional neural networks. In
Mohammad Rastegari et al · 2016
Earlier work this paper cites.
BranchyNet: Fast Inference via Early Exiting from Deep Neural Networks. In
Surat Teerapittayanon, Bradley McDanel, and HT Kung. 2016 · 2016
Earlier work this paper cites.
Do convolutional neural networks learn class hierarchy?
Alsallakh Bilal et al · 2017
Earlier work this paper cites.
On Calibration of Modern Neural Networks. In
Chuan Guo et al · 2017
Earlier work this paper cites.
MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications
Andrew G. Howard et al · 2017
Earlier work this paper cites.
Neurosurgeon: Collaborative Intelligence Between the Cloud and Mobile Edge. In
Yiping Kang et al · 2017
Earlier work this paper cites.
Not All Pixels Are Equal: Difficulty-Aware Semantic Segmentation via Deep Layer Cascade. In
Xiaoxiao Li et al · 2017
Earlier work this paper cites.
Runtime Neural Pruning. In
Ji Lin et al · 2017
Earlier work this paper cites.
Mobile Sensing at the Service of Mental Well-Being: A Large-Scale Longitudinal Study. WWW
Sandra Servia-Rodríguez et al · 2017
Earlier work this paper cites.
Attention is All you Need. In
Ashish Vaswani et al · 2017
Earlier work this paper cites.
Idk cascades: Fast deep learning by learning not to overthink
Xin Wang et al · 2017
Earlier work this paper cites.
Feedback networks. In
Amir R Zamir et al · 2017
Earlier work this paper cites.
NestDNN: Resource-Aware Multi-Tenant On-Device Deep Learning for Continuous Mobile Vision. In
Biyi Fang et al · 2018
Earlier work this paper cites.
Multi-Scale Dense Networks for Resource Efficient Image Classification. In
Gao Huang et al · 2018
Earlier work this paper cites.
To Trust Or Not To Trust A Classifier
Heinrich Jiang, Been Kim, Melody Guan, and Maya Gupta. 2018 · 2018
Earlier work this paper cites.
CascadeCNN: Pushing the Performance Limits of Quantisation in Convolutional Neural Networks. In
Alexandros Kouris et al · 2018
Earlier work this paper cites.
Dynamic Deep Neural Networks: Optimizing accuracy-efficiency trade-offs by selective execution. In
Lanlan Liu and Jia Deng. 2018 · 2018
Earlier work this paper cites.
Hydranets: Specialized dynamic architectures for efficient inference. In
Ravi Teja Mullapudi et al · 2018
Earlier work this paper cites.
Adaptive deep learning model selection on embedded systems
Ben Taylor et al · 2018
Earlier work this paper cites.
Convolutional networks with adaptive inference graphs. In
Andreas Veit and Serge Belongie. 2018 · 2018
Earlier work this paper cites.
Skipnet: Learning dynamic routing in convolutional networks. In
Xin Wang, Fisher Yu, Zi-Yi Dou, Trevor Darrell, and Joseph E Gonzalez. 2018 · 2018
Cited alongside, same era.
Blockdrop: Dynamic inference paths in residual networks. In
Zuxuan Wu et al · 2018
Cited alongside, same era.
Netadapt: Platform-aware neural network adaptation for mobile applications. In
Tien-Ju Yang et al · 2018
Cited alongside, same era.
EmBench: Quantifying Performance Variations of Deep Neural Networks Across Modern Commodity Devices. In
Mario Almeida et al · 2019
Cited alongside, same era.
Dynamically sacrificing accuracy for reduced computation: Cascaded inference based on softmax confidence. In
Konstantin Berestizshevsky et al · 2019
Cited alongside, same era.
You look twice: Gaternet for dynamic filter selection in CNNs. In
Depth-Adaptive Transformer. In
Maha Elbayad et al · 2020
Later among the works it cites.
FlexDNN: Input-Adaptive On-Device Deep Learning for Efficient Mobile Vision. In
Biyi Fang et al · 2020
Later among the works it cites.
Resource Efficient Domain Adaptation. In
Junguang Jiang, Ximei Wang, Mingsheng Long, and Jianmin Wang. 2020 · 2020
Later among the works it cites.
Low Cost Early Exit Decision Unit Design for CNN Accelerator. In
Geonho Kim and Jongsun Park. 2020 · 2020
Later among the works it cites.
A 0.22–0.89 mW Low-Power and Highly-Secure Always-On Face Recognition Processor With Adversarial Attack Prevention
Youngwoo Kim et al · 2020
Later among the works it cites.
A throughput-latency co-optimised cascade of convolutional neural network classifiers. In
Alexandros Kouris et al · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhourong Chen, Yang Li, Samy Bengio, and Si Si. 2019 · 2019
Cited alongside, same era.
A novel design of adaptive and hierarchical convolutional neural networks using partial reconfiguration on FPGA. In
Mohammad Farhadi et al · 2019
Cited alongside, same era.
Dynamic channel pruning: Feature boosting and suppression
Xitong Gao, Yiren Zhao, Łukasz Dudziak, Robert Mullins, and Cheng-zhong Xu. 2019 · 2019
Cited alongside, same era.
Learning anytime predictions in neural networks via adaptive loss balancing. In
Hanzhang Hu et al · 2019
Cited alongside, same era.
AI Benchmark: All About Deep Learning on Smartphones in 2019. In
Andrey Ignatov et al · 2019
Cited alongside, same era.
Shallow-Deep Networks: Understanding and Mitigating Network Overthinking. In
Yigitcan Kaya, Sanghyun Hong, and Tudor Dumitras. 2019 · 2019
Cited alongside, same era.
MobiSR: Efficient On-Device Super-Resolution Through Heterogeneous Mobile Processors. In
Royson Lee et al · 2019
Cited alongside, same era.
Later among the works it cites.
Edge AI: On-Demand Accelerating Deep Neural Network Inference via Edge Computing. In
E. Li et al · 2020
Later among the works it cites.
Pruning Algorithms to Accelerate Convolutional Neural Networks for Edge Applications: A Survey
Jiayi Liu et al · 2020
Later among the works it cites.
Differentiable branching in deep networks for fast inference. In
Simone Scardapane et al · 2020
Later among the works it cites.
The Right Tool for the Job: Matching Model and Instance Complexities. In
Roy Schwartz et al · 2020
Later among the works it cites.
Fractional skipping: Towards finer-grained dynamic CNN inference. In
Jianghao Shen et al · 2020
Later among the works it cites.
The Cascade Transformer: an Application for Efficient Answer Sentence Selection. In
Luca Soldaini et al · 2020
Later among the works it cites.
Dual Dynamic Inference: Enabling more efficient, adaptive, and controllable deep inference
Yue Wang et al · 2020
Later among the works it cites.
Early Exiting BERT for Efficient Document Ranking. In
Ji Xin et al · 2020
Later among the works it cites.
Early exit or not: resource-efficient blind quality enhancement for compressed images. In
Qunliang Xing et al · 2020
Later among the works it cites.
Resolution adaptive networks for efficient inference. In
Le Yang, Yizeng Han, Xi Chen, Shiji Song, Jifeng Dai, and Gao Huang. 2020 · 2020
Later among the works it cites.
BERT Loses Patience: Fast and Robust Inference with Early Exit. In
Wangchunshu Zhou et al · 2020
Later among the works it cites.
Class-specific early exit design methodology for convolutional neural networks
Vanderlei Bonato and Christos Bouganis. 2021 · 2021
Closest in time.
Don’t shoot butterfly with rifles: Multi-channel Continuous Speech Separation with Early Exit Transformer. In
Sanyuan Chen et al · 2021
Closest in time.
A Survey of Quantization Methods for Efficient Neural Network Inference
Amir Gholami et al · 2021
Closest in time.
Class Means as an Early Exit Decision Mechanism
Alperen Gormez and Erdem Koyuncu. 2021 · 2021
Closest in time.
Dynamic neural networks: A survey
Yizeng Han et al · 2021
Closest in time.
A Panda? No, It’s a Sloth: Slowdown Attacks on Adaptive Multi-Exit Neural Network Inference. In
Sanghyun Hong et al · 2021
Closest in time.
FjORD: Fair and Accurate Federated Learning under heterogeneous targets with Ordered Dropout
Samuel Horvath, Stefanos Laskaridis, et al · 2021
Closest in time.
Multi-Exit Semantic Segmentation Networks
Alexandros Kouris, Stylianos I. Venieris, Stefanos Laskaridis, and Nicholas D. Lane. 2021 · 2021
Closest in time.
It’s Always Personal: Using Early Exits for Efficient On-Device CNN Personalisation
Ilias Leontiadis et al · 2021
Closest in time.
Split computing and early exiting for deep learning applications: Survey and research challenges
Yoshitomo Matsubara et al · 2021
Closest in time.