Fetching the paper…
Reading the bibliography…
Model compression techniques on Deep Neural Network (DNN) have been widely acknowledged as an effective way to achieve acceleration on a variety of platforms, and DNN weight pruning is a straightforward and effective method.
Nothing clear enough to list yet.
https://github.com/alibaba/MNN
Cited in the paper.
https://www.tensorflow.org/mobile/tflite/
Cited in the paper.
Nothing clear enough to list yet.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…