Fetching the paper…
Reading the bibliography…
With the recent demand of deploying neural network models on mobile and edge devices, it is desired to improve the model's generalizability on unseen testing data, as well as enhance the model's robustness under fixed-point quantization for efficient deployment.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava et al · 1958
Earlier work this paper cites.
A simple weight decay can improve generalization. In NeurIPS
Anders Krogh and John A Hertz. 1991 · 1991
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In ICCV
Jia Deng et al · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton. 2009 · 2009
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Yoshua Bengio et al · 2013
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Ian J Goodfellow et al · 2014
Earlier work this paper cites.
1.1 computing’s energy problem (and what we can do about it). In ISSCC
Mark Horowitz. 2014 · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman. 2014 · 2014
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition. In ICCV
Kaiming He et al · 2016
Earlier work this paper cites.
Deep networks with stochastic depth. In ECCV
Gao Huang et al · 2016
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang et al · 2016
Cited alongside, same era.
DoReFa-Net: Training low bitwidth convolutional neural networks with low bitwidth gradients
Shuchang Zhou et al · 2016
Cited alongside, same era.
Visualizing the loss landscape of neural nets
Hao Li et al · 2017
Cited alongside, same era.
mixup: Beyond empirical risk minimization
Hongyi Zhang et al · 2017
Cited alongside, same era.
Model compression via distillation and quantization
Antonio Polino et al · 2018
Later among the works it cites.
MobileNetV2: Inverted residuals and linear bottlenecks. In ICCV
Mark Sandler et al · 2018
Later among the works it cites.
LQ-Nets: Learned quantization for highly accurate and compact deep neural networks. In ECCV . 365–382
Dongqing Zhang et al · 2018
Later among the works it cites.
Robustness via curvature regularization, and vice versa. In ICCV
Seyed-Mohsen Moosavi-Dezfooli et al · 2019
Later among the works it cites.
Improving neural network quantization without retraining using outlier channel splitting. In ICML . 7543–7552
Ritchie Zhao et al · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ron Banner et al · 2018
Cited alongside, same era.
AutoAugment: Learning augmentation policies from data
Ekin D Cubuk et al · 2018
Cited alongside, same era.
Empirical study of the topology and geometry of deep networks. In ICCV
Alhussein Fawzi et al · 2018
Cited alongside, same era.
Towards deep learning models resistant to adversarial attacks. In ICLR
Aleksander Madry et al · 2018
Cited alongside, same era.
Milad Alizadeh et al · 2020
Later among the works it cites.
Sharpness-aware minimization for efficiently improving generalization
Pierre Foret et al · 2020
Later among the works it cites.
DivideMix: Learning with noisy labels as semi-supervised learning
Junnan Li et al · 2020
Later among the works it cites.
BSQ: Exploring bit-level sparsity for mixed-Precision neural network quantization
Huanrui Yang et al · 2021
Closest in time.