Quantization and training of neural networks for efficient integer-arithmetic-only inference
Benoit Jacob, Skirmantas Kligys, Bo Chen, Menglong Zhu, Matthew Tang, Andrew Howard, Hartwig Adam, and Dmitry Kalenichenko · 2018
Later among the works it cites.
Modeling task relationships in multi-task learning with multi-gate mixture-of-experts
Jiaqi Ma, Zhe Zhao, Xinyang Yi, Jilin Chen, Lichan Hong, and Ed H Chi · 2018
Later among the works it cites.
Knowledge transfer with jacobian matching
Original
Suraj Srinivas and François Fleuret · 2018
Later among the works it cites.
Ranking distillation: Learning compact ranking models with high performance for recommender system
Jiaxi Tang and Ke Wang · 2018
Later among the works it cites.
Heated-up softmax embedding
Original
Xu Zhang, Felix Xinnan Yu, Svebor Karaman, Wei Zhang, and Shih-Fu Chang · 2018
Later among the works it cites.
Bag of tricks for image classification with convolutional neural networks
Tong He, Zhi Zhang, Hang Zhang, Zhongyue Zhang, Junyuan Xie, and Mu Li · 2019
Later among the works it cites.
Knowledge distillation with adversarial samples supporting decision boundary
Byeongho Heo, Minsik Lee, Sangdoo Yun, and Jin Young Choi · 2019
Later among the works it cites.
Improved knowledge distillation via teacher assistant: Bridging the gap between student and teacher
Original
Seyed-Iman Mirzadeh, Mehrdad Farajtabar, Ang Li, and Hassan Ghasemzadeh · 2019
Later among the works it cites.
When does label smoothing help?
Original
Rafael Müller, Simon Kornblith, and Geoffrey Hinton · 2019
Later among the works it cites.
Towards understanding knowledge distillation
Mary Phuong and Christoph Lampert · 2019
Later among the works it cites.
Revisit knowledge distillation: a teacher-free framework
Original
Li Yuan, Francis EH Tay, Guilin Li, Tao Wang, and Jiashi Feng · 2019
Later among the works it cites.