Fetching the paper…
Reading the bibliography…
In recent years, deep neural networks have been successful in both industry and academia, especially for computer vision tasks.
1902
Earlier work this paper cites.
1902
Earlier work this paper cites.
1903
Earlier work this paper cites.
1904
Earlier work this paper cites.
1904
Earlier work this paper cites.
1905
Earlier work this paper cites.
1905
Earlier work this paper cites.
1906
Earlier work this paper cites.
1908
Earlier work this paper cites.
1908
Earlier work this paper cites.
1909
Earlier work this paper cites.
1909
Earlier work this paper cites.
1909
Earlier work this paper cites.
1910
Earlier work this paper cites.
1910
Earlier work this paper cites.
1911
Earlier work this paper cites.
1911
Earlier work this paper cites.
1911
Earlier work this paper cites.
1912
Earlier work this paper cites.
Li, J., Fu, K., Zhao, S. & Ge, S. (2019). Spatiotemporal knowledge distillation for efficient estimation of aerial video saliency. IEEE TIP
1914
Earlier work this paper cites.
Yuan, M., & Peng, Y. (2020). CKD: Cross-task knowledge distillation for text-to-image synthesis. IEEE TMM
1968
Earlier work this paper cites.
2002
Earlier work this paper cites.
2004
Earlier work this paper cites.
2006
Earlier work this paper cites.
Bucilua, C., Caruana, R. & Niculescu-Mizil, A. (2006). Model compression. In: SIGKDD
2006
Earlier work this paper cites.
Deng, J., Dong, W., Socher, R., Li, L. J., Li, K., & Fei-Fei, L. (2009). Imagenet: A large-scale hierarchical image database. In: CVPR
2009
Earlier work this paper cites.
Krizhevsky, A., & Hinton, G. (2009). Learning multiple layers of features from tiny images
2009
Earlier work this paper cites.
2011
Earlier work this paper cites.
Urner, R., Shalev-Shwartz, S., Ben-David, S. (2011). Access to unlabeled data can speed up prediction time. In ICML
2011
Earlier work this paper cites.
Krizhevsky, A., Sutskever, I. & Hinton, G. E. (2012). Imagenet classification with deep convolutional neural networks. In: NeurIPS
2012
Earlier work this paper cites.
Bengio, Y., Courville, A., & Vincent, P. (2013). Representation learning: A review and new perspectives. IEEE TPAMI
2013
Earlier work this paper cites.
Ba, J. & Caruana, R. (2014). Do deep nets really need to be deep? In: NeurIPS
2014
Earlier work this paper cites.
Denton, E. L., Zaremba, W., Bruna, J., LeCun, Y. & Fergus, R. (2014). Exploiting linear structure within convolutional networks for efficient evaluation. In: NeurIPS
2014
Earlier work this paper cites.
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., & Bengio, Y. (2014). Generative adversarial nets. In: NeurIPS
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
Courbariaux, M., Bengio, Y. & David, J. P. (2015). Binaryconnect: Training deep neural networks with binary weights during propagations. In: NeurIPS
2015
Earlier work this paper cites.
Han, S., Pool, J., Tran, J. & Dally, W. (2015). Learning both weights and connections for efficient neural network. In: NeurIPS
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
Ioffe, S., & Szegedy, C. (2015). Batch normalization: Accelerating deep network training by reducing internal covariate shift. In: ICML
2015
Earlier work this paper cites.
Romero, A., Ballas, N., Kahou, S. E., Chassang, A., Gatta, C., & Bengio, Y. (2015). Fitnets: Hints for thin deep nets. In: ICLR
2015
Earlier work this paper cites.
Sindhwani, V., Sainath, T. & Kumar, S. (2015). Structured transforms for small-footprint deep learning. In: NeurIPS
2015
Earlier work this paper cites.
Chebotar, Y. & Waters, A. (2016). Distilling knowledge from ensembles of neural networks for speech recognition. In: Interspeech
2016
Earlier work this paper cites.
Chen, T., Goodfellow, I. & Shlens, J. (2016) Net2net: Accelerating learning via knowledge transfer. In: ICLR
2016
Earlier work this paper cites.
Gupta, S., Hoffman, J. & Malik, J. (2016). Cross modal distillation for supervision transfer. In: CVPR
2016
Earlier work this paper cites.
He, K., Zhang, X., Ren, S. & Sun, J. (2016) Deep residual learning for image recognition. In: CVPR
2016
Earlier work this paper cites.
Hoffman, J., Gupta, S. & Darrell, T. (2016). Learning with side information through modality hallucination. In: CVPR
2016
Earlier work this paper cites.
Kim, Y., Rush & A. M. (2016). Sequence-level knowledge distillation. In: EMNLP
2016
Earlier work this paper cites.
Kuncoro, A., Ballesteros, M., Kong, L., Dyer, C. & Smith, N. A. (2016). Distilling an ensemble of greedy dependency parsers into one mst parser. In: EMNLP
2016
Earlier work this paper cites.
Lopez-Paz, D., Bottou, L., Schölkopf, B. & Vapnik, V. (2016). Unifying distillation and privileged information. In: ICLR
2016
Earlier work this paper cites.
Luo, P., Zhu, Z., Liu, Z., Wang, X. & Tang, X. (2016). Face model compression by distilling knowledge from neurons. In: AAAI
2016
Earlier work this paper cites.
Mou, L., Jia, R., Xu, Y., Li, G., Zhang, L. & Jin, Z. (2016). Distilling word embeddings: An encoding approach. In: CIKM
2016
Earlier work this paper cites.
Papernot, N., McDaniel, P., Wu, X., Jha, S. & Swami, A. (2016). Distillation as a defense to adversarial perturbations against deep neural networks. In: IEEE SP
2016
Earlier work this paper cites.
Price, R., Iso, K. & Shinoda, K. (2016). Wise teachers train better dnn acoustic models. EURASIP Journal on Audio, Speech, and Music Processing
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., … & Dieleman, S. (2016). Mastering the game of Go with deep neural networks and tree search. Nature
2016
Earlier work this paper cites.
Wong, J. H. & Gales, M. (2016). Sequence student-teacher training of deep neural networks. In: Interspeech
2016
Earlier work this paper cites.
Wu, J., Leng, C., Wang, Y., Hu, Q. & Cheng, J. (2016). Quantized convolutional neural networks for mobile devices. In: CVPR
2016
Earlier work this paper cites.
Zhai, S., Cheng, Y., Zhang, Z. M. & Lu, W. (2016). Doubly convolutional neural networks. In: NeurIPS
2016
Earlier work this paper cites.
Asami, T., Masumura, R., Yamaguchi, Y., Masataki, H. & Aono, Y. (2017). Domain adaptation of dnn acoustic models using knowledge distillation. In: ICASSP
2017
Earlier work this paper cites.
Chen, G., Choi, W., Yu, X., Han, T., & Chandraker, M. (2017). Learning efficient object detection models with knowledge distillation. In: NeurIPS
2017
Earlier work this paper cites.
Chollet, F. (2017). Xception: Deep learning with depthwise separable convolutions. In: CVPR
2017
Earlier work this paper cites.
Cui, J., Kingsbury, B., Ramabhadran, B., Saon, G., Sercu, T., Audhkhasi, K. & et al. (2017). Knowledge distillation across ensembles of multilingual models for low-resource languages. In: ICASSP
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Fukuda, T., Suzuki, M., Kurata, G., Thomas, S., Cui, J. & Ramabhadran, B. (2017). Efficient knowledge distillation from an ensemble of teachers. In: Interspeech
2017
Earlier work this paper cites.
Gong, C., Tao, D., Liu, W., Liu, L., & Yang, J. (2017). Label propagation via teaching-to-learn and learning-to-teach. TNNLS
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Huang, G., Liu, Z., Van, Der Maaten, L. & Weinberger, K. Q. (2017). Densely connected convolutional networks. In: CVPR
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Kim, S. W. & Kim, H. E. (2017). Transferring knowledge to smaller network with class-distance loss. In: ICLRW
2017
Earlier work this paper cites.
Li, Q., Jin, S. & Yan, J. (2017). Mimicking very efficient network for object detection. In: CVPR
2017
Earlier work this paper cites.
Li, Z. & Hoiem, D. (2017). Learning without forgetting. IEEE TPAMI
2017
Earlier work this paper cites.
Lopes, R. G., Fenu, S. & Starner, T. (2017). Data-free knowledge distillation for deep neural networks. In: NeurIPS
2017
Earlier work this paper cites.
Lu, L., Guo, M. & Renals, S. (2017). Knowledge distillation for small-footprint highway networks. In: ICASSP
2017
Earlier work this paper cites.
Nakashole, N. & Flauger, R. (2017). Knowledge distillation for bilingual dictionary induction. In: EMNLP
2017
Earlier work this paper cites.
Papernot, N., Abadi, M., Erlingsson, U., Goodfellow, I. & Talwar, K. (2017). Semi-supervised knowledge transfer for deep learning from private training data. In: ICLR
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Shmelkov, K., Schmid, C. & Alahari, K. (2017). Incremental learning of object detectors without catastrophic forgetting. In: ICCV
2017
Earlier work this paper cites.
Su, J. C. & Maji, S. (2017). Adapting models to signal degradation using distillation. In: BMVC
2017
Earlier work this paper cites.
Tarvainen, A., & Valpola, H. (2017). Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results. In: NeurIPS
2017
Earlier work this paper cites.
Urban, G., Geras, K. J., Kahou, S. E., Aslan, O., Wang, S., Caruana, R. & et al. (2017). Do deep convolutional nets really need to be deep and convolutional? In: ICLR
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Watanabe, S., Hori, T., Le Roux, J. & Hershey, J. R. (2017). Student-teacher network learning with enhanced features. In: ICASSP
2017
Earlier work this paper cites.
Yim, J., Joo, D., Bae, J. & Kim, J. (2017). A gift from knowledge distillation: Fast optimization, network minimization and transfer learning. In: CVPR
2017
Earlier work this paper cites.
You, S., Xu, C., Xu, C. & Tao, D. (2017). Learning from multiple teacher networks. In: SIGKDD
2017
Earlier work this paper cites.
Yu, X., Liu, T., Wang, X., & Tao, D. (2017). On compressing deep models by low rank and sparse decomposition. In: CVPR
2017
Earlier work this paper cites.
Zagoruyko, S. & Komodakis, N. (2017). Paying more attention to attention: Improving the performance of convolutional neural networks via attention transfer. In: ICLR
2017
Earlier work this paper cites.
Albanie, S., Nagrani, A., Vedaldi, A. & Zisserman, A. (2018). Emotion recognition in speech using cross-modal transfer in the wild. In: ACM MM
2018
Earlier work this paper cites.
Anil, R., Pereyra, G., Passos, A., Ormandi, R., Dahl, G. E.. & Hinton, G. E. (2018). Large scale distributed neural network training through online distillation. In: ICLR
2018
Cited alongside, same era.
Arora, S., Cohen, N., & Hazan, E. (2018). On the optimization of deep networks: Implicit acceleration by overparameterization. In: ICML
2018
Cited alongside, same era.
Ashok, A., Rhinehart, N., Beainy, F. & Kitani, K. M. (2018). N2N learning: Network to network compression via policy gradient reinforcement learning. In: ICLR
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Belagiannis, V., Farshad, A. & Galasso, F. (2018). Adversarial network compression. In: ECCV
Tan, X., Ren, Y., He, D., Qin, T., Zhao, Z. & Liu, T. Y. (2019). Multilingual neural machine translation with knowledge distillation. In: ICLR
2019
Later among the works it cites.
Thoker, F. M. & Gall, J. (2019). Cross-modal knowledge distillation for action recognition. In: ICIP
2019
Later among the works it cites.
Tung, F. & Mori, G. (2019). Similarity-preserving knowledge distillation. In: ICCV
2019
Later among the works it cites.
Vongkulbhisal, J., Vinayavekhin, P. & Visentini-Scarzanella, M. (2019). Unifying heterogeneous classifiers with distillation. In: CVPR
2019
Later among the works it cites.
Wei, H. R., Huang, S., Wang, R., Dai, X. & Chen, J. (2019). Online distilling from checkpoints for neural machine translation. In: NAACL-HLT
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Chen, Z. & Liu, B. (2018). Lifelong machine learning. Synthesis Lectures on Artificial Intelligence and Machine Learning
2018
Cited alongside, same era.
Cheng, Y., Wang, D., Zhou, P. & Zhang, T. (2018). Model compression and acceleration for deep neural networks: The principles, progress, and challenges. IEEE Signal Proc Mag
2018
Cited alongside, same era.
Crowley, E. J., Gray, G. & Storkey, A. J. (2018). Moonshine: Distilling with cheap convolutions. In: NeurIPS
2018
Cited alongside, same era.
Furlanello, T., Lipton, Z., Tschannen, M., Itti, L. & Anandkumar, A. (2018). Born again neural networks. In: ICML
2018
Cited alongside, same era.
Garcia, N. C., Morerio, P. & Murino, V. (2018). Modality distillation with multiple stream networks for action recognition. In: ECCV
2018
Cited alongside, same era.
Ghorbani, S., Bulut, A. E. & Hansen, J. H. (2018). Advancing multi-accented lstm-ctc speech recognition using a domain specific student-teacher learning paradigm. In: SLTW
2018
Cited alongside, same era.
Gong, C., Chang, X., Fang, M. & Yang, J. (2018). Teaching semi-supervised classifier via generalized distillation. In: IJCAI
2018
Cited alongside, same era.
Wu, B., Dai, X., Zhang, P., Wang, Y., Sun, F., Wu, Y., … & Keutzer, K. (2019). Fbnet: Hardware-aware efficient convnet design via differentiable neural architecture search. In: CVPR
2019
Later among the works it cites.
Xu, T. B., & Liu, C. L. (2019). Data-distortion guided self-distillation for deep neural networks. In: AAAI
2019
Later among the works it cites.
Yan, M., Zhao, M., Xu, Z., Zhang, Q., Wang, G. & Su, Z. (2019). Vargfacenet: An efficient variable group convolutional neural network for lightweight face recognition. In: ICCVW
2019
Later among the works it cites.
Ye, J., Ji, Y., Wang, X., Ou, K., Tao, D. & Song, M. (2019). Student becoming the master: Knowledge amalgamation for joint scene parsing, depth estimation, and more. In: CVPR
2019
Later among the works it cites.
Yoo, J., Cho, M., Kim, T., & Kang, U. (2019). Knowledge extraction with no observable data. In: NeurIPS
2019
Later among the works it cites.
You, Y., Li, J., Reddi, S., Hseu, J., Kumar, S., Bhojanapalli, S., … & Hsieh, C. J. (2019). Large batch optimization for deep learning: Training bert in 76 minutes. In: ICLR
2019
Later among the works it cites.
Yu, L., Yazici, V. O., Liu, X., Weijer, J., Cheng, Y. & Ramisa, A. (2019). Learning metrics from teachers: Compact networks for image embedding. In: CVPR
2019
Later among the works it cites.
Zhai, M., Chen, L., Tung, F., He, J., Nawhal, M. & Mori, G. (2019). Lifelong gan: Continual learning for conditional image generation. In: ICCV
2019
Later among the works it cites.
Zhu, M., Han, K., Zhang, C., Lin, J. & Wang, Y. (2019). Low-resolution visual recognition via deep feature distillation. In: ICASSP
2019
Later among the works it cites.
Aguilar, G., Ling, Y., Zhang, Y., Yao, B., Fan, X. & Guo, E. (2020). Knowledge distillation from internal representations. In: AAAI
2020
Closest in time.
Asif, U., Tang, J. & Harrer, S. (2020). Ensemble knowledge distillation for learning improved and efficient networks. In: ECAI
2020
Closest in time.
Bai, H., Wu, J., King, I. & Lyu, M. (2020). Few shot network compression via cross distillation. In: AAAI
2020
Closest in time.
Bergmann, P., Fauser, M., Sattlegger, D., & Steger, C. (2020). Uninformed students: Student-teacher anomaly detection with discriminative latent embeddings. In: CVPR
2020
Closest in time.
Bistritz, I., Mann, A., & Bambos, N. (2020). Distributed Distillation for On-Device Learning. In: NeurIPS
2020
Closest in time.
Caccia, M., Rodriguez, P., Ostapenko, O., Normandin, F., Lin, M., Caccia, L., Laradji, I., Rish, I., Lacoste, A., Vazquez D., & Charlin, L. (2020). Online Fast Adaptation and Knowledge Accumulation (OSAKA): a New Approach to Continual Learning. In: NeurIPS
2020
Closest in time.
Cheng, X., Rao, Z., Chen, Y., & Zhang, Q. (2020). Explaining Knowledge Distillation by Quantifying the Knowledge. In: CVPR
2020
Closest in time.
Chung, I., Park, S., Kim, J. & Kwak, N. (2020). Feature-map-level online adversarial knowledge distillation. In: ICML
2020
Closest in time.
Cui, Z., Song, T., Wang, Y., & Ji, Q. (2020). Knowledge Augmented Deep Neural Networks for Joint Facial Expression and Action Unit Recognition. In: NeurIPS
2020
Closest in time.
Cun, X., & Pun, C. M. (2020). Defocus Blur Detection via Depth Distillation. In: ECCV
2020
Closest in time.
Dou, Q., Liu, Q., Heng, P. A., & Glocker, B. (2020). Unpaired multi-modal segmentation via knowledge distillation. IEEE TMI
2020
Closest in time.
Du, S., You, S., Li, X., Wu, J., Wang, F., Qian, C., & Zhang, C. (2020). Agree to Disagree: Adaptive Ensemble Knowledge Distillation in Gradient Space. In: NeurIPS
2020
Closest in time.
Fakoor, R., Mueller, J. W., Erickson, N., Chaudhari, P., & Smola, A. J. (2020). Fast, Accurate, and Simple Models for Tabular Data via Augmented Distillation. In: NeurIPS
2020
Closest in time.
Gao, Z., Chung, J., Abdelrazek, M., Leung, S., Hau, W. K., Xian, Z., Zhang, H., & Li, S. (2020). Privileged modality distillation for vessel border detection in intracoronary imaging. IEEE TMI
2020
Closest in time.
Ge, S., Zhao, S., Li, C., Zhang, Y., & Li, J. (2020). Efficient Low-Resolution Face Recognition via Bridge Distillation. IEEE TIP
2020
Closest in time.
Goldblum, M., Fowl, L., Feizi, S. & Goldstein, T. (2020). Adversarially robust distillation. In: AAAI
2020
Closest in time.
Gu, J., & Tresp, V. (2020). Search for Better Students to Learn Distilled Knowledge. In: ECAI
2020
Closest in time.
Guan, Y., Zhao, P., Wang, B., Zhang, Y., Yao, C., Bian, K., & Tang, J. (2020). Differentiable Feature Aggregation Search for Knowledge Distillation. In: ECCV
2020
Closest in time.
Guo, Q., Wang, X., Wu, Y., Yu, Z., Liang, D., Hu, X., & Luo, P. (2020). Online Knowledge Distillation via Collaborative Learning. In: CVPR
2020
Closest in time.
Haroush, M., Hubara, I., Hoffer, E., & Soudry, D. (2020). The knowledge within: Methods for data-free model compression. In: CVPR
2020
Closest in time.
Hou, Y., Ma, Z., Liu, C., Hui, T. W., & Loy, C. C. (2020). Inter-Region Affinity Distillation for Road Marking Segmentation. In: CVPR
2020
Closest in time.
Hu, H., Xie, L., Hong, R., & Tian, Q. (2020). Creating Something from Nothing: Unsupervised Knowledge Distillation for Cross-Modal Hashing. In: CVPR
2020
Closest in time.
Huang, Z., Zou, Y., Bhagavatula, V., & Huang, D. (2020). Comprehensive Attention Self-Distillation for Weakly-Supervised Object Detection. In: NeurIPS
2020
Closest in time.
Ji, G., & Zhu, Z. (2020). Knowledge Distillation in Wide Neural Networks: Risk Bound, Data Efficiency and Imperfect Teacher. In: NeurIPS
2020
Closest in time.
Jiao, X., Yin, Y., Shang, L., Jiang, X., Chen, X., Li, L. & et al. (2020). Tinybert: Distilling bert for natural language understanding. In: EMNLP
2020
Closest in time.
Kang, M., Mun, J. & Han, B. (2020). Towards oracle knowledge distillation with neural architecture search. In: AAAI
2020
Closest in time.
Kwon, K., Na, H., Lee, H., & Kim, N. S. (2020). Adaptive Knowledge Distillation Based on Entropy. In: ICASSP
2020
Closest in time.
Lai, K. H., Zha, D., Li, Y., & Hu, X. (2020). Dual Policy Distillation. In: IJCAI
2020
Closest in time.
Lin, T., Kong, L., Stich, S. U., & Jaggi, M. (2020). Ensemble distillation for robust model fusion in federated learning. In: NeurIPS
2020
Closest in time.
Luo, S., Pan, W., Wang, X., Wang, D., Tang, H., & Song, M. (2020). Collaboration by Competition: Self-coordinated Knowledge Amalgamation for Multi-talent Student Learning. In: ECCV
2020
Closest in time.
Mirzadeh, S. I., Farajtabar, M., Li, A. & Ghasemzadeh, H. (2020). Improved knowledge distillation via teacher assistant. In: AAAI
2020
Closest in time.
Mobahi, H., Farajtabar, M., & Bartlett, P. L. (2020). Self-distillation amplifies regularization in hilbert space. In: NeurIPS
2020
Closest in time.
Pan, B., Cai, H., Huang, D. A., Lee, K. H., Gaidon, A., Adeli, E., & Niebles, J. C. (2020). Spatio-Temporal Graph for Video Captioning with Knowledge Distillation. In: CVPR
2020
Closest in time.
Park, S. & Kwak, N. (2020). Feature-level Ensemble Knowledge Distillation for Aggregating Knowledge from Multiple Networks. In: ECAI
2020
Closest in time.
Passalis, N., Tzelepi, M., & Tefas, A. (2020a). Probabilistic Knowledge Transfer for Lightweight Deep Representation Learning. TNNLS
2020
Closest in time.
Peng, H., Du, H., Yu, H., Li, Q., Liao, J., & Fu, J. (2020). Cream of the Crop: Distilling Prioritized Paths For One-Shot Neural Architecture Search. In: NeurIPS
2020
Closest in time.
Perez, A., Sanguineti, V., Morerio, P. & Murino, V. (2020). Audio-visual model distillation using acoustic images. In: WACV
2020
Closest in time.
Radosavovic, I., Kosaraju, R. P., Girshick, R., He, K., & Dollar P. (2020). Designing network design spaces. In: CVPR
2020
Closest in time.
Shen, P., Lu, X., Li, S., & Kawai, H. (2020). Knowledge Distillation-Based Representation Learning for Short-Utterance Spoken Language Identification. IEEE/ACM T AUDIO SPE
2020
Closest in time.
Tian, Y., Krishnan, D. & Isola, P. (2020). Contrastive representation distillation. In: ICLR
2020
Closest in time.
Tu, Z., He, F., & Tao, D. (2020). Understanding Generalization in Recurrent Neural Networks. In International Conference on Learning Representations. In: ICLR
2020
Closest in time.
Walawalkar, D., Shen, Z., & Savvides, M. (2020). Online Ensemble Model Compression using Knowledge Distillation. In: ECCV
2020
Closest in time.
Wu, X., He, R., Hu, Y., & Sun, Z. (2020). Learning an evolutionary embedding via massive knowledge distillation. International Journal of Computer Vision , 1-18
2020
Closest in time.
Xie, Q., Hovy, E., Luong, M. T., & Le, Q. V. (2020). Self-training with Noisy Student improves ImageNet classification. In: CVPR
2020
Closest in time.
Yao, A., & Sun, D. (2020). Knowledge Transfer via Dense Cross-Layer Mutual-Distillation. In: ECCV
2020
Closest in time.
Yao, H., Zhang, C., Wei, Y., Jiang, M., Wang, S., Huang, J., Chawla, N. V., & Li, Z. (2020). Graph Few-shot Learning via Knowledge Transfer. In: AAAI
2020
Closest in time.
Ye, J., Ji, Y., Wang, X., Gao, X., & Song, M. (2020). Data-Free Knowledge Amalgamation via Group-Stack Dual-GAN. In: CVPR
2020
Closest in time.
Yin, H., Molchanov, P., Alvarez, J. M., Li, Z., Mallya, A., Hoiem, D., Jha, Niraj K., & Kautz, J. (2020). Dreaming to distill: Data-free knowledge transfer via DeepInversion. In: CVPR
2020
Closest in time.
Yuan, L., Tay, F. E., Li, G., Wang, T. & Feng, J. (2020). Revisit knowledge distillation: a teacher-free framework. In: CVPR
2020
Closest in time.
Yue, K., Deng, J., & Zhou, F. (2020). Matching Guided Distillation. In: ECCV
2020
Closest in time.
Yun, S., Park, J., Lee, K. & Shin, J. (2020). Regularizing Class-wise Predictions via Self-knowledge Distillation. In: CVPR
2020
Closest in time.
Zhao, C., & Hospedales, T. (2020). Robust Domain Randomised Reinforcement Learning through Peer-to-Peer Distillation. In: NeurIPS
2020
Closest in time.
Zhao, H., Sun, X., Dong, J., Chen, C., & Dong, Z. (2020a). Highlight every step: Knowledge distillation via collaborative teaching. IEEE TCYB
2020
Closest in time.
Zhang, Z., & Sabuncu, M. R. (2020). Self-Distillation as Instance-Specific Label Smoothing. In: NeurIPS
2020
Closest in time.
Zhou, P., Mai, L., Zhang, J., Xu, N., Wu, Z. & Davis, L. S. (2020). M2KD: Multi-model and multi-level knowledge distillation for incremental learning. In: BMVC
2020
Closest in time.
Boo, Y., Shin, S., Choi, J., & Sung, W. (2021). Stochastic Precision Ensemble: Self-Knowledge Distillation for Quantized Deep Neural Networks. In: AAAI
2021
Closest in time.
Chawla, A., Yin, H., Molchanov, P., & Alvarez, J. (2021). Data-Free Knowledge Distillation for Object Detection. In: WACV
2021
Closest in time.
Chen, D., Mei, J. P., Zhang, Y., Wang, C., Wang, Z., Feng, Y., & Chen, C. (2021). Cross-Layer Distillation with Semantic Calibration. In: AAAI
2021
Closest in time.
Chen, H., Wang, Y., Xu, C., Xu, C. & Tao, D. (2021). Learning student networks via feature embedding. IEEE TNNLS
2021
Closest in time.
Fu, H., Zhou, S., Yang, Q., Tang, J., Liu, G., Liu, K., & Li, X. (2021). LRC-BERT: Latent-representation Contrastive Knowledge Distillation for Natural Language Understanding. In: AAAI
2021
Closest in time.
Gao, M., Wang, Y., & Wan, L. (2021). Residual Error Based Knowledge Distillation. Neurocomputing
2021
Closest in time.
Li, B., Wang, Z., Liu, H., Du, Q., Xiao, T., Zhang, C., & Zhu, J. (2021). Learning Light-Weight Translation Models from Deep Transformer. In: AAAI
2021
Closest in time.
Nayak, G. K., Mopuri, K. R., & Chakraborty, A. (2021). Effectiveness of Arbitrary Transfer Sets for Data-free Knowledge Distillation. In: WACV
2021
Closest in time.
Passban, P., Wu, Y., Rezagholizadeh, M., & Liu, Q. (2021). ALP-KD: Attention-Based Layer Projection for Knowledge Distillation. In: AAAI
2021
Closest in time.
Shen, C., Wang, X., Yin, Y., Song, J., Luo, S., & Song, M. (2021). Progressive Network Grafting for Few-Shot Knowledge Distillation. In: AAAI
2021
Closest in time.
2021
Closest in time.
Tan, H., Liu, X., Liu, M., Yin, B., & Li, X. (2021). KT-GAN: Knowledge-Transfer Generative Adversarial Network for Text-to-Image Synthesis. IEEE TIP
2021
Closest in time.
Wang, Z. R., & Du, J. (2021). Joint architecture and knowledge distillation in CNN for Chinese text recognition. Pattern Recognition
2021
Closest in time.
Wu, G., & Gong, S. (2021). Peer Collaborative Learning for Online Knowledge Distillation. In: AAAI
2021
Closest in time.
Yuan, F., Shou, L., Pei, J., Lin, W., Gong, M., Fu, Y., & Jiang, D. (2021). Reinforced Multi-Teacher Selection for Knowledge Distillation. In: AAAI
2021
Closest in time.
Vapnik, V. & Izmailov, R. (2015). Learning using privileged information: similarity control and knowledge transfer. J Mach Learn Res
2049
Closest in time.
Ge, S., Zhao, S., Li, C. & Li, J. (2018). Low-resolution face recognition in the wild via selective knowledge distillation. IEEE TIP
2062
Closest in time.
Xia, S., Wang, G., Chen, Z., & Duan, Y. (2018). Complete random forest based class noise filtering learning for improving the generalizability of classifiers. IEEE TKDE 31(11): 2063-2078
2078
Closest in time.