Rotating your face using multi-task deep neural network. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 676–684
Junho Yim, Heechul Jung, ByungIn Yoo, Changkyu Choi, Dusik Park, and Junmo Kim. 2015 · 2015
Cited alongside, same era.
Sequence-Level Knowledge Distillation. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing, EMNLP 2016, Austin, Texas, USA, November 1-4, 2016 . 1317–1327
Yoon Kim and Alexander M. Rush. 2016 · 2016
Cited alongside, same era.
Distilling an Ensemble of Greedy Dependency Parsers into One MST Parser. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing, EMNLP 2016, Austin, Texas, USA, November 1-4, 2016 . 1744–1753
Adhiguna Kuncoro, Miguel Ballesteros, Lingpeng Kong, Chris Dyer, and Noah A. Smith. 2016 · 2016
Cited alongside, same era.
Channel Pruning for Accelerating Very Deep Neural Networks. In IEEE International Conference on Computer Vision, ICCV 2017, Venice, Italy, October 22-29, 2017 . 1398–1406
Yihui He, Xiangyu Zhang, and Jian Sun. 2017 · 2017
Cited alongside, same era.
Semi-supervised Knowledge Transfer for Deep Learning from Private Training Data. In 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Conference Track Proceedings
Nicolas Papernot, Martín Abadi, Úlfar Erlingsson, Ian J. Goodfellow, and Kunal Talwar. 2017 · 2017
Cited alongside, same era.
Multi-Task Learning with Labeled and Unlabeled Tasks
Anastasia Pentina and Christoph H Lampert. 2017 · 2017
Cited alongside, same era.
Learning from Multiple Teacher Networks. In Proceedings of the 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, Halifax, NS, Canada, August 13 - 17, 2017 . 1285–1294
Shan You, Chang Xu, Chao Xu, and Dacheng Tao. 2017 · 2017
Cited alongside, same era.
Transfer and Multi-Task Learning for Noun-Noun Compound Interpretation. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, Brussels, Belgium, October 31 - November 4, 2018 . 1488–1498
Murhaf Fares, Stephan Oepen, and Erik Velldal. 2018 · 2018
Cited alongside, same era.
The lottery ticket hypothesis: Finding sparse, trainable neural networks
Original
Jonathan Frankle and Michael Carbin. 2018 · 2018
Cited alongside, same era.
Universal Language Model Fine-tuning for Text Classification. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics, ACL 2018, Melbourne, Australia, July 15-20, 2018, Volume 1: Long Papers . 328–339
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Cited alongside, same era.