Fetching the paper…
Reading the bibliography…
Deep pre-training and fine-tuning models (like BERT, OpenAI GPT) have demonstrated excellent results in question answering areas.
A vector space model for automatic indexing
Gerard Salton, Anita Wong, and Chung-Shu Yang. 1975 · 1975
Earlier work this paper cites.
Optimal brain damage
Yann LeCun, John S. Denker, and Sara A. Solla. 1989 · 1989
Earlier work this paper cites.
Efficient and accurate approximations of nonlinear convolutional networks
Xiangyu Zhang, Jianhua Zou, Xiang Ming, Kaiming He, and Jian Sun. 2015 · 1992
Earlier work this paper cites.
Second order derivatives for network pruning: Optimal brain surgeon
Babak Hassibi and David G. Stork. 1993 · 1993
Earlier work this paper cites.
Okapi at trec-7: automatic ad hoc, filtering, vlc and interactive track
Stephen E Robertson, Steve Walker, Micheline Beaulieu, and Peter Willett. 1999 · 1999
Earlier work this paper cites.
Retrieval models for question and answer archives
Xiaobing Xue, Jiwoon Jeon, and W Bruce Croft. 2008 · 2008
Earlier work this paper cites.
Learning deep structured semantic models for web search using clickthrough data
Po-Sen Huang, Xiaodong He, Jianfeng Gao, Li Deng, Alex Acero, and Larry P. Heck. 2013 · 2013
Earlier work this paper cites.
Ontology-based interpretation of natural language
Philipp Cimiano, Christina Unger, and John McCrae. 2014 · 2014
Earlier work this paper cites.
Exploiting linear structure within convolutional networks for efficient evaluation
Emily L. Denton, Wojciech Zaremba, Joan Bruna, Yann LeCun, and Rob Fergus. 2014 · 2014
Earlier work this paper cites.
Speeding up convolutional neural networks with low rank expansions
Max Jaderberg, Andrea Vedaldi, and Andrew Zisserman. 2014 · 2014
Cited alongside, same era.
Semantic modelling with long-short-term memory for information retrieval
Hamid Palangi, Li Deng, Yelong Shen, Jianfeng Gao, Xiaodong He, Jianshu Chen, Xinying Song, and Rabab K. Ward. 2014 · 2014
Cited alongside, same era.
A latent semantic model with convolutional-pooling structure for information retrieval
Yelong Shen, Xiaodong He, Jianfeng Gao, Li Deng, and Grégoire Mesnil. 2014 · 2014
Cited alongside, same era.
Distilling the knowledge in a neural network
Geoffrey E Hinton, Oriol Vinyals, and Jeffrey Dean. 2015 · 2015
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Adversarial transfer learning for chinese named entity recognition with self-attention mechanism
Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao, and Shengping Liu. 2018 · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Later among the works it cites.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Later among the works it cites.
Unsupervised cross-dataset person re-identification by transfer learning of spatial-temporal patterns
Jianming Lv, Weihang Chen, Qing Li, and Can Yang. 2018 · 2018
Later among the works it cites.
Image to image translation for domain adaptation
Zak Murez, Soheil Kolouri, David J. Kriegman, Ravi Ramamoorthi, and Kyungnam Kim. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mu Li, Wangmeng Zuo, and David Zhang. 2016 · 2016
Cited alongside, same era.
Unsupervised cross-dataset transfer learning for person re-identification
Peixi Peng, Tao Xiang, Yaowei Wang, Massimiliano Pontil, Shaogang Gong, Tiejun Huang, and Yonghong Tian. 2016 · 2016
Cited alongside, same era.
Channel pruning for accelerating very deep neural networks
Yihui He, Xiangyu Zhang, and Jian Sun. 2017 · 2017
Cited alongside, same era.
Zero-shot object detection
Ankan Bansal, Karan Sikka, Gaurav Sharma, Rama Chellappa, and Ajay Divakaran. 2018 · 2018
Cited alongside, same era.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
Model compression via distillation and quantization
Antonio Polino, Razvan Pascanu, and Dan Alistarh. 2018 · 2018
Later among the works it cites.
Improving language understanding by generative pre-training
Alec Radford. 2018 · 2018
Later among the works it cites.