Fetching the paper…
Reading the bibliography…
Inductive transfer learning has greatly impacted computer vision, but existing approaches in NLP still require task-specific modifications and training from scratch.
Estimation of dependences based on empirical data
Vladimir Naumovich Vapnik and Samuel Kotz. 1982 · 1982
Earlier work this paper cites.
Multitask learning: A knowledge-based source of inductive bias
Rich Caruana. 1993 · 1993
Earlier work this paper cites.
The trec-8 question answering track evaluation
Ellen M Voorhees and Dawn M Tice. 1999 · 1999
Earlier work this paper cites.
A Model of Inductive Bias Learning
Jonathan Baxter. 2000 · 2000
Earlier work this paper cites.
Review spam detection
Nitin Jindal and Bing Liu. 2007 · 2007
Earlier work this paper cites.
Deep boltzmann machines
Ruslan Salakhutdinov and Geoffrey Hinton. 2009 · 2009
Earlier work this paper cites.
Why does unsupervised pre-training help deep learning?
Dumitru Erhan, Yoshua Bengio, Aaron Courville, Pierre-Antoine Manzagol, Pascal Vincent, and Samy Bengio. 2010 · 2010
Earlier work this paper cites.
A survey on transfer learning
Sinno Jialin Pan and Qiang Yang. 2010 · 2010
Earlier work this paper cites.
Document categorization in legal electronic discovery: computer classification vs. manual review
Herbert L Roitblat, Anne Kershaw, and Patrick Oot. 2010 · 2010
Earlier work this paper cites.
Biographies, bollywood, boom-boxes and blenders: Domain adaptation for sentiment classification
John Blitzer, Mark Dredze, and Fernando Pereira. 2007 · 2011
Earlier work this paper cites.
Classifying text messages for the haiti earthquake
Cornelia Caragea, Nathan McNeese, Anuj Jaiswal, Greg Traylor, Hyun-Woo Kim, Prasenjit Mitra, Dinghao Wu, Andrea H Tapia, Lee Giles, Bernard J Jansen, et al. 2011 · 2011
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Andrew L Maas, Raymond E Daly, Peter T Pham, Dan Huang, Andrew Y Ng, and Christopher Potts. 2011 · 2011
Earlier work this paper cites.
The application of data mining techniques in financial fraud detection: A classification framework and an academic review of literature
EWT Ngai, Yong Hu, YH Wong, Yijun Chen, and Xin Sun. 2011 · 2011
Earlier work this paper cites.
Detecting automation of twitter accounts: Are you a human, bot, or cyborg?
Zi Chu, Steven Gianvecchio, Haining Wang, and Sushil Jajodia. 2012 · 2012
Earlier work this paper cites.
Distributed Representations of Words and Phrases and their Compositionality
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Decaf: A deep convolutional activation feature for generic visual recognition
Jeff Donahue, Yangqing Jia, Oriol Vinyals, Judy Hoffman, Ning Zhang, Eric Tzeng, and Trevor Darrell. 2014 · 2014
Earlier work this paper cites.
Cnn features off-the-shelf: an astounding baseline for recognition
Ali Sharif Razavian, Hossein Azizpour, Josephine Sullivan, and Stefan Carlsson. 2014 · 2014
Earlier work this paper cites.
How transferable are features in deep neural networks?
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson. 2014 · 2014
Earlier work this paper cites.
Semi-supervised Sequence Learning
Andrew M. Dai and Quoc V. Le. 2015 · 2015
Cited alongside, same era.
Hypercolumns for object segmentation and fine-grained localization
Bharath Hariharan, Pablo Arbeláez, Ross Girshick, and Jitendra Malik. 2015 · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy. 2015 · 2015
Cited alongside, same era.
Discriminative neural sentence modeling by tree-based convolution
Lili Mou, Hao Peng, Ge Li, Yan Xu, Lu Zhang, and Zhi Jin. 2015 · 2015
Cited alongside, same era.
Improving neural machine translation models with monolingual data
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2015 · 2015
Cited alongside, same era.
Deep Biaffine Attention for Neural Dependency Parsing
Timothy Dozat and Christopher D. Manning. 2017 · 2017
Later among the works it cites.
Using millions of emoji occurrences to learn any-domain representations for detecting sentiment, emotion and sarcasm
Bjarke Felbo, Alan Mislove, Anders Søgaard, Iyad Rahwan, and Sune Lehmann. 2017 · 2017
Later among the works it cites.
Densely Connected Convolutional Networks
Gao Huang, Zhuang Liu, Kilian Q. Weinberger, and Laurens van der Maaten. 2017 · 2017
Later among the works it cites.
Deep pyramid convolutional neural networks for text categorization
Rie Johnson and Tong Zhang. 2017 · 2017
Later among the works it cites.
SGDR: Stochastic Gradient Descent with Warm Restarts
Ilya Loshchilov and Frank Hutter. 2017 · 2017
Later among the works it cites.
Learned in Translation: Contextualized Word Vectors
Bryan McCann, James Bradbury, Caiming Xiong, and Richard Socher. 2017 · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
UNITN: Training Deep Convolutional Neural Network for Twitter Sentiment Classification
Aliaksei Severyn and Alessandro Moschitti. 2015 · 2015
Cited alongside, same era.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015 · 2015
Cited alongside, same era.
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
What makes ImageNet good for transfer learning?
Minyoung Huh, Pulkit Agrawal, and Alexei A Efros. 2016 · 2016
Cited alongside, same era.
Supervised and semi-supervised text categorization using lstm for region embeddings
Rie Johnson and Tong Zhang. 2016 · 2016
Cited alongside, same era.
Assessing the ability of lstms to learn syntax-sensitive dependencies
Tal Linzen, Emmanuel Dupoux, and Yoav Goldberg. 2016 · 2016
Cited alongside, same era.
Adversarial training methods for semi-supervised text classification
Takeru Miyato, Andrew M Dai, and Ian Goodfellow. 2016 · 2016
Cited alongside, same era.
Later among the works it cites.
Pointer Sentinel Mixture Models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher. 2017b · 2017
Later among the works it cites.
Question Answering through Transfer Learning from Large Fine-grained Supervision Data
Sewon Min, Minjoon Seo, and Hannaneh Hajishirzi. 2017 · 2017
Later among the works it cites.
Semi-supervised sequence tagging with bidirectional language models
Matthew E Peters, Waleed Ammar, Chandra Bhagavatula, and Russell Power. 2017 · 2017
Later among the works it cites.
Learning to generate reviews and discovering sentiment
Alec Radford, Rafal Jozefowicz, and Ilya Sutskever. 2017 · 2017
Later among the works it cites.
Semi-supervised multitask learning for sequence labeling
Marek Rei. 2017 · 2017
Later among the works it cites.
Cyclical learning rates for training neural networks
Leslie N Smith. 2017 · 2017
Later among the works it cites.
Revisiting Recurrent Networks for Paraphrastic Sentence Embeddings
John Wieting and Kevin Gimpel. 2017 · 2017
Later among the works it cites.
Colorless green recurrent networks dream hierarchically
Kristina Gulordava, Piotr Bojanowski, Edouard Grave, Tal Linzen, and Marco Baroni. 2018 · 2018
Closest in time.
Empower sequence labeling with task-aware neural language model
Liyuan Liu, Jingbo Shang, Frank Xu, Xiang Ren, Huan Gui, Jian Peng, and Jiawei Han. 2018 · 2018
Closest in time.
Exploring the Limits of Weakly Supervised Pretraining
Dhruv Mahajan, Ross Girshick, Vignesh Ramanathan, Kaiming He, Manohar Paluri, Yixuan Li, Ashwin Bharambe, and Laurens van der Maaten. 2018 · 2018
Closest in time.
Deep contextualized word representations
Matthew E Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Closest in time.