Fetching the paper…
Reading the bibliography…
Today when many practitioners run basic NLP on the entire web and large-volume traffic, faster methods are paramount to saving time and energy costs.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and J urgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
The vanishing gradient problem during learning recurrent neural nets and problem solutions
Sepp Hochreiter. 1998 · 1998
Earlier work this paper cites.
Text chunking using transformation-based learning
Lance A Ramshaw and Mitchell P Marcus. 1999 · 1999
Earlier work this paper cites.
Conditional random fields: Probabilistic models for segmenting and labeling sequence data
John D. Lafferty, Andrew McCallum, and Fernando C. N. Pereira. 2001 · 2001
Earlier work this paper cites.
Named entity recognition with long short-term memory
James Hammerton. 2003 · 2003
Earlier work this paper cites.
Introduction to the CoNLL-2003 shared task: Language-independent named entity recognition
Erik F Tjong Kim Sang and Fien De Meulder. 2003 · 2003
Earlier work this paper cites.
Collective information extraction with relational markov networks
Razvan Bunescu and Raymond J. Mooney. 2004 · 2004
Earlier work this paper cites.
Collective segmentation and labeling of distant entities in information extraction
Charles Sutton and Andrew McCallum. 2004 · 2004
Earlier work this paper cites.
Incorporating non-local information into information extraction systems by gibbs sampling
Jenny Rose Finkel, Trond Grenager, and Christopher Manning. 2005 · 2005
Earlier work this paper cites.
Ontonotes: the 90% solution
Eduard Hovy, Mitchell Marcus, Martha Palmer, Lance Ramshaw, and Ralph Weischedel. 2006 · 2006
Earlier work this paper cites.
Towards robust linguistic analysis using ontonotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Hwee Tou Ng, Anders Bj orkelund, Olga Uryupina, Yuchen Zhang, and Zhi Zhong. 2006 · 2006
Earlier work this paper cites.
Structure compilation: trading structure for features
Percy Liang, Hal Daumé III, and Dan Klein. 2008 · 2008
Earlier work this paper cites.
Search-based structured prediction
Hal Daumé III, John Langford, and Daniel Marcu. 2009 · 2009
Earlier work this paper cites.
Design challenges and misconceptions in named entity recognition
Lev Ratinov and Dan Roth. 2009 · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio. 2010 · 2010
Earlier work this paper cites.
Word representations: a simple and general method for semi-supervised learning
Joseph Turian, Lev Ratinov, and Yoshua Bengio. 2010 · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Ronan Collobert, Jason Weston, Léon Bottou, Michael Karlen, Koray Kavukcuoglu, and Pavel Kuksa. 2011 · 2011
Cited alongside, same era.
Deep sparse rectifier neural networks
Xavier Glorot, Antoine Bordes, and Yoshua Bengio. 2011 · 2011
Cited alongside, same era.
Conll-2012 shared task: Modeling multilingual unrestricted coreference in ontonotes
Sameer Pradhan, Alessandro Moschitti, Nianwen Xue, Olga Uryupina, and Yuchen Zhang. 2012 · 2012
Cited alongside, same era.
Not all contexts are created equal: Better word representations with variable attention
Wang Ling, Lin Chu-Cheng, Yulia Tsvetkov, and Silvio Amir. 2013 · 2013
Cited alongside, same era.
A joint model for entity analysis: Coreference, typing and linking
Greg Durrett and Dan Klein. 2014 · 2014
Cited alongside, same era.
A convolutional neural network for modelling sentences
Nal Kalchbrenner, Edward Grefenstette, and Phil Blunsom. 2014 · 2014
Molding cnns for text: non-linear, non-consecutive convolutions
Tao Lei, Regina Barzilay, and Tommi Jaakkola. 2015 · 2015
Later among the works it cites.
Finding Function in Form: Compositional Character Models for Open Vocabulary Word Representation
Wang Ling, Tiago Luís, Luís Marujo, Ramón Fernandez Astudillo, Silvio Amir, Chris Dyer, Alan W Black, and Isabel Trancoso. 2015 · 2015
Later among the works it cites.
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich. 2015 · 2015
Later among the works it cites.
Representing text for joint embedding of text and knowledge bases
Kristina Toutanova, Danqi Chen, Patrick Pantel, Hoifung Poon, Pallavi Choudhury, and Michael Gamon. 2015 · 2015
Later among the works it cites.
Structured training for neural network transition-based parsing
David Weiss, Chris Alberti, Michael Collins, and Slav Petrov. 2015 · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Convolutional neural networks for sentence classification
Yoon Kim. 2014 · 2014
Cited alongside, same era.
Lexicon infused phrase embeddings for named entity resolution
Alexandre Passos, Vineet Kumar, and Andrew McCallum. 2014 · 2014
Cited alongside, same era.
Fitnets: Hints for thin deep nets
Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou, Antoine Chassang, Carlo Gatta, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Tensorflow: Large-scale machine learning on heterogeneous systems, 2015
Martın Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, et al. 2015 · 2015
Cited alongside, same era.
Semantic image segmentation with deep convolutional nets and fully connected crfs
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L. Yuille. 2015 · 2015
Cited alongside, same era.
Semi-supervised sequence learning
Andrew M. Dai and Quoc V. Le. 2015 · 2015
Cited alongside, same era.
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015 · 2015
Later among the works it cites.
Named entity recognition with bidirectional lstm-cnns
Jason PC Chiu and Eric Nichols. 2016 · 2016
Later among the works it cites.
Knowledge matters: Importance of prior information for optimization
Çalar Gülçehre and Yoshua Bengio. 2016 · 2016
Later among the works it cites.
Neural machine translation in linear time
Nal Kalchbrenner, Lasse Espeholt, Karen Simonyan, Aaron van den Oord, Alex Graves, and Koray Kavukcuoglu. 2016 · 2016
Later among the works it cites.
Neural architectures for named entity recognition
Guillaume Lample, Miguel Ballesteros, Sandeep Subramanian, Kazuya Kawakami, and Chris Dyer. 2016 · 2016
Later among the works it cites.
Stability and generalization in structured prediction
Ben London, Bert Huang, and Lise Getoor. 2016 · 2016
Later among the works it cites.
End-to-end sequence labeling via bi-directional lstm-cnns-crf
Xuezhe Ma and Eduard Hovy. 2016 · 2016
Later among the works it cites.
Wavenet: A generative model for raw audio
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu. 2016 · 2016
Later among the works it cites.
Multi-task cross-lingual sequence tagging from scratch
Zhilin Yang, Ruslan Salakhutdinov, and William Cohen. 2016 · 2016
Later among the works it cites.
Multi-scale context aggregation by dilated convolutions
Fisher Yu and Vladlen Koltun. 2016 · 2016
Later among the works it cites.
Dropout with expectation-linear regularization
Xuezhe Ma, Yingkai Gaom, Zhiting Hu, Yaoliang Yu, Yuntian Deng, and Eduard Hovy. 2017 · 2017
Closest in time.