Fetching the paper…
Reading the bibliography…
Unsupervised part of speech (POS) tagging is often framed as a clustering problem, but practical taggers need to \textit{ground} their clusters as well.
Cross-lingual name tagging and linking for 282 languages
Xiaoman Pan, Boliang Zhang, Jonathan May, Joel Nothman, Kevin Knight, and Heng Ji. 2017 · 1958
Earlier work this paper cites.
Class-based n
Peter F. Brown, Vincent J. Della Pietra, Peter V. deSouza, Jenifer C. Lai, and Robert L. Mercer. 1992 · 1992
Earlier work this paper cites.
Tagging English text with a probabilistic model
Bernard Merialdo. 1994 · 1994
Earlier work this paper cites.
Carmel finite-state toolkit
Jonathan Graehl. 1997 · 1997
Earlier work this paper cites.
Algorithms for bigram and trigram word clustering
Sven Martin, Jörg Liermann, and Hermann Ney. 1998 · 1998
Earlier work this paper cites.
SRILM-an extensible language modeling toolkit
Andreas Stolcke. 2002 · 2002
Earlier work this paper cites.
Combining distributional and morphological information for part of speech induction
Alexander Clark. 2003 · 2003
Earlier work this paper cites.
Contrastive estimation: Training log-linear models on unlabeled data
Noah A. Smith and Jason Eisner. 2005 · 2005
Earlier work this paper cites.
Prototype-driven learning for sequence models
Aria Haghighi and Dan Klein. 2006 · 2006
Earlier work this paper cites.
A fully Bayesian approach to unsupervised part-of-speech tagging
Sharon Goldwater and Tom Griffiths. 2007 · 2007
Earlier work this paper cites.
Why doesn’t EM find good HMM POS-taggers?
Mark Johnson. 2007 · 2007
Earlier work this paper cites.
EM can find pretty good HMM POS-taggers (when given a good start)
Yoav Goldberg, Meni Adler, and Michael Elhadad. 2008 · 2008
Earlier work this paper cites.
Evaluating unsupervised part-of-speech tagging for grammar induction
William P. Headden III, David McClosky, and Eugene Charniak. 2008 · 2008
Earlier work this paper cites.
Posterior vs parameter sparsity in latent variable models
Kuzman Ganchev, Ben Taskar, Fernando Pereira, and Joao Graca. 2009 · 2009
Earlier work this paper cites.
Minimized models for unsupervised part-of-speech tagging
Sujith Ravi and Kevin Knight. 2009 · 2009
Earlier work this paper cites.
Painless unsupervised learning with features
Taylor Berg-Kirkpatrick, Alexandre Bouchard-Côté, John DeNero, and Dan Klein. 2010 · 2010
Earlier work this paper cites.
Two decades of unsupervised POS induction: How far have we come?
Christos Christodoulopoulos, Sharon Goldwater, and Mark Steedman. 2010 · 2010
Cited alongside, same era.
A statistical model for lost language decipherment
Benjamin Snyder, Regina Barzilay, and Kevin Knight. 2010 · 2010
Cited alongside, same era.
Unsupervised part-of-speech tagging with bilingual graph-based projections
Dipanjan Das and Slav Petrov. 2011 · 2011
Cited alongside, same era.
Crisis MT: Developing a cookbook for MT in crisis situations
Will Lewis, Robert Munro, and Stephan Vogel. 2011 · 2011
Cited alongside, same era.
Deciphering foreign language
Sujith Ravi and Kevin Knight. 2011 · 2011
Cited alongside, same era.
Wiki-ly supervised part-of-speech tagging
Shen Li, João Graça, and Ben Taskar. 2012 · 2012
Cited alongside, same era.
Many languages, one parser
Waleed Ammar, George Mulcaire, Miguel Ballesteros, Chris Dyer, and Noah A Smith. 2016 · 2016
Later among the works it cites.
Named entity recognition with bidirectional LSTM-CNNs
Jason Chiu and Eric Nichols. 2016 · 2016
Later among the works it cites.
A representation learning framework for multi-source transfer parsing
Jiang Guo, Wanxiang Che, David Yarowsky, Haifeng Wang, and Ting Liu. 2016 · 2016
Later among the works it cites.
Fasttext.zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski, Matthijs Douze, Hervé Jégou, and Tomas Mikolov. 2016 · 2016
Later among the works it cites.
Very-large scale parsing and normalization of wiktionary morphological paradigms
Christo Kirov, John Sylak-Glassman, Roger Que, and David Yarowsky. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Short message communications: users, topics, and in-language processing
Robert Munro and Christopher Manning. 2012 · 2012
Cited alongside, same era.
Universal dependency annotation for multilingual parsing
Ryan McDonald, Joakim Nivre, Yvonne Quirmbach-Brundage, Yoav Goldberg, Dipanjan Das, Kuzman Ganchev, Keith Hall, Slav Petrov, Hao Zhang, Oscar Täckström, Claudia Bedini, Núria Bertomeu Castelló, and Jungmee Lee. 2013 · 2013
Cited alongside, same era.
Token and type constraints for cross-lingual part-of-speech tagging
Oscar Täckström, Dipanjan Das, Slav Petrov, Ryan McDonald, and Joakim Nivre. 2013 · 2013
Cited alongside, same era.
What can we get from 1000 tokens? a case study of multilingual POS tagging for resource-poor languages
Long Duong, Trevor Cohn, Karin Verspoor, Steven Bird, and Paul Cook. 2014 · 2014
Cited alongside, same era.
Unifying bayesian inference and vector space models for improved decipherment
Qing Dou, Ashish Vaswani, Kevin Knight, and Chris Dyer. 2015 · 2015
Cited alongside, same era.
Transition-based dependency parsing with stack long short-term memory
Chris Dyer, Miguel Ballesteros, Wang Ling, Austin Matthews, and Noah A. Smith. 2015 · 2015
Cited alongside, same era.
Karl Stratos, Michael Collins, and Daniel Hsu. 2016 · 2016
Later among the works it cites.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Later among the works it cites.
Deep biaffine attention for neural dependency parsing
Timothy Dozat and Christopher D Manning. 2017 · 2017
Later among the works it cites.
Model transfer for tagging low-resource languages using a bilingual dictionary
Meng Fang and Trevor Cohn. 2017 · 2017
Later among the works it cites.
Deciphering related languages
Nima Pourdamghani and Kevin Knight. 2017 · 2017
Later among the works it cites.
Tokenizing, POS tagging, lemmatizing and parsing UD 2.0 with UDPipe
Milan Straka and Jana Straková. 2017 · 2017
Later among the works it cites.
CoNLL 2017 shared task: Multilingual parsing from raw text to universal dependencies
Daniel Zeman, Martin Popel, Milan Straka, Jan Hajic, Joakim Nivre, Filip Ginter, Juhani Luotolahti, Sampo Pyysalo, Slav Petrov, Martin Potthast, Francis Tyers, Elena Badmaeva, Memduh Gokirmak, Anna Nedoluzhko, Silvie Cinkova, Jan Hajic jr., Jaroslava Hlavacova, Václava Kettnerová, Zdenka Uresova, Jenna Kanerva, Stina Ojala, Anna Missilä, Christopher D. Manning, Sebastian Schuster, Siva Reddy, Dima Taji, Nizar Habash, Herman Leung, Marie-Catherine de Marneffe, Manuela Sanguinetti, Maria Simi, Hiroshi Kanayama, Valeria dePaiva, Kira Droganova, Héctor Martínez Alonso, Çağrı Çöltekin, Umut Sulubacak, Hans Uszkoreit, Vivien Macketanz, Aljoscha Burchardt, Kim Harris, Katrin Marheinecke, Georg Rehm, Tolga Kayadelen, Mohammed Attia, Ali Elkahky, Zhuoran Yu, Emily Pitler, Saran Lertpradit, Michael Mandl, Jesse Kirchner, Hector Fernandez Alcalde, Jana Strnadová, Esha Banerjee, Ruli Manurung, Antonio Stella, Atsuko Shimada, Sookyoung Kwak, Gustavo Mendonca, Tatiana Lando, Rattima Nitisaroj, and Josie Li. 2017 · 2017
Later among the works it cites.
Unsupervised neural machine translation
Mikel Artetxe, Gorka Labaka, Eneko Agirre, and Kyunghyun Cho. 2018 · 2018
Later among the works it cites.
Out-of-the-box universal romanization tool uroman
Ulf Hermjakob, Jonathan May, and Kevin Knight. 2018 · 2018
Later among the works it cites.
CoNLL 2018 shared task: Multilingual parsing from raw text to universal dependencies
Daniel Zeman, Jan Hajič, Martin Popel, Martin Potthast, Milan Straka, Filip Ginter, Joakim Nivre, and Slav Petrov. 2018 · 2018
Later among the works it cites.