Fetching the paper…
Reading the bibliography…
We address part-of-speech (POS) induction by maximizing the mutual information between the induced label and its context.
Statistical inference for probabilistic functions of finite state Markov chains
Leonard E. Baum and Ted Petrie. 1966 · 1966
Earlier work this paper cites.
Word association norms, mutual information, and lexicography
Kenneth Ward Church and Patrick Hanks. 1990 · 1990
Earlier work this paper cites.
Class-based n n -gram models of natural language
Peter F. Brown, Peter V. Desouza, Robert L. Mercer, Vincent J. Della Pietra, and Jenifer C. Lai. 1992 · 1992
Earlier work this paper cites.
Tagging English text with a probabilistic model
Bernard Merialdo. 1994 · 1994
Earlier work this paper cites.
The information bottleneck method
Naftali Tishby, Fernando C Pereira, and William Bialek. 2000 · 2000
Earlier work this paper cites.
V-measure: A conditional entropy-based external cluster evaluation measure
Andrew Rosenberg and Julia Hirschberg. 2007 · 2007
Earlier work this paper cites.
Simple semi-supervised dependency parsing
Terry Koo, Xavier Carreras, and Michael Collins. 2008 · 2008
Earlier work this paper cites.
Painless unsupervised learning with features
Taylor Berg-Kirkpatrick, Alexandre Bouchard-Côté, John DeNero, and Dan Klein. 2010 · 2010
Earlier work this paper cites.
Two decades of unsupervised POS induction: How far have we come?
Christos Christodoulopoulos, Sharon Goldwater, and Mark Steedman. 2010 · 2010
Earlier work this paper cites.
Deep canonical correlation analysis
Galen Andrew, Raman Arora, Jeff Bilmes, and Karen Livescu. 2013 · 2013
Cited alongside, same era.
Universal dependency annotation for multilingual parsing
Ryan T. McDonald, Joakim Nivre, Yvonne Quirmbach-Brundage, Yoav Goldberg, Dipanjan Das, Kuzman Ganchev, Keith B. Hall, Slav Petrov, Hao Zhang, Oscar Täckström, Claudia Bedini, Núria B. Castelló, and Jungmee Lee. 2013 · 2013
Cited alongside, same era.
Improved part-of-speech tagging for online conversational text with word clusters
Olutobi Owoputi, Brendan O’Connor, Chris Dyer, Kevin Gimpel, Nathan Schneider, and Noah A Smith. 2013 · 2013
Cited alongside, same era.
Conditional random field autoencoders for unsupervised structured prediction
Waleed Ammar, Chris Dyer, and Noah A Smith. 2014 · 2014
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Stochastic optimization for deep cca via nonlinear orthogonal iterations
Weiran Wang, Raman Arora, Karen Livescu, and Nathan Srebro. 2015 · 2015
Later among the works it cites.
Unsupervised part-of-speech tagging with anchor hidden markov models
Karl Stratos, Michael Collins, and Daniel Hsu. 2016 · 2016
Later among the works it cites.
Unsupervised neural hidden markov models
Ke M. Tran, Yonatan Bisk, Ashish Vaswani, Daniel Marcu, and Kevin Knight. 2016 · 2016
Later among the works it cites.
Inter-annotator agreement
Ron Artstein. 2017 · 2017
Later among the works it cites.
Mutual information neural estimation
Mohamed Ishmael Belghazi, Aristide Baratin, Sai Rajeshwar, Sherjil Ozair, Yoshua Bengio, Aaron Courville, and Devon Hjelm. 2018 · 2018
Closest in time.
Unsupervised learning of syntactic structure with invertible neural projections
Junxian He, Graham Neubig, and Taylor Berg-Kirkpatrick. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Justin B Kinney and Gurinder S Atwal. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D Manning. 2014 · 2014
Cited alongside, same era.
Unsupervised pos induction with word embeddings
Chu-Cheng Lin, Waleed Ammar, Chris Dyer, and Lori Levin. 2015 · 2015
Cited alongside, same era.
Deep learning and the information bottleneck principle
Naftali Tishby and Noga Zaslavsky. 2015 · 2015
Cited alongside, same era.
Closest in time.
Learning deep representations by mutual information estimation and maximization
R Devon Hjelm, Alex Fedorov, Samuel Lavoie-Marchildon, Karan Grewal, Adam Trischler, and Yoshua Bengio. 2018 · 2018
Closest in time.
Information theoretic co-training
David McAllester. 2018 · 2018
Closest in time.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Closest in time.