Fetching the paper…
Reading the bibliography…
The goal of this paper is to investigate the connection between the performance gain that can be obtained by selftraining and the similarity between the corpora used in this approach.
S. Kullback and R. A. Leibler, “On information and sufficiency,” The Annals of Mathematical Statistics , vol. 22, no. 1, pp. 79–86, 1951
1951
Earlier work this paper cites.
A. Rényi, “On measures of information and entropy,” in Proceedings of the 4 t h 4^{th} Berkeley Symposium on Mathematics, Statistics and Probability , vol. 1. Berkeley, California, USA: University of California Press, 1961, pp. 547–561
1961
Earlier work this paper cites.
D. Biber, Variation across speech and writing . Cambridge, UK: Cambridge University Press, 1988
1988
Earlier work this paper cites.
E. W. Noreen, Computer-intensive methods for testing hypotheses . New York, NY, USA: John Wiley, 1989
1989
Earlier work this paper cites.
J. Lin, “Divergence measures based on the Shannon entropy,” IEEE Transactions on Information Theory , vol. 37, no. 1, pp. 145–151, 1991
1991
Earlier work this paper cites.
E. Charniak, “Statistical parsing with a context-free grammar and word statistics,” in Proceedings of the Fourteenth National Conference on Artificial Intelligence and Ninth Innovative Applications of Artificial Intelligence Conference . Rhode Island, USA: MIT Press, 1997, pp. 598–603
1997
Earlier work this paper cites.
S. Della Pietra, V. Della Pietra, and J. Lafferty, “Inducing features of random fields,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 19, no. 4, pp. 380–393, 1997
1997
Earlier work this paper cites.
L. Lee, “Measures of distributional similarity,” in Proceedings of the 37th Annual Meeting of the Association for Computational Linguistics . Maryland, USA: Association for Computational Linguistics, 1999, pp. 25–32
1999
Earlier work this paper cites.
T. Joachims, “Making large-scale support vector machine learning practical,” in Advances in kernel methods: support vector learning . Cambridge, MA, USA: MIT Press, 1999, pp. 169–184
1999
Earlier work this paper cites.
A. Yeh, “More accurate tests for the statistical significance of result differences,” in Proceedings of the 18th International Conference on Computational Linguistics , vol. 2. Saarbrücken, Germany: Association for Computational Linguistics, 2000, pp. 947–953
2000
Earlier work this paper cites.
D. Y. W. Lee, “Genres, registers, text types, domain, and styles: Clarifying the concepts and navigating a path through the BNC jungle,” Language Learning & Technology , vol. 5, no. 3, pp. 37–72, 2001
2001
Earlier work this paper cites.
P. Mitra, C. Murthy, and S. K. Pal, “Unsupervised feature selection using feature similarity,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 24, pp. 301–312, 2002
2002
Earlier work this paper cites.
J. Gao, J. Goodman, M. Li, and K.-F. Lee, “Toward a unified approach to statistical language modeling for chinese,” Transactions on Asian Language Information Processing , vol. 1, no. 1, pp. 3–33, 2002
2002
Earlier work this paper cites.
W. Yuan, J. Gao, and H. Suzuki, “An empirical study on language model adaptation using a metric of domain similarity,” in Natural Language Processing Ð IJCNLP 2005 , ser. Lecture Notes in Computer Science, R. Dale, K.-F. Wong, J. Su, and O. Kwong, Eds. Berlin Heidelberg: Springer, 2005, vol. 3651, pp. 957–968
2005
Cited alongside, same era.
H. Daumé III and D. Marcu, “Domain adaptation for statistical classifiers,” Journal of Artificial Intelligence Research , vol. 26, pp. 101–126, 2006
2006
Cited alongside, same era.
J. Jiang and C. Zhai, “Instance weighting for domain adaptation in NLP,” in Proceedings of the 45th Annual Meeting of the Association of Computational Linguistics . Prague, Czech Republic: Association for Computational Linguistics, June 2007, pp. 264–271
2007
Cited alongside, same era.
J. Blitzer, M. Dredze, and F. Pereira, “Biographies, Bollywood, Boom-boxes and Blenders: Domain adaptation for sentiment classification,” in Proceedings of the 45 t h 45^{th} Annual Meeting of the Association of Computational Linguistics . Prague, Czech Republic: Association for Computational Linguistics, 2007, pp. 440–447
R. C. Moore and W. Lewis, “Intelligent selection of language model training data,” in Proceedings of the ACL 2010 Conference Short Papers . Uppsala, Sweden: Association for Computational Linguistics, 2010, pp. 220–224
2010
Later among the works it cites.
V. Van Asch and W. Daelemans, “Using domain similarity for performance estimation,” in Proceedings of the 2010 Workshop on Domain Adaptation for Natural Language Processing . Uppsala, Sweden: Association for Computational Linguistics, July 2010, pp. 31–36
2010
Later among the works it cites.
2010
Later among the works it cites.
J. Sinclair and J. Ball, “Preliminary recommendations on text typology,” Consiglio Nazionale delle Ricerche, Istituto di Linguistica Computazionale, Pisa, Italy, Expert Advisory Group on Language Engineering Standards (EAGLES) EAG—TCWG—TTYP/P, 1996, www.ilc.cnr.it/EAGLES96/texttyp/ texttyp.html (Last accessed: June 2011)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2007
Cited alongside, same era.
S. Ravi, K. Knight, and R. Soricut, “Automatic prediction of parser accuracy,” in Proceedings of the 2008 Conference on Empirical Methods in Natural Language Processing . Honolulu, Hawaii: Association for Computational Linguistics, October 2008, pp. 887–896
2008
Cited alongside, same era.
K. Verspoor, K. B. Cohen, and L. Hunter, “The textual characteristics of traditional and open access scientific journals are similar,” BMC Bioinformatics , vol. 10, pp. 1–16, 2009
2009
Cited alongside, same era.
B. Chen, W. Lam, I. Tsang, and T.-L. Wong, “Extracting discriminative concepts for domain adaptation in text mining,” in Proceedings of the 15th ACM SIGKDD international conference on Knowledge discovery and data mining , ser. KDD ’09. Paris, France: ACM, 2009, pp. 179–188
2009
Cited alongside, same era.
Y. Mansour, M. Mohri, and A. Rostamizadeh, “Multiple source adaptation and the Rényi divergence,” in Proceedings of the Twenty-Fifth Conference on Uncertainty in Artificial Intelligence . Montreal, Quebec, Canada: AUAI Press, 2009, pp. 367–374
2009
Cited alongside, same era.
Y. Zhang and R. Wang, “Correlating natural language parser performance with statistical measures of the text,” in Proceedings of the 32nd annual German conference on Advances in artificial intelligence . Paderborn, Germany: Springer-Verlag, 2009, pp. 217–224
2009
Cited alongside, same era.
K. Sagae, “Self-training without reranking for parser domain adaptation and its impact on semantic role labeling,” in Proceedings of the 2010 Workshop on Domain Adaptation for Natural Language Processing . Uppsala, Sweden: Association for Computational Linguistics, July 2010, pp. 37–44
2010
Cited alongside, same era.
D. McClosky, “Any domain parsing: Automatic domain adaptation for natural language parsing,” Ph.D. dissertation, Department of Computer Science, Brown University, Rhode Island, USA, 2010
2010
Cited alongside, same era.
D. Biber and B. Gray, “Challenging stereotypes about academic writing: Complexity, elaboration, explicitness,” Journal of English for Academic Purposes , vol. 9, no. 1, pp. 2–20, 2010
2010
Cited alongside, same era.
2011
Later among the works it cites.
C. Dong and U. Schäfer, “Ensemble-style self-training on citation classification,” in Proceedings of 5th International Joint Conference on Natural Language Processing . Chiang Mai, Thailand: Asian Federation of Natural Language Processing, November 2011, pp. 623–631
2011
Later among the works it cites.
B. Plank, “Domain adaptation for parsing,” Ph.D. dissertation, University of Groningen, the Netherlands, 2011, groningen Dissertations in Linguistics 96
2011
Later among the works it cites.
N. Ponomareva and M. Thelwall, “Biographies or blenders: Which resource is best for cross-domain sentiment analysis?” in Computational Linguistics and Intelligent Text Processing , ser. Lecture Notes in Computer Science, A. Gelbukh, Ed. Berlin Heidelberg: Springer, 2012, vol. 7181, pp. 488–499
2012
Later among the works it cites.
R. Remus, “Domain adaptation using domain similarity- and domain complexity-based instance selection for cross-domain sentiment analysis,” in Proceedings of the IEEE 12th International Conference on Data Mining Workshops . Brussels, Belgium: IEEE, 2012, pp. 717–723
2012
Later among the works it cites.
Z. Liu, X. Dong, Y. Guan, and J. Yang, “Reserved self-training: A semi-supervised sentiment classification method for chinese microblogs,” in Proceedings of the Sixth International Joint Conference on Natural Language Processing . Nagoya, Japan: Asian Federation of Natural Language Processing, October 2013, pp. 455–462
2013
Later among the works it cites.
L. Lee, “On the effectiveness of the skew divergence for statistical language analysis,” in 8 t h 8^{th} International Workshop on Artificial Intelligence and Statistics (AISTATS 2001) . Florida, USA: AISTATS, 2001, pp. 65–72, online repository http://www.gatsby.ucl.ac.uk/aistats/aistats2001 (Last accessed: March 2013)
2013
Later among the works it cites.
E. Agirre, D. Cer, M. Diab, A. Gonzalez-Agirre, and W. Guo, “*SEM 2013 shared task: Semantic textual similarity,” in Second Joint Conference on Lexical and Computational Semantics, Volume 1: Proceedings of the Main Conference and the Shared Task: Semantic Textual Similarity . Atlanta, Georgia, USA: Association for Computational Linguistics, June 2013, pp. 32–43
2013
Later among the works it cites.
D. Bär, T. Zesch, and I. Gurevych, “DKPro similarity: An open source framework for text similarity,” in Proceedings of the 51st Annual Meeting of the Association for Computational Linguistics: System Demonstrations . Sofia, Bulgaria: Association for Computational Linguistics, August 2013, pp. 121–126
2013
Later among the works it cites.