Fetching the paper…
Reading the bibliography…
The Moore-Lewis method of "intelligent selection of language model training data" is very effective, cheap, efficient...
A Universal Prior for Integers and Estimation by Minimum Description Length
Rissanen, J. (1983) · 1983
Earlier work this paper cites.
Scalable Backoff Language Models
Seymore, K. and Rosenfeld, R. (1996) · 1996
Earlier work this paper cites.
Entropy-Based Pruning of Backoff Language Models
Stolcke, A. (1998) · 1998
Earlier work this paper cites.
Comparing Corpora
Kilgarriff, A. (2001) · 2001
Earlier work this paper cites.
Growing an N-gram Language Model
Siivola, V. and Pellom, B. L. (2005) · 2005
Cited alongside, same era.
Intelligent Selection of Language Model Training Data
Moore, R. C. and Lewis, W. D. (2010) · 2010
Cited alongside, same era.
Domain Adaptation Via Pseudo In-Domain Data Selection
Axelrod, A., He, X., and Gao, J. (2011) · 2011
Cited alongside, same era.
Data Selection for Statistical Machine Translation
Axelrod, A. (2014) · 2014
Cited alongside, same era.
Selecting Relevant Text Subsets from Web-Data for Building Topic Specific Language Models
Sethy, A., Georgiou, P. G., and Narayanan, S. (2006a)
Cited in the paper.
Text Data Acquisition for Domain-Specific Language Models
Sethy, A., Georgiou, P. G., and Narayanan, S. (2006b)
Cited in the paper.
Submodularity for Data Selection in Statistical Machine Translation
Bilmes, J. and Kirchhoff, K. (2014) · 2014
Later among the works it cites.
Submodular Subset Selection for Large-Scale Speech Training Data
Wei, K., Liu, Y., Kirchhoff, K., Bartels, C., and Bilmes, J. (2014) · 2014
Later among the works it cites.
Class-Based N-gram Language Difference Models for Data Selection
Axelrod, A., Vyas, Y., Martindale, M., and Carpuat, M. (2015) · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…