Fetching the paper…
Reading the bibliography…
We present and apply two methods for addressing the problem of selecting relevant training data out of a general pool for use in tasks such as machine translation.
P. F. Brown, V. J. Della Pietra, P. V. DeSouza, J. C. Lai, and R. L. Mercer, “Class-Based N-gram Models of Natural Language,” Computational Linguistics , vol. 18, no. 4, pp. 467–479, 1992
1992
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-j. Zhu, “BLEU: A Method for Automatic Evaluation of Machine Translation,” ACL (Association for Computational Linguistics) , 2002
2002
Earlier work this paper cites.
F. J. Och and H. Ney, “A Systematic Comparison of Various Statistical Alignment Models,” Computational Linguistics , vol. 29, no. 1, pp. 19–51, mar 2003
2003
Earlier work this paper cites.
A. Sethy, P. G. Georgiou, and S. Narayanan, “Text Data Acquisition for Domain-Specific Language Models,” EMNLP (Empirical Methods in Natural Language Processing) , 2006
2006
Earlier work this paper cites.
P. Koehn, H. Hoang, A. Birch-Mayne, C. Callison-Burch, M. Federico, N. Bertoldi, B. Cowan, W. Shen, C. Moran, R. Zens, C. Dyer, O. Bojar, A. Constantin, and E. Herbst, “Moses: Open Source Toolkit for Statistical Machine Translation,” ACL (Association for Computational Linguistics) Interactive Poster and Demonstration Sessions , 2007
2007
Earlier work this paper cites.
C. Biemann, U. Quasthoff, G. Heyer, and F. Holz, “ASV Toolbox: a Modular Collection of Language Exploration Tools,” in LREC (Language Resources and Evaluation) , 2008
2008
Earlier work this paper cites.
D. Chiang, Y. Marton, and P. Resnik, “Online Large-Margin Training of Syntactic and Structural Translation Features,” in EMNLP (Empirical Methods in Natural Language Processing) , 2008
2008
Cited alongside, same era.
R. C. Moore and W. D. Lewis, “Intelligent Selection of Language Model Training Data,” ACL (Association for Computational Linguistics) , 2010
2010
Cited alongside, same era.
A. Axelrod, X. He, and J. Gao, “Domain Adaptation Via Pseudo In-Domain Data Selection,” EMNLP (Empirical Methods in Natural Language Processing) , 2011
2011
Cited alongside, same era.
K. Heafield, “KenLM : Faster and Smaller Language Model Queries,” WMT (Workshop on Statistical Machine Translation) , 2011
2011
Cited alongside, same era.
M. Cettolo, C. Girardi, and M. Federico, “WITˆ3 : Web Inventory of Transcribed and Translated Talks,” EAMT (European Association for Machine Translation) , 2012
C. Dyer, V. Chahuneau, and N. A. Smith, “A Simple, Fast, and Effective Reparameterization of IBM Model 2,” in NAACL (North American Association for Computational Linguistics) , 2013
2013
Later among the works it cites.
A. Axelrod, Y. Vyas, M. Martindale, and M. Carpuat, “Class-Based N-gram Language Difference Models for Data Selection,” IWSLT (International Workshop on Spoken Language Translation) , 2015
2015
Later among the works it cites.
M. Kazi, E. Salesky, B. Thompson, J. Taylor, J. Gwinnup, T. Anderson, G. Erdmann, E. Hansen, B. Ore, K. Young, and M. Hutt, “The MITLL-AFRL IWSLT 2016 Systems,” Proceedings of the International Workshop on Spoken Language Translation (IWSLT) , 2016
2016
Later among the works it cites.
A. Axelrod, “Cynical Selection of Language Model Training Data,” arXiv [cs.CL] , pp. 1–19, 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2012
Cited alongside, same era.
J. Ganitkevitch, Y. Cao, J. Weese, M. Post, and C. Callison-Burch, “Joshua 4.0: Packing, PRO, and Paraphrases,” in WMT (Workshop on Statistical Machine Translation) , 2012
2012
Cited alongside, same era.
N.-Q. Pham, J. Niehues, T.-L. Ha, E. Cho, M. Sperber, and A. Waibel, “The Karlsruhe Institute of Technology Systems for the News Translation Task in WMT 2017,” WMT Conference on Statistical Machine Translation , 2017
2017
Later among the works it cites.