Mining a comparable text corpus for a vietnamese-french statistical machine translation system
Thi-Ngoc-Diep Do, Viet-Bac Le, Brigitte Bigi, Laurent Besacier, and Eric Castelli. 2009 · 2009
Cited alongside, same era.
United nations general assembly resolutions: A six-language parallel corpus
Alexandre Rafalovitch, Robert Dale, et al. 2009 · 2009
Cited alongside, same era.
Mint: A method for effective and scalable mining of named entity transliterations from large comparable corpora
Raghavendra Udupa, K Saravanan, A Kumaran, and Jagadeesh Jagarlamudi. 2009 · 2009
Cited alongside, same era.
Combining content-based and url-based heuristics to harvest aligned bitexts from multilingual sites with bitextor
Miquel Espla-Gomis and Mikel Forcada. 2010 · 2010
Cited alongside, same era.
Computing Krippendorff’s alpha-reliability
Klaus Krippendorff. 2011 · 2011
Cited alongside, same era.
Dirt cheap web-scale parallel text from the common crawl
Jason R Smith, Herve Saint-Amand, Magdalena Plamada, Philipp Koehn, Chris Callison-Burch, and Adam Lopez. 2013 · 2013
Cited alongside, same era.
N-gram counts and language models from the common crawl
Christian Buck, Kenneth Heafield, and Bas van Ooyen. 2014 · 2014
Cited alongside, same era.
Findings of the wmt 2016 bilingual document alignment shared task
Christian Buck and Philipp Koehn. 2016a · 2016
Cited alongside, same era.
Yoda system for wmt16 shared task: Bilingual document alignment
Aswarth Abhilash Dara and Yiu-Chang Lin. 2016 · 2016
Cited alongside, same era.
First steps towards coverage-based document alignment
Luís Gomes and Gabriel Pereira Lopes. 2016 · 2016
Cited alongside, same era.
The United Nations parallel corpus v1. 0
Michał Ziemski, Marcin Junczys-Dowmunt, and Bruno Pouliquen. 2016 · 2016
Cited alongside, same era.
Bag of tricks for efficient text classification
Armand Joulin, Edouard Grave, and Piotr Bojanowski Tomas Mikolov. 2017 · 2017
Cited alongside, same era.