Fetching the paper…
Reading the bibliography…
In this paper, we present an approach to learn multilingual sentence embeddings using a bi-directional dual-encoder with additive margin softmax.
Mining the web for bilingual text
Philip Resnik · 1999
Earlier work this paper cites.
Parallel web text mining for cross-language ir
Jiang Chen and Jian-Yun Nie · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
Mining english/chinese parallel documents from the world wide web
Christopher C Yang and Kar Wing Li · 2002
Earlier work this paper cites.
The web as a parallel corpus
Philip Resnik and Noah A Smith · 2003
Earlier work this paper cites.
Reliable measures for aligning japanese-english news articles and sentences
Masao Utiyama and Hitoshi Isahara · 2003
Earlier work this paper cites.
Improving machine translation performance by exploiting non-parallel corpora
Dragos Stefan Munteanu and Daniel Marcu · 2005
Earlier work this paper cites.
Extracting parallel sub-sentential fragments from non-parallel corpora
Dragos Stefan Munteanu and Daniel Marcu · 2006
Earlier work this paper cites.
A dom tree alignment model for mining parallel data from the web
Lei Shi, Cheng Niu, Ming Zhou, and Jianfeng Gao · 2006
Earlier work this paper cites.
Mining a comparable text corpus for a vietnamese-french statistical machine translation system
Thi-Ngoc-Diep Do, Viet-Bac Le, Brigitte Bigi, Laurent Besacier, and Eric Castelli · 2009
Earlier work this paper cites.
Large scale parallel document mining for machine translation
Jakob Uszkoreit, Jay M. Ponte, Ashok C. Popat, and Moshe Dubiner · 2010
Earlier work this paper cites.
Building a web-based parallel corpus and filtering out machine-translated text
Alexandra Antonova and Alexey Misyurev · 2011
Cited alongside, same era.
Japanese and korean voice search
M. Schuster and K. Nakajima · 2012
Cited alongside, same era.
Findings of the 2013 Workshop on Statistical Machine Translation
Ondřej Bojar, Christian Buck, Chris Callison-Burch, Christian Federmann, Barry Haddow, Philipp Koehn, Christof Monz, Matt Post, Radu Soricut, and Lucia Specia · 2013
Cited alongside, same era.
Nearest neighbor search in google correlate
Dan Vanderkam, Rob Schonberger, Henry Rowley, and Sanjiv Kumar · 2013
Cited alongside, same era.
Findings of the 2014 workshop on statistical machine translation
Ondřej Bojar, Christian Buck, Christian Federmann, Barry Haddow, Philipp Koehn, Johannes Leveling, Christof Monz, Pavel Pecina, Matt Post, Herve Saint-Amand, et al · 2014
Cited alongside, same era.
Abcnn: Attention-based convolutional neural network for modeling sentence pairs
H2@bucc18: Parallel sentence extraction from comparable corpora using multilingual sentence embeddings
Houda Bouamor and Hassan Sajjad · 2018
Later among the works it cites.
Learning cross-lingual sentence representations via a multi-task dual-encoder model
Muthuraman Chidambaram, Yinfei Yang, Daniel Cer, Steve Yuan, Yun-Hsuan Sung, Brian Strope, and Ray Kurzweil · 2018
Later among the works it cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Later among the works it cites.
Effective parallel corpus mining using bilingual sentence embeddings
Mandy Guo, Qinlan Shen, Yinfei Yang, Heming Ge, Daniel Cer, Gustavo Hernandez Abrego, Keith Stevens, Noah Constant, Yun-hsuan Sung, Brian Strope, and Ray Kurzweil · 2018
Later among the works it cites.
Achieving human parity on automatic chinese to english news translation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Wenpeng Yin, Hinrich Schütze, Bing Xiang, and Bowen Zhou · 2015
Cited alongside, same era.
Cícero Nogueira dos Santos, Ming Tan, Bing Xiang, and Bowen Zhou · 2016
Cited alongside, same era.
The united nations parallel corpus v1. 0
Michal Ziemski, Marcin Junczys-Dowmunt, and Bruno Pouliquen · 2016
Cited alongside, same era.
A deep neural network approach to parallel sentence extraction
Francis Grégoire and Philippe Langlais · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Margin-based parallel corpus mining with multilingual sentence embeddings
Mikel Artetxe and Holger Schwenk · 2018
Cited alongside, same era.
Hany Hassan, Anthony Aue, Chang Chen, Vishal Chowdhary, Jonathan Clark, Christian Federmann, Xuedong Huang, Marcin Junczys-Dowmunt, William Lewis, Mu Li, et al · 2018
Later among the works it cites.
Filtering and mining parallel data in a joint multilingual space
Holger Schwenk · 2018
Later among the works it cites.
Additive margin softmax for face verification
Feng Wang, Jian Cheng, Weiyang Liu, and Haijun Liu · 2018
Later among the works it cites.
Denoising neural machine translation training with trusted data and online data selection
Wei Wang, Taro Watanabe, Macduff Hughes, Tetsuji Nakagawa, and Ciprian Chelba · 2018
Later among the works it cites.
Learning semantic textual similarity from conversations
Yinfei Yang, Steve Yuan, Daniel Cer, Sheng-Yi Kong, Noah Constant, Petr Pilar, Heming Ge, Yun-hsuan Sung, Brian Strope, and Ray Kurzweil · 2018
Later among the works it cites.
Overview of the third bucc shared task: Spotting parallel sentences in comparable corpora
Pierre Zweigenbaum, Serge Sharoff, and Reinhard Rapp · 2018
Later among the works it cites.