Fetching the paper…
Reading the bibliography…
In this paper, we introduced our joint team SJTU-NICT 's participation in the WMT 2020 machine translation shared task.
Context-aware learning for neural machine translation
Sébastien Jean and Kyunghyun Cho. 2019 · 1903
Earlier work this paper cites.
Pkuseg: A toolkit for multi-domain chinese word segmentation
Ruixuan Luo, Jingjing Xu, Yi Zhang, Xuancheng Ren, and Xu Sun. 2019 · 1906
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Revealing the dark secrets of bert
Olga Kovaleva, Alexey Romanov, Anna Rogers, and Anna Rumshisky. 2019 · 1908
Earlier work this paper cites.
Revisiting self-training for neural sequence generation
Junxian He, Jiatao Gu, Jiajun Shen, and Marc’Aurelio Ranzato. 2019 · 1909
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1911
Earlier work this paper cites.
Improving conditioning in context-aware sequence to sequence models
Xinyi Wang, Jason Weston, Michael Auli, and Yacine Jernite. 2019 · 1911
Earlier work this paper cites.
Explicit sentence compression for neural machine translation
Zuchao Li, Rui Wang, Kehai Chen, Masao Utiyama, Eiichiro Sumita, Zhuosheng Zhang, and Hai Zhao. 2019a · 1912
Earlier work this paper cites.
A survey on document-level machine translation: Methods and evaluation
Sameen Maruf, Fahimeh Saleh, and Gholamreza Haffari. 2019 · 1912
Earlier work this paper cites.
The present status of automatic translation of languages
Yehoshua Bar-Hillel. 1960 · 1960
Earlier work this paper cites.
Probability of error of some adaptive pattern-recognition machines
H Scudder. 1965 · 1965
Earlier work this paper cites.
A program for aligning sentences in bilingual corpora
William A Gale and Kenneth Church. 1993 · 1993
Earlier work this paper cites.
Unsupervised word sense disambiguation rivaling supervised methods
David Yarowsky. 1995 · 1995
Earlier work this paper cites.
Speech & language processing
Dan Jurafsky. 2000 · 2000
Earlier work this paper cites.
Longformer: The long-document transformer
Iz Beltagy, Matthew E. Peters, and Arman Cohan. 2020 · 2004
Earlier work this paper cites.
Reference language based unsupervised neural machine translation
Zuchao Li, Hai Zhao, Rui Wang, Masao Utiyama, and Eiichiro Sumita. 2020b · 2004
Earlier work this paper cites.
Self-training for unsupervised neural machine translation in unbalanced training data scenarios
Haipeng Sun, Rui Wang, Kehai Chen, Masao Utiyama, Eiichiro Sumita, and Tiejun Zhao. 2020a · 2004
Earlier work this paper cites.
Moses: Open source toolkit for statistical machine translation
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, et al. 2007 · 2007
Earlier work this paper cites.
A comparison of pivot methods for phrase-based statistical machine translation
Masao Utiyama and Hitoshi Isahara. 2007 · 2007
Earlier work this paper cites.
Pivot language approach for phrase-based statistical machine translation
Hua Wu and Haifeng Wang. 2007 · 2007
Earlier work this paper cites.
A survey on transfer learning
Sinno Jialin Pan and Qiang Yang. 2009 · 2009
Cited alongside, same era.
On the importance of pivot language selection for statistical machine translation
Michael Paul, Hirofumi Yamamoto, Eiichiro Sumita, and Satoshi Nakamura. 2009 · 2009
Cited alongside, same era.
The probabilistic relevance framework: BM25 and beyond
Stephen Robertson and Hugo Zaragoza. 2009 · 2009
Cited alongside, same era.
Introduction to semi-supervised learning
Xiaojin Zhu and Andrew B Goldberg. 2009 · 2009
Cited alongside, same era.
Transfer learning
Lisa Torrey and Jude Shavlik. 2010 · 2010
Cited alongside, same era.
langid. py: An off-the-shelf language identification tool
Marco Lui and Timothy Baldwin. 2012 · 2012
Cited alongside, same era.
Context-aware neural machine translation learns anaphora resolution
Elena Voita, Pavel Serdyukov, Rico Sennrich, and Ivan Titov. 2018 · 2018
Later among the works it cites.
Minimum divergence vs. maximum margin: an empirical comparison on seq2seq models
Huan Zhang and Hai Zhao. 2018 · 2018
Later among the works it cites.
Modeling multi-turn conversation with deep utterance aggregation
Zhuosheng Zhang, Jiangtong Li, Pengfei Zhu, Hai Zhao, and Gongshen Liu. 2018 · 2018
Later among the works it cites.
On the use of bert for neural machine translation
Stéphane Clinchant, Kweon Woo Jung, and Vassilina Nikoulina. 2019 · 2019
Later among the works it cites.
Cross-lingual language model pretraining
Alexis Conneau and Guillaume Lample. 2019 · 2019
Later among the works it cites.
Transformer-xl: Attentive language models beyond a fixed-length context
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime G Carbonell, Quoc Le, and Ruslan Salakhutdinov. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lexical chain based cohesion models for document-level statistical machine translation
Deyi Xiong, Yang Ding, Min Zhang, and Chew Lim Tan. 2013 · 2013
Cited alongside, same era.
Dual learning for machine translation
Di He, Yingce Xia, Tao Qin, Liwei Wang, Nenghai Yu, Tie-Yan Liu, and Wei-Ying Ma. 2016 · 2016
Cited alongside, same era.
Wavenet: A generative model for raw audio
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Transfer learning for low-resource neural machine translation
Barret Zoph, Deniz Yuret, Jonathan May, and Kevin Knight. 2016 · 2016
Cited alongside, same era.
spacy 2: Natural language understanding with bloom embeddings, convolutional neural networks and incremental parsing
Matthew Honnibal and Ines Montani. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Recycling a pre-trained bert encoder for neural machine translation
Kenji Imamura and Eiichiro Sumita. 2019 · 2019
Later among the works it cites.
Effective cross-lingual transfer of neural machine translation models without shared vocabularies
Yunsu Kim, Yingbo Gao, and Hermann Ney. 2019 · 2019
Later among the works it cites.
Reformer: The efficient transformer
Nikita Kitaev, Lukasz Kaiser, and Anselm Levskaya. 2019 · 2019
Later among the works it cites.
ALBERT: A lite bert for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2019 · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
Analysing concatenation approaches to document-level nmt in two different domains
Yves Scherrer, Jörg Tiedemann, and Sharid Loáiciga. 2019 · 2019
Later among the works it cites.
Lattice-based transformer encoder for neural machine translation
Fengshun Xiao, Jiangtong Li, Hai Zhao, Rui Wang, and Kehai Chen. 2019 · 2019
Later among the works it cites.
XLNet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
Head-Driven Phrase Structure Grammar parsing on Penn Treebank
Junru Zhou and Hai Zhao. 2019 · 2019
Later among the works it cites.
Bipartite flat-graph network for nested named entity recognition
Ying Luo and Hai Zhao. 2020 · 2020
Closest in time.
Incorporating bert into neural machine translation
Jinhua Zhu, Yingce Xia, Lijun Wu, Di He, Tao Qin, Wengang Zhou, Houqiang Li, and Tieyan Liu. 2020 · 2020
Closest in time.
Syntax for semantic role labeling, to be, or not to be
Shexia He, Zuchao Li, Hai Zhao, and Hongxiao Bai. 2018 · 2071
Closest in time.