Fetching the paper…
Reading the bibliography…
The primary objective of our work is to build a large-scale English-Thai dataset for machine translation.
Multilingual universal sentence encoder for semantic retrieval
Yang, Y., Cer, D. M., Ahmad, A., Guo, M., Law, J., Constant, N., Ábrego, G. H., Yuan, S., Tar, C., Sung, Y.-H., Strope, B., and Kurzweil, R. (2019) · 1907
Earlier work this paper cites.
Taskmaster-1: Toward a realistic and diverse dialog dataset
Byrne, B., Krishnamoorthi, K., Sankar, C., Neelakantan, A., Duckworth, D., Yavuz, S., Goodrich, B., Dubey, A., Cedilnik, A., and Kim, K.-Y. (2019) · 1909
Earlier work this paper cites.
Nltk: The natural language toolkit
Loper, E. and Bird, S. (2002) · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J. (2002) · 2002
Earlier work this paper cites.
Automatically constructing a corpus of sentential paraphrases
Dolan, W. B. and Brockett, C. (2005) · 2005
Earlier work this paper cites.
Europarl: A parallel corpus for statistical machine translation
Koehn, P. (2005) · 2005
Earlier work this paper cites.
Thoughts on word and sentence segmentation in thai
Aroonmanakun, W. et al. (2007) · 2007
Earlier work this paper cites.
Bitextor, a free/open-source software to harvest translation memories from multilingual websites
Espl, M. and Transducens, G. (2009) · 2009
Earlier work this paper cites.
Creating a live, public short message service corpus: The nus sms corpus
Chen, T. and Kan, M.-Y. (2011) · 2011
Earlier work this paper cites.
Parallel data, tools and interfaces in OPUS
Tiedemann, J. (2012) · 2012
Earlier work this paper cites.
Crowdsourcing annotation for machine learning in natural language processing tasks (non-final version! proofread version will be uploaded april 30, 2012.)
Zaidan, O. (2012) · 2012
Cited alongside, same era.
The AMARA corpus: Building parallel language resources for the educational domain
Abdelali, A., Guzman, F., Sajjad, H., and Vogel, S. (2014) · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y. (2014) · 2014
Cited alongside, same era.
The iwslt 2015 evaluation campaign
Cettolo, M., Niehues, J., Stüker, S., Bentivogli, L., Cattoni, R., and Federico, M. (2015) · 2015
Cited alongside, same era.
A massively parallel corpus: The bible in 100 languages
Christodouloupoulos, C. and Steedman, M. (2015) · 2015
Cited alongside, same era.
OpenSubtitles2016: Extracting large parallel corpora from movie and TV subtitles
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L., and Polosukhin, I. (2017) · 2017
Later among the works it cites.
Achieving human parity on automatic chinese to english news translation
Hassan, H., Aue, A., Chen, C., Chowdhary, V., Clark, J., Federmann, C., Huang, X., Junczys-Dowmunt, M., Lewis, W., Li, M., Liu, S., Liu, T.-Y., Luo, R., Menezes, A., Qin, T., Seide, F., Tan, X., Tian, F., Wu, L., Wu, S., Xia, Y., Zhang, D., Zhang, Z., and Zhou, M. (2018) · 2018
Later among the works it cites.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing
Kudo, T. and Richardson, J. (2018) · 2018
Later among the works it cites.
Scaling neural machine translation
Ott, M., Edunov, S., Grangier, D., and Auli, M. (2018) · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lison, P. and Tiedemann, J. (2016) · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Y., Schuster, M., Chen, Z., Le, Q. V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., Macherey, K., Klingner, J., Shah, A., Johnson, M., Liu, X., Kaiser, L., Gouws, S., Kato, Y., Kudo, T., Kazawa, H., Stevens, K., Kurian, G., Patil, N., Wang, W., Young, C., Smith, J., Riesa, J., Rudnick, A., Vinyals, O., Corrado, G. S., Hughes, M., and Dean, J. (2016) · 2016
Cited alongside, same era.
The united nations parallel corpus v1.0
Ziemski, M., Junczys-Dowmunt, M., and Pouliquen, B. (2016) · 2016
Cited alongside, same era.
Convolutional sequence to sequence learning
Gehring, J., Auli, M., Grangier, D., Yarats, D., and Dauphin, Y. N. (2017) · 2017
Cited alongside, same era.
Six challenges for neural machine translation
Koehn, P. and Knowles, R. (2017) · 2017
Cited alongside, same era.
Top sites in thailand the sites in the top sites lists are ordered by their 1 month alexa traffic rank.the 1 month rank is calculated using a combination of average daily visitors and pageviews over the past month. the site with the highest combination of visitors and pageviews is ranked #1
Cited in the paper.
Post, M. (2018) · 2018
Later among the works it cites.
JW300: A wide-coverage parallel corpus for low-resource languages
Agić, Ž. and Vulić, I. (2019) · 2019
Later among the works it cites.
ParaCrawl: Web-scale parallel corpora for the languages of the EU
Esplà, M., Forcada, M., Ramírez-Sánchez, G., and Hoang, H. (2019) · 2019
Later among the works it cites.
Ctrl: A conditional transformer language model for controllable generation
Keskar, N. S., McCann, B., Varshney, L. R., Xiong, C., and Socher, R. (2019) · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Ott, M., Edunov, S., Baevski, A., Fan, A., Gross, S., Ng, N., Grangier, D., and Auli, M. (2019) · 2019
Later among the works it cites.
Pythainlp/pythainlp: Pythainlp 2.1.4
Phatthiyaphaibun, W., Chaovavanich, K., Polpanumas, C., Suriyawongkul, A., Lowphansirikul, L., and Chormai, P. (2020) · 2020
Closest in time.