Fetching the paper…
Reading the bibliography…
Multimodal machine translation involves drawing information from more than one modality, based on the assumption that the additional modalities will contain useful alternative views of the input data.
1902
Earlier work this paper cites.
1904
Earlier work this paper cites.
1907
Earlier work this paper cites.
1907
Earlier work this paper cites.
1909
Earlier work this paper cites.
1910
Earlier work this paper cites.
1910
Earlier work this paper cites.
1911
Earlier work this paper cites.
Calixto I, Liu Q, Campbell N (2017) Doubly-attentive decoder for multi-modal neural machine translation. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Association for Computational Linguistics, pp 1913–1924
1924
Earlier work this paper cites.
Unal ME, Citamak B, Yagcioglu S, Erdem A, Erdem E, Cinbis NI, Cakici R (2016) Tasviret: A benchmark dataset for automatic Turkish description generation from images. In: 2016 24th Signal Processing and Communication Application Conference (SIU), pp 1977–1980
1980
Earlier work this paper cites.
Morimoto T (1990) Automatic interpreting telephony research at ATR. In: Proceedings of a Workshop on Machine Translation, UMIST
1990
Earlier work this paper cites.
Caruana R (1997) Multitask learning. Machine Learning 28(1):41–75
1997
Earlier work this paper cites.
Hochreiter S, Schmidhuber J (1997) Long Short-Term Memory. Neural Computation 9(8):1735–1780
1997
Earlier work this paper cites.
Lavie A, Waibel A, Levin L, Finke M, Gates D, Gavalda M, Zeppenfeld T, Zhan P (1997) JANUS-III: Speech-to-speech translation in multiple languages. In: 1997 IEEE International Conference on Acoustics, Speech, and Signal Processing, IEEE Comput. Soc. Press, Munich, Germany, vol 1, pp 99–102
1997
Earlier work this paper cites.
Schuster M, Paliwal KK (1997) Bidirectional recurrent neural networks. IEEE Transactions on Signal Processing 45(11):2673–2681
1997
Earlier work this paper cites.
Vidal E (1997) Finite-state speech-to-speech translation. In: Proceedings of the 1997 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), IEEE, vol 1, pp 111–114
1997
Earlier work this paper cites.
Takezawa T, Morimoto T, Sagisaka Y, Campbell N, Iida H, Sugaya F, Yokoo A, Yamamoto S (1998) A Japanese-to-English Speech Translation System: ATR-MATRIX. In: Fifth International Conference on Spoken Language Processing
1998
Earlier work this paper cites.
Ney H (1999) Speech translation: Coupling of recognition and translation. In: Proceedings of the 1999 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), IEEE, vol 1, pp 517–520
1999
Earlier work this paper cites.
Wahlster W (2000) Mobile Speech-to-Speech Translation of Spontaneous Dialogs: An Overview of the Final Verbmobil System. In: Wahlster W (ed) Verbmobil: Foundations of Speech-to-Speech Translation, Springer Berlin Heidelberg, Berlin, Heidelberg, pp 3–21
2000
Earlier work this paper cites.
Papineni K, Roukos S, Ward T, Zhu WJ (2001) BLEU: a method for automatic evaluation of machine translation. In: Proceedings of the 40th Annual Meeting on Association for Computational Linguistics - ACL ’02, Association for Computational Linguistics, Philadelphia, Pennsylvania
2001
Earlier work this paper cites.
Chesterman A, Wagner E (2002) Can Theory Help Translators? A Dialogue Between the Ivory Tower and the Wordface. Routledge
2002
Earlier work this paper cites.
Bengio Y, Ducharme R, Vincent P, Jauvin C (2003) A neural probabilistic language model. Journal of machine learning research 3(Feb):1137–1155
2003
Earlier work this paper cites.
Och FJ (2003) Minimum Error Rate Training in Statistical Machine Translation. In: Proceedings of the 41st Annual Meeting on Association for Computational Linguistics - Volume 1, Association for Computational Linguistics, Stroudsburg, PA, USA, ACL ’03, pp 160–167
2003
Earlier work this paper cites.
Akiba Y, Federico M, Kando N, Nakaiwa H, Paul M, Tsujii J (2004) Overview of the IWSLT 2004 evaluation campaign. In: Proceedings of the 2004 International Workshop on Spoken Language Translation, Kyoto, Japan
2004
Earlier work this paper cites.
Gella S, Elliott D, Keller F (2019) Cross-lingual visual verb sense disambiguation. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers), Association for Computational Linguistics, Minneapolis, Minnesota, pp 1998–2004
2004
Earlier work this paper cites.
Clough P, Grubinger M, Deselaers T, Hanbury A, Müller H (2006) Overview of the ImageCLEF 2006 photographic retrieval and object annotation tasks. In: Proceedings of the 7th International Conference on Cross-Language Evaluation Forum (CLEF), Springer, pp 579–594
2006
Earlier work this paper cites.
Grubinger M, Clough P, Müller H, Deselaers T (2006) The IAPR TC-12 Benchmark: A New Evaluation Resource for Visual Information Systems. In: Proceedings of the OntoImage Workshop on Language Resources for Content-based Image Retrieval, Genoa, Italy, pp 13–23
2006
Earlier work this paper cites.
Matusov E, Kanthak S, Ney H (2006) Integrating speech recognition and machine translation: Where do we stand? In: 2006 IEEE International Conference on Acoustics Speech and Signal Processing Proceedings, vol 5, pp V–V
2006
Earlier work this paper cites.
Schwenk H, Dechelotte D, Gauvain JL (2006) Continuous space language models for statistical machine translation. In: Proceedings of the COLING/ACL 2006 Main Conference Poster Sessions, Association for Computational Linguistics, pp 723–730
2006
Earlier work this paper cites.
Snover M, Dorr B, Schwartz R, Micciulla L, Makhoul J (2006) A Study of Translation Edit Rate with Targeted Human Annotation. In: Proceedings of Association for Machine Translation in the Americas, 6
2006
Earlier work this paper cites.
Koehn P, Zens R, Dyer C, Bojar O, Constantin A, Herbst E, Hoang H, Birch A, Callison-Burch C, Federico M, Bertoldi N, Cowan B, Shen W, Moran C (2007) Moses: open source toolkit for statistical machine translation. In: Proceedings of the 45th Annual Meeting of the ACL on Interactive Poster and Demonstration Sessions - ACL ’07, Association for Computational Linguistics, Prague, Czech Republic
2007
Earlier work this paper cites.
Lavie A, Agarwal A (2007) Meteor: an automatic metric for MT evaluation with high levels of correlation with human judgments. In: Proceedings of the Second Workshop on Statistical Machine Translation - StatMT ’07, Association for Computational Linguistics, Prague, Czech Republic, pp 228–231
2007
Earlier work this paper cites.
Ramirez J, Gorriz JM, Segura JC (2007) Voice activity detection. fundamentals and speech recognition system robustness. In: Grimm M, Kroschel K (eds) Robust Speech, IntechOpen, Rijeka, chap 1
2007
Earlier work this paper cites.
Deng J, Dong W, Socher R, Li LJ, Li K, Fei-Fei L (2009) Imagenet: A large-scale hierarchical image database. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, IEEE, pp 248–255
2009
Earlier work this paper cites.
Koehn P (2009) Statistical machine translation. Cambridge University Press
2009
Earlier work this paper cites.
Stein BE, Stanford TR, Rowland BA (2009) The neural basis of multisensory integration in the midbrain: Its organization and maturation. Hearing Research 258(1):4 – 15, multisensory integration in auditory and auditory-related areas of cortex
2009
Earlier work this paper cites.
Mikolov T, Karafiát M, Burget L, Cernocký J, Khudanpur S (2010) Recurrent neural network based language model. In: Kobayashi T, Hirose K, Nakamura S (eds) INTERSPEECH, ISCA, pp 1045–1048
2010
Earlier work this paper cites.
Paul M, Federico M, Stüker S (2010) Overview of the IWSLT 2010 evaluation campaign. In: Proceedings of the 2010 International Workshop on Spoken Language Translation
2010
Earlier work this paper cites.
Rashtchian C, Young P, Hodosh M, Hockenmaier J (2010) Collecting image annotations using Amazon’s Mechanical Turk. In: Proceedings of the Workshop on Creating Speech and Language Data with Amazon’s Mechanical Turk, Association for Computational Linguistics, pp 139–147
2010
Earlier work this paper cites.
Xiao J, Hays J, Ehinger KA, Oliva A, Torralba A (2010) SUN database: Large-scale scene recognition from abbey to zoo. In: The Twenty-Third IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2010, San Francisco, CA, USA, 13-18 June 2010, pp 3485–3492
2010
Earlier work this paper cites.
He X, Deng L, Acero A (2011) Why word error rate is not a good metric for speech recognizer training for the speech translation task? In: 2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp 5632–5635
2011
Earlier work this paper cites.
Yao B, Jiang X, Khosla A, Lin AL, Guibas L, Fei-Fei L (2011) Human action recognition by learning bases of action attributes and parts. In: 2011 International Conference on Computer Vision, pp 1331–1338
2011
Earlier work this paper cites.
Anguera X, Bozonnet S, Evans N, Fredouille C, Friedland G, Vinyals O (2012) Speaker diarization: A review of recent research. IEEE Transactions on Audio, Speech, and Language Processing 20(2):356–370
2012
Earlier work this paper cites.
Cettolo M, Girardi C, Federico M (2012) WIT3: Web Inventory of Transcribed and Translated Talks. In: Proceedings of the 16th Conference of the European Association for Machine Translation, Trento, Italy, pp 261–268
2012
Earlier work this paper cites.
Mohamed A, Hinton G, Penn G (2012) Understanding how deep belief networks perform acoustic modelling. In: 2012 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp 4273–4276
2012
Earlier work this paper cites.
Peitz S, Wiesler S, Nußbaum-Thom M, Ney H (2012) Spoken language translation using automatically transcribed text in training. In: Proceedings of the 9th International Workshop on Spoken Language Translation, pp 276–283
2012
Earlier work this paper cites.
Tiedemann J (2012) Parallel Data, Tools and Interfaces in OPUS. In: Chair) NCC, Choukri K, Declerck T, Doğan MU, Maegaard B, Mariani J, Moreno A, Odijk J, Piperidis S (eds) Proceedings of the Eight International Conference on Language Resources and Evaluation (LREC’12), European Language Resources Association (ELRA), Istanbul, Turkey
2012
Earlier work this paper cites.
Drugan J (2013) Quality in Professional Translation: Assessment and Improvement. Continuum Advances in Translation, Bloomsbury Academic
2013
Earlier work this paper cites.
Graham Y, Baldwin T, Moffat A, Zobel J (2013) Continuous measurement scales in human evaluation of machine translation. In: Proceedings of the 7th Linguistic Annotation Workshop and Interoperability with Discourse, Association for Computational Linguistics, Sofia, Bulgaria, pp 33–41
2013
Earlier work this paper cites.
Graves A, Mohamed Ar, Hinton G (2013) Speech recognition with deep recurrent neural networks. In: 2013 IEEE International Conference on Acoustics, Speech and Signal Processing, IEEE, Vancouver, BC, Canada, pp 6645–6649
2013
Earlier work this paper cites.
Guzman F, Sajjad H, Vogel S, Abdelali A (2013) The AMARA Corpus: Building resources for translating the web’s educational content. In: Proceedings of the 10th International Workshop on Spoken Language Translation (IWSLT), Heidelberg, Germany
2013
Earlier work this paper cites.
He X, Deng L (2013) Speech-centric information processing: An optimization-oriented approach. Proceedings of the IEEE 101(5):1116–1135
2013
Earlier work this paper cites.
Kalchbrenner N, Blunsom P (2013) Recurrent Continuous Translation Models. In: Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing, Association for Computational Linguistics, Seattle, Washington, USA, pp 1700–1709
2013
Earlier work this paper cites.
Post M, Kumar G, Lopez A, Karakos D, Callison-Burch C, Khudanpur S (2013) Improved speech-to-text translation with the Fisher and Callhome Spanish-English speech translation corpus. In: Proceedings of the 10th International Workshop on Spoken Language Translation (IWSLT), Heidelberg, Germany
2013
Earlier work this paper cites.
Saon G, Soltau H, Nahamoo D, Picheny M (2013) Speaker adaptation of neural network acoustic models using i-vectors. In: 2013 IEEE Workshop on Automatic Speech Recognition and Understanding, pp 55–59
2013
Earlier work this paper cites.
Zhou B (2013) Statistical machine translation for speech: A perspective on structures, learning, and decoding. Proceedings of the IEEE 101(5):1180–1202
2013
Earlier work this paper cites.
Abdelali A, Guzman F, Sajjad H, Vogel S (2014) The AMARA Corpus: Building parallel language resources for the educational domain. In: Proceedings of the 9th International Conference on Language Resources and Evaluation (LREC), Reykjavík, Iceland, pp 1856–1862
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
Denkowski M, Lavie A (2014) Meteor universal: Language specific translation evaluation for any target language. In: Proceedings of the Ninth Workshop on Statistical Machine Translation, Association for Computational Linguistics, pp 376–380
2014
Earlier work this paper cites.
Ghahremani P, BabaAli B, Povey D, Riedhammer K, Trmal J, Khudanpur S (2014) A pitch extraction algorithm tuned for automatic speech recognition. In: 2014 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp 2494–2498
2014
Earlier work this paper cites.
Girshick R, Donahue J, Darrell T, Malik J (2014) Rich feature hierarchies for accurate object detection and semantic segmentation. In: The IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2014
Earlier work this paper cites.
Jansen D, Alcala A, Guzman F (2014) AMARA: A sustainable, global solution for accessibility, powered by communities of volunteers. In: Universal Access in Human-Computer Interaction. Design for All and Accessibility Practice, Springer, pp 401–411
2014
Earlier work this paper cites.
Kiros R, Salakhutdinov R, Zemel R (2014) Multimodal Neural Language Models. In: Proceedings of the 31st International Conference on Machine Learning
2014
Earlier work this paper cites.
Ramanathan V, Joulin A, Liang P, Fei-Fei L (2014) Linking people in videos with “their” names using coreference resolution. In: European conference on computer vision, Springer, pp 95–110
2014
Earlier work this paper cites.
Ruiz N, Federico M (2014) Assessing the impact of speech recognition errors on machine translation quality. AMTA 2014: proceedings of the eleventh conference of the Association for Machine Translation in the Americas, Vancouver, BC pp 261–274
2014
Earlier work this paper cites.
Sutskever I, Vinyals O, Le QV (2014) Sequence to Sequence Learning with Neural Networks. In: Proceedings of the 27th International Conference on Neural Information Processing Systems, MIT Press, Cambridge, MA, USA, NIPS’14, pp 3104–3112
2014
Earlier work this paper cites.
Tsvetkov Y, Metze F, Dyer C (2014) Augmenting translation models with simulated acoustic confusions for improved spoken language translation. In: Proceedings of the 14th Conference of the European Chapter of the Association for Computational Linguistics, Association for Computational Linguistics, Gothenburg, Sweden, pp 616–625
2014
Earlier work this paper cites.
Young P, Lai A, Hodosh M, Hockenmaier J (2014) From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions. Transactions of the Association for Computational Linguistics 2:67–78
2014
Earlier work this paper cites.
Antol S, Agrawal A, Lu J, Mitchell M, Batra D, Lawrence Zitnick C, Parikh D (2015) VQA: Visual Question Answering. In: Proceedings of the IEEE international conference on computer vision, pp 2425–2433
2015
Cited alongside, same era.
Bahdanau D, Cho K, Bengio Y (2015) Neural Machine Translation by Jointly Learning to Align and Translate. In: Proceedings of the 3rd International Conference on Learning Representations, pp San Diego, CA, USA
2015
Cited alongside, same era.
2015
Cited alongside, same era.
Dong D, Wu H, He W, Yu D, Wang H (2015) Multi-task learning for multiple language translation. In: Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 1: Long Papers), Association for Computational Linguistics, Beijing, China, pp 1723–1732
Yoshikawa Y, Shigeto Y, Takeuchi A (2017) STAIR captions: Constructing a large-scale Japanese image caption dataset. In: Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), Association for Computational Linguistics, Vancouver, Canada, pp 417–421
2017
Later among the works it cites.
Yu D, Li J (2017) Recent progresses in deep learning based acoustic models. IEEE/CAA Journal of Automatica Sinica 4(3):396–409
2017
Later among the works it cites.
Zhang J, Utiyama M, Sumita E, Neubig G, Nakamura S (2017) NICT-NAIST system for WMT17 multimodal translation task. In: Proceedings of the Second Conference on Machine Translation, Volume 2: Shared Task Papers, Association for Computational Linguistics, Copenhagen, Denmark, pp 477–482
2017
Later among the works it cites.
Anastasopoulos A, Chiang D (2018) Tied multitask learning for neural speech translation. In: Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers), Association for Computational Linguistics, New Orleans, Louisiana, pp 82–91
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
2015
Cited alongside, same era.
Ioffe S, Szegedy C (2015) Batch normalization: Accelerating deep network training by reducing internal covariate shift. In: Proceedings of The 32nd International Conference on Machine Learning, pp 448–456
2015
Cited alongside, same era.
Ling ZH, Kang SY, Zen H, Senior A, Schuster M, Qian XJ, Meng HM, Deng L (2015) Deep learning for acoustic modeling in parametric speech generation: A systematic review of existing techniques and future trends. IEEE Signal Processing Magazine 3(32):35–52
2015
Cited alongside, same era.
2015
Cited alongside, same era.
Mao J, Xu W, Yang Y, Wang J, Huang Z, Yuille A (2015) Deep captioning with multimodal recurrent neural networks (m-rnn). In: International Conference on Learning Representations (ICLR)
2015
Cited alongside, same era.
Panayotov V, Chen G, Povey D, Khudanpur S (2015) Librispeech: an ASR corpus based on public domain audio books. In: Acoustics, Speech and Signal Processing (ICASSP), 2015 IEEE International Conference on, IEEE, pp 5206–5210
2015
Cited alongside, same era.
Pulkki V, Karjalainen M (2015) Communication acoustics: an introduction to speech, audio and psychoacoustics. John Wiley & Sons
2015
Cited alongside, same era.
Ruiz N, Federico M (2015) Phonetically-oriented word error alignment for speech recognition error analysis in speech translation. In: 2015 IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU), pp 296–302
2015
Cited alongside, same era.
2018
Later among the works it cites.
2018
Later among the works it cites.
Bansal S, Kamper H, Livescu K, Lopez A, Goldwater S (2018) Low-resource speech-to-text translation. In: Interspeech 2018, pp 1298–1302
2018
Later among the works it cites.
Barrault L, Bougares F, Specia L, Lala C, Elliott D, Frank S (2018) Findings of the Third Shared Task on Multimodal Machine Translation. In: Proceedings of the Third Conference on Machine Translation, Volume 2: Shared Task Papers, Association for Computational Linguistics, Belgium, Brussels, pp 308–327
2018
Later among the works it cites.
Belinkov Y, Bisk Y (2018) Synthetic and natural noise both break neural machine translation. In: International Conference on Learning Representations
2018
Later among the works it cites.
Bentivogli L, Cettolo M, Federico M, Federmann C (2018) Machine Translation Human Evaluation: An investigation of evaluation based on Post-Editing and its relation with Direct Assessment. In: Proceedings of the 2018 International Workshop on Spoken Language Translation, Bruges, Belgium, pp 62–69
2018
Later among the works it cites.
Bérard A, Besacier L, Kocabiyikoglu AC, Pietquin O (2018) End-to-end automatic speech translation of audiobooks. In: International Conference on Acoustics, Speech and Signal Processing (ICASSP), IEEE
2018
Later among the works it cites.
Caglayan O, Bardet A, Bougares F, Barrault L, Wang K, Masana M, Herranz L, van de Weijer J (2018) LIUM-CVC submissions for WMT18 multimodal translation task. In: Proceedings of the Third Conference on Machine Translation, Association for Computational Linguistics, Belgium, Brussels, pp 603–608
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
Castilho S, Doherty S, Gaspari F, Moorkens J (2018) Approaches to human and machine translation quality assessment. In: Translation Quality Assessment: From Principles to Practice, Machine Translation: Technologies and Applications, Springer International Publishing, pp 9–38
2018
Later among the works it cites.
Cheng Y, Tu Z, Meng F, Zhai J, Liu Y (2018) Towards robust neural machine translation. In: Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Association for Computational Linguistics, Melbourne, Australia, pp 1756–1766
2018
Later among the works it cites.
Chiu C, Sainath TN, Wu Y, Prabhavalkar R, Nguyen P, Chen Z, Kannan A, Weiss RJ, Rao K, Gonina E, Jaitly N, Li B, Chorowski J, Bacchiani M (2018) State-of-the-art speech recognition with sequence-to-sequence models. In: 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp 4774–4778
2018
Later among the works it cites.
Delbrouck JB, Dupont S (2018) UMONS Submission for WMT18 Multimodal Translation Task. In: Proceedings of the Third Conference on Machine Translation, Volume 2: Shared Task Papers, Association for Computational Linguistics, Belgium, Brussels, pp 643–647
2018
Later among the works it cites.
Dong L, Xu S, Xu B (2018) Speech-transformer: a no-recurrence sequence-to-sequence model for speech recognition. In: 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), IEEE, pp 5884–5888
2018
Later among the works it cites.
Elaraby M, Tawfik AY, Khaled M, Hassan H, Osama A (2018) Gender aware spoken language translation applied to english-arabic. In: 2018 2nd International Conference on Natural Language and Speech Processing (ICNLSP), IEEE, pp 1–6
2018
Later among the works it cites.
Elliott D (2018) Adversarial evaluation of multimodal machine translation. In: Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, Association for Computational Linguistics, pp 2974–2978
2018
Later among the works it cites.
Frank S, Elliott D, Specia L (2018) Assessing multilingual multimodal image description: Studies of native speaker preferences and translator choices. Natural Language Engineering 24(03):393–413
2018
Later among the works it cites.
Grönroos SA, Huet B, Kurimo M, Laaksonen J, Merialdo B, Pham P, Sjöberg M, Sulubacak U, Tiedemann J, Troncy R, Vázquez R (2018) The memad submission to the wmt18 multimodal translation task. In: Proceedings of the Third Conference on Machine Translation, Volume 2: Shared Task Papers, Association for Computational Linguistics, Belgium, Brussels, pp 609–617
2018
Later among the works it cites.
Gwinnup J, Sandvick J, Hutt M, Erdmann G, Duselis J, Davis J (2018) The AFRL-Ohio State WMT18 multimodal system: Combining visual with traditional. In: Proceedings of the Third Conference on Machine Translation, Association for Computational Linguistics, Belgium, Brussels, pp 618–621
2018
Later among the works it cites.
Junczys-Dowmunt M, Grundkiewicz R, Dwojak T, Hoang H, Heafield K, Neckermann T, Seide F, Germann U, Aji AF, Bogoychev N, Martins AFT, Birch A (2018) Marian: Fast neural machine translation in C++. In: Proceedings of ACL 2018, System Demonstrations, Association for Computational Linguistics, Melbourne, Australia, pp 116–121
2018
Later among the works it cites.
Kádár Á, Elliott D, Côté MA, Chrupała G, Alishahi A (2018) Lessons learned in multilingual grounded language learning. In: Proceedings of the 22nd Conference on Computational Natural Language Learning, Association for Computational Linguistics, Brussels, Belgium, pp 402–412
2018
Later among the works it cites.
Kocabiyikoglu AC, Besacier L, Kraif O (2018) Augmenting Librispeech with French Translations: A Multimodal Corpus for Direct Speech Translation Evaluation. In: Proceedings of the 11th Conference on Language Resources and Evaluation (LREC), European Language Resources Association (ELRA)
2018
Later among the works it cites.
Lala C, Specia L (2018) Multimodal Lexical Translation. In: Proceedings of the 11th Conference on Language Resources and Evaluation, Miyazaki, Japan
2018
Later among the works it cites.
Lala C, Madhyastha PS, Scarton C, Specia L (2018) Sheffield submissions for WMT18 multimodal translation shared task. In: Proceedings of the Third Conference on Machine Translation, Association for Computational Linguistics, Belgium, Brussels, pp 630–637
2018
Later among the works it cites.
Libovický J, Helcl J, Mareček D (2018) Input combination strategies for multi-source transformer decoder. In: Proceedings of the Third Conference on Machine Translation: Research Papers, Association for Computational Linguistics, Belgium, Brussels, pp 253–260
2018
Later among the works it cites.
Liu D, Liu J, Guo W, Xiong S, Ma Z, Song R, Wu C, Liu Q (2018) The ustc-nel speech translation system at iwslt 2018. In: Proceedings of the 15th International Workshop on Spoken Language Translation, pp 70–75
2018
Later among the works it cites.
Ma Q, Bojar O, Graham Y (2018) Results of the WMT18 Metrics Shared Task: Both characters and embeddings achieve good performance. In: Proceedings of the Third Conference on Machine Translation, Volume 2: Shared Task Papers, Association for Computational Linguistics, Belgium, Brussels, pp 682–701
2018
Later among the works it cites.
Niehues J, Cattoni R, Stüker S, Cettolo M, Turchi M, Federico M (2018) The IWSLT 2018 Evaluation Campaign. In: Proceedings of the 2018 International Workshop on Spoken Language Translation, Bruges, Belgium
2018
Later among the works it cites.
Osamura K, Kano T, Sakti S, Sudoh K, Nakamura S (2018) Using spoken word posterior features in neural machine translation. In: Proceedings of the 15th International Workshop on Spoken Language Translation, pp 189–195
2018
Later among the works it cites.
Sanabria R, Caglayan O, Palaskar S, Elliott D, Barrault L, Specia L, Metze F (2018) How2: A large-scale dataset for multimodal language understanding. In: Proceedings of the Workshop on Visually Grounded Interaction and Language (NeurIPS 2018)
2018
Later among the works it cites.
Specia L, Blain F, Logacheva V, Astudillo RF, Martins A (2018) Findings of the WMT 2018 Shared Task on Quality Estimation. In: Proceedings of the Third Conference on Machine Translation, Volume 2: Shared Task Papers, Association for Computational Linguistics, Belgium, Brussels, pp 702–722
2018
Later among the works it cites.
Vaswani A, Bengio S, Brevdo E, Chollet F, Gomez A, Gouws S, Jones L, Kaiser Ł, Kalchbrenner N, Parmar N, Sepassi R, Shazeer N, Uszkoreit J (2018) Tensor2Tensor for neural machine translation. In: Proceedings of the 13th Conference of the Association for Machine Translation in the Americas (Volume 1: Research Papers), Association for Machine Translation in the Americas, Boston, MA, pp 193–199
2018
Later among the works it cites.
Wang A, Singh A, Michael J, Hill F, Levy O, Bowman S (2018a) GLUE: A multi-task benchmark and analysis platform for natural language understanding. In: Proceedings of the 2018 EMNLP Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP, Association for Computational Linguistics, Brussels, Belgium, pp 353–355
2018
Later among the works it cites.
Wang X, Pham H, Dai Z, Neubig G (2018b) SwitchOut: an efficient data augmentation algorithm for neural machine translation. In: Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, Association for Computational Linguistics, Brussels, Belgium, pp 856–861
2018
Later among the works it cites.
Wang Y, Shi L, Wei L, Zhu W, Chen J, Wang Z, Wen S, Chen W, Wang Y, Jia J (2018c) The Sogou-TIIC Speech Translation System for IWSLT 2018. In: Proceedings of the 2018 International Workshop on Spoken Language Translation, pp 112–117
2018
Later among the works it cites.
Watanabe S, Hori T, Karita S, Hayashi T, Nishitoba J, Unno Y, Soplin NEY, Heymann J, Wiesner M, Chen N, et al (2018) Espnet: End-to-end speech processing toolkit. In: Interspeech 2018, pp 2207–2211
2018
Later among the works it cites.
Zheng R, Yang Y, Ma M, Huang L (2018) Ensemble sequence level training for multimodal MT: OSU-Baidu WMT18 multimodal machine translation system report. In: Proceedings of the Third Conference on Machine Translation, Association for Computational Linguistics, Belgium, Brussels, pp 638–642
2018
Later among the works it cites.
Bahar P, Zeyer A, Schlüter R, Ney H (2019) On using specaugment for end-to-end speech translation. In: Proceedings of the 16th International Workshop on Spoken Language Translation
2019
Closest in time.
Bansal S, Kamper H, Livescu K, Lopez A, Goldwater S (2019) Pre-training on high-resource speech recognition improves low-resource speech-to-text translation. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers), Association for Computational Linguistics, Minneapolis, Minnesota, pp 58–68
2019
Closest in time.
Caglayan O (2019) Multimodal Machine Translation. Theses, Université du Maine
2019
Closest in time.
Caglayan O, Madhyastha P, Specia L, Barrault L (2019) Probing the need for visual context in multimodal machine translation. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers), Association for Computational Linguistics, Minneapolis, Minnesota, pp 4159–4170
2019
Closest in time.
Di Gangi M, Negri M, Nguyen VN, Tebbifakhr A, Turchi M (2019) Data augmentation for end-to-end speech translation: FBK@IWSLT ’19. In: Proceedings of the 16th International Workshop on Spoken Language Translation
2019
Closest in time.
Di Gangi MA, Negri M, Turchi M (2019b) Adapting transformer to end-to-end spoken language translation. In: INTERSPEECH 2019, International Speech Communication Association (ISCA), pp 1133–1137
2019
Closest in time.
Di Gangi MA, Negri M, Turchi M (2019c) One-to-many multilingual end-to-end speech translation. In: 2019 IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU)
2019
Closest in time.
Dutta Chowdhury K, Elliott D (2019) Understanding the effect of textual adversaries in multimodal machine translation. In: Proceedings of the Beyond Vision and LANguage: inTEgrating Real-world kNowledge (LANTERN), Association for Computational Linguistics, Hong Kong, China, pp 35–40
2019
Closest in time.
Inaguma H, Duh K, Kawahara T, Watanabe S (2019a) Multilingual end-to-end speech translation. In: 2019 IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU)
2019
Closest in time.
Inaguma H, Kiyono S, Soplin NEY, Suzuki J, Duh K, Watanabe S (2019b) Espnet how2 speech translation system for iwslt 2019: Pre-training, knowledge distillation, and going deeper. In: Proceedings of the 16th International Workshop on Spoken Language Translation
2019
Closest in time.
Ive J, Madhyastha P, Specia L (2019) Distilling translations with visual awareness. In: Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, Association for Computational Linguistics, Florence, Italy, pp 6525–6538
2019
Closest in time.
Jia Y, Johnson M, Macherey W, Weiss RJ, Cao Y, Chiu CC, Ari N, Laurenzo S, Wu Y (2019) Leveraging weakly supervised data to improve end-to-end speech-to-text translation. In: ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), IEEE, pp 7180–7184
2019
Closest in time.
Li J, Lavrukhin V, Ginsburg B, Leary R, Kuchaiev O, Cohen JM, Nguyen H, Gadde RT (2019) Jasper: An End-to-End Convolutional Neural Acoustic Model. In: Proc. Interspeech 2019, pp 71–75
2019
Closest in time.
Libovický J (2019) Multimodality in Machine Translation. PhD thesis, Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics, Praha
2019
Closest in time.
Liu Y, Xiong H, He Z, Zhang J, Wu H, Wang H, Zong C (2019) End-to-end speech translation with knowledge distillation. In: Interspeech 2019
2019
Closest in time.
Lüscher C, Beck E, Irie K, Kitza M, Michel W, Zeyer A, Schlüter R, Ney H (2019) RWTH ASR Systems for LibriSpeech: Hybrid vs Attention. In: Proc. Interspeech 2019, pp 231–235
2019
Closest in time.
Ma Q, Wei J, Bojar O, Graham Y (2019) Results of the WMT19 metrics shared task: Segment-level and strong MT systems pose big challenges. In: Proceedings of the 4th Conference on Machine Translation (WMT), Association for Computational Linguistics, Florence, Italy, pp 62–90
2019
Closest in time.
Madhyastha P, Wang J, Specia L (2019) VIFIDEL: Evaluating the visual fidelity of image descriptions. In: Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, Florence, Italy, pp 6539–6550
2019
Closest in time.
Niehues J, Cattoni R, Stüker S, Negri M, Turchi M, Ha TL, Salesky E, Sanabria R, Barrault L, Specia L, Federico M (2019) The IWSLT 2019 evaluation campaign. In: Proceedings of the 16th International Workshop on Spoken Language Translation (IWSLT)
2019
Closest in time.
Ott M, Edunov S, Baevski A, Fan A, Gross S, Ng N, Grangier D, Auli M (2019) fairseq: A fast, extensible toolkit for sequence modeling. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics (Demonstrations), Association for Computational Linguistics, Minneapolis, Minnesota, pp 48–53
2019
Closest in time.
Pham NQ, Nguyen TS, Ha TL, Hussain J, Schneider F, Niehues J, Stüker S, Waibel A (2019) The iwslt 2019 kit speech translation system. In: Proceedings of the 16th International Workshop on Spoken Language Translation
2019
Closest in time.
Pino J, Puzon L, Gu J, Ma X, McCarthy AD, Gopinath D (2019) Harnessing indirect training data for end-to-end automatic speech translation: Tricks of the trade. In: Proceedings of the 16th International Workshop on Spoken Language Translation (IWSLT)
2019
Closest in time.
Salesky E, Sperber M, Waibel A (2019) Fluent translations from disfluent speech in end-to-end speech translation. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers), Association for Computational Linguistics, Minneapolis, Minnesota, pp 2786–2792
2019
Closest in time.
Schneider F, Waibel A (2019) KIT’s submission to the IWSLT 2019 shared task on text translation. In: Proceedings of the 16th International Workshop on Spoken Language Translation
2019
Closest in time.
Sperber M, Neubig G, Niehues J, Waibel A (2019) Attention-passing models for robust and data-efficient end-to-end speech translation. Transactions of the Association for Computational Linguistics 7:313–325
2019
Closest in time.
Zhang P, Ge N, Chen B, Fan K (2019) Lattice transformer for speech translation. In: Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, Association for Computational Linguistics, Florence, Italy, pp 6475–6484
2019
Closest in time.
Graves A, Schmidhuber J (2005) Framewise phoneme classification with bidirectional LSTM networks. In: Proceedings. 2005 IEEE International Joint Conference on Neural Networks, 2005., IEEE, Montreal, Que., Canada, vol 4, pp 2047–2052
2052
Closest in time.
Xu K, Ba J, Kiros R, Cho K, Courville A, Salakhudinov R, Zemel R, Bengio Y (2015) Show, attend and tell: Neural image caption generation with visual attention. In: Proceedings of the 32nd International Conference on Machine Learning (ICML-15), JMLR Workshop and Conference Proceedings, pp 2048–2057
2057
Closest in time.