Fetching the paper…
Reading the bibliography…
We present the Multilingual TEDx corpus, built to support speech recognition (ASR) and speech translation (ST) research across many non-English source languages.
F. W. Stentiford and M. G. Steer, “Machine translation of speech,”
1988
Earlier work this paper cites.
A. Waibel, A. N. Jain, A. E. McNair
1991
Earlier work this paper cites.
J. C. Wells, “Computer-coding the IPA: A proposed extension of SAMPA,” 1995/2000
2000
Earlier work this paper cites.
A. Stolcke, “SRILM - an extensible language modeling toolkit,” in
2002
Earlier work this paper cites.
E. Matusov, G. Leusch, O. Bender, and H. Ney, “Evaluating machine translation output with automatic sentence segmentation,” in
2005
Earlier work this paper cites.
T. Kiss and J. Strunk, “Unsupervised multilingual sentence boundary detection,”
2006
Earlier work this paper cites.
2010
Earlier work this paper cites.
F. Braune and A. Fraser, “Improved unsupervised sentence alignment for symmetrical and asymmetrical parallel corpora,” in
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne
2011
Earlier work this paper cites.
M. Cettolo, C. Girardi, and M. Federico, “Wit3: Web inventory of transcribed and translated talks,” in
2012
Earlier work this paper cites.
M. Post, G. Kumar, A. Lopez, D. Karakos, C. Callison-Burch, and S. Khudanpur, “Improved speech-to-text translation with the Fisher and Callhome Spanish–English speech translation corpus,” 2013
2013
Earlier work this paper cites.
M. Cettolo, J. Niehues, S. Stüker, L. Bentivogli, and M. Federico, “Report on the 11th iwslt evaluation campaign,” in
2014
Earlier work this paper cites.
J. R. Novak, N. Minematsu, and K. Hirose, “Phonetisaurus: Exploring grapheme-to-phoneme conversion with joint n-gram models in the WFST framework,”
2016
Earlier work this paper cites.
D. Povey, V. Peddinti, D. Galvez
2016
Earlier work this paper cites.
T.-L. Ha, J. Niehues, and A. Waibel, “Toward multilingual neural machine translation with universal encoder and decoder,” in
2016
Earlier work this paper cites.
L. Duong, A. Anastasopoulos, D. Chiang, S. Bird, and T. Cohn, “An attentional model for speech translation without transcription,” in
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Cited alongside, same era.
M. Cettolo, M. Federico
2017
Cited alongside, same era.
R. J. Weiss, J. Chorowski, N. Jaitly, Y. Wu, and Z. Chen, “Sequence-to-sequence models can directly translate foreign speech,”
2017
Cited alongside, same era.
Y. Zhang, W. Chan, and N. Jaitly, “Very deep convolutional networks for end-to-end speech recognition,”
2017
Cited alongside, same era.
A. Kocabiyikoglu, L. Besacier, and O. Kraif, “Augmenting Librispeech with French Translations: A Multimodal Corpus for Direct Speech Translation Evaluation,” in
2018
Cited alongside, same era.
M. Post, “A call for clarity in reporting BLEU scores,” in
2019
Later among the works it cites.
J. Niehues, R. Cattoni, S. Stüker, M. Negri, M. Turchi, E. Salesky, R. Sanabria, L. Barrault, L. Specia, and M. Federico, “The IWSLT 2019 evaluation campaign,” in
2019
Later among the works it cites.
J. Pino, L. Puzon, J. Gu, X. Ma, A. D. McCarthy, and D. Gopinath, “Harnessing indirect training data for end-to-end automatic speech translation: Tricks of the trade,” in
2019
Later among the works it cites.
D. S. Park, W. Chan, Y. Zhang, C.-C. Chiu, B. Zoph, E. D. Cubuk, and Q. V. Le, “Specaugment: A simple data augmentation method for automatic speech recognition,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
D. R. Mortensen, S. Dalmia, and P. Littell, “Epitran: Precision G2P for many languages,” in
2018
Cited alongside, same era.
T. Kudo and J. Richardson, “SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing,” in
2018
Cited alongside, same era.
M. A. Di Gangi, M. Negri, and M. Turchi, “Adapting transformer to end-to-end spoken language translation,” in
2019
Cited alongside, same era.
H. Inaguma, K. Duh, T. Kawahara, and S. Watanabe, “Multilingual end-to-end speech translation,” in
2019
Cited alongside, same era.
M. Sperber, G. Neubig, J. Niehues, and A. Waibel, “Attention-passing models for robust and data-efficient end-to-end speech translation,”
2019
Cited alongside, same era.
M. A. Di Gangi, R. Cattoni, L. Bentivogli, M. Negri, and M. Turchi, “MuST-C: a Multilingual Speech Translation Corpus,” in
2019
Cited alongside, same era.
2019
Later among the works it cites.
M. Z. Boito, W. N. Havard, M. Garnerin
2020
Later among the works it cites.
J. Iranzo-Sánchez, J. A. Silvestre-Cerdà, J. Jorge
2020
Later among the works it cites.
C. Wang, J. Pino, A. Wu, and J. Gu, “CoVoST: A diverse multilingual speech-to-text translation corpus,” in
2020
Later among the works it cites.
C. Wang, A. Wu, and J. Pino, “CoVoST 2: A massively multilingual speech-to-text translation corpus,” 2020
2020
Later among the works it cites.
R. Ardila, M. Branson
2020
Later among the works it cites.
J. L. Lee, L. F. Ashby
2020
Later among the works it cites.
C. Wang, Y. Tang, X. Ma, A. Wu, D. Okhonko, and J. Pino, “fairseq S2T: Fast speech-to-text modeling with fairseq,” in
2020
Later among the works it cites.
A. Fan, S. Bhosale, H. Schwenk
2020
Later among the works it cites.
E. Ansari, N. Bach, O. Bojar
2020
Later among the works it cites.
E. Salesky and A. W. Black, “Phone features improve speech translation,”
2020
Later among the works it cites.
N.-Q. Pham, T.-L. Ha, T.-N. Nguyen, T.-S. Nguyen, E. Salesky, S. Stueker, J. Niehues, and A. Waibel, “Relative positional encoding for speech recognition and direct translation,” 2020
2020
Later among the works it cites.
R. Cattoni, M. A. Di Gangi, L. Bentivogli, M. Negri, and M. Turchi, “Must-c: A multilingual corpus for end-to-end speech translation,”
2021
Closest in time.