Fetching the paper…
Reading the bibliography…
Recently, representation learning for text and speech has successfully improved many language related tasks.
“cloze procedure”: A new tool for measuring readability
Taylor, W. L · 1953
Earlier work this paper cites.
Europarl: A parallel corpus for statistical machine translation
Koehn, P · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A., Fernández, S., Gomez, F., and Schmidhuber, J · 2006
Earlier work this paper cites.
Tera: Self-supervised learning of transformer encoder representation for speech
Liu, A. T., Li, S.-W., and Lee, H.-y · 2007
Earlier work this paper cites.
Bridging the modality gap for speech-to-text translation
Liu, Y., Zhu, J., Zhang, J., and Zong, C · 2010
Earlier work this paper cites.
The kaldi speech recognition toolkit
Povey, D., Ghoshal, A., Boulianne, G., Goel, N., Hannemann, M., Qian, Y., Schwarz, P., and Stemmer, G · 2011
Earlier work this paper cites.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing
Kudo, T. and Richardson, J · 2012
Earlier work this paper cites.
Librispeech: an asr corpus based on public domain audio books
Panayotov, V., Chen, G., Povey, D., and Khudanpur, S · 2015
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Rajpurkar, P., Zhang, J., Lopyrev, K., and Liang, P · 2016
Earlier work this paper cites.
Neural lattice-to-sequence models for uncertain inputs
Sperber, M., Neubig, G., Niehues, J., and Waibel, A · 2017
Earlier work this paper cites.
Sequence-to-sequence models can directly translate foreign speech
Weiss, R. J., Chorowski, J., Jaitly, N., Wu, Y., and Chen, Z · 2017
Earlier work this paper cites.
End-to-end automatic speech translation of audiobooks
Bérard, A., Besacier, L., Kocabiyikoglu, A. C., and Pietquin, O · 2018
Earlier work this paper cites.
Universal language model fine-tuning for text classification
Howard, J. and Ruder, S · 2018
Cited alongside, same era.
Augmenting librispeech with french translations: A multimodal corpus for direct speech translation evaluation
Kocabiyikoglu, A. C., Besacier, L., and Kraif, O · 2018
Cited alongside, same era.
Deep contextualized word representations
Peters, M. E., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., and Zettlemoyer, L · 2018
Cited alongside, same era.
Improving language understanding by generative pre-training
Radford, A., Narasimhan, K., Salimans, T., and Sutskever, I · 2018
Cited alongside, same era.
Glue: A multi-task benchmark and analysis platform for natural language understanding
Wang, A., Singh, A., Michael, J., Hill, F., Levy, O., and Bowman, S. R · 2018
Ernie: Enhanced representation through knowledge integration
Sun, Y., Wang, S., Li, Y., Feng, S., Chen, X., Zhang, H., Tian, X., Zhu, D., Tian, H., and Wu, H · 2019
Later among the works it cites.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Baevski, A., Zhou, H., Mohamed, A., and Auli, M · 2020
Later among the works it cites.
Mam: Masked acoustic modeling for end-to-end speech-to-text translation
Chen, J., Ma, M., Zheng, R., and Huang, L · 2020
Later among the works it cites.
Dong, Q., Ye, R., Wang, M., Zhou, H., Xu, S., Xu, B., and Li, L · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Chuang, Y.-S., Liu, C.-L., Lee, H.-y., and Lee, L.-s · 2019
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2019
Cited alongside, same era.
MuST-C: a Multilingual Speech Translation Corpus
Di Gangi, M. A., Cattoni, R., Bentivogli, L., Negri, M., and Turchi, M · 2019
Cited alongside, same era.
Leveraging weakly supervised data to improve end-to-end speech-to-text translation
Jia, Y., Johnson, M., Macherey, W., Weiss, R. J., Cao, Y., Chiu, C.-C., Ari, N., Laurenzo, S., and Wu, Y · 2019
Cited alongside, same era.
Cross-lingual language model pretraining
Lample, G. and Conneau, A · 2019
Cited alongside, same era.
End-to-end speech translation with knowledge distillation
Liu, Y., Xiong, H., He, Z., Zhang, J., Wu, H., Wang, H., and Zong, C · 2019
Cited alongside, same era.
Inaguma, H., Kiyono, S., Duh, K., Karita, S., Soplin, N. E. Y., Hayashi, T., and Watanabe, S · 2020
Later among the works it cites.
Libri-light: A benchmark for asr with limited or no supervision
Kahn, J., Rivière, M., Zheng, W., Kharitonov, E., Xu, Q., Mazaré, P. E., Karadayi, J., Liptchinsky, V., Collobert, R., Fuegen, C., Likhomanenko, T., Synnaeve, G., Joulin, A., Mohamed, A., and Dupoux, E · 2020
Later among the works it cites.
Dual-decoder transformer for joint automatic speech recognition and multilingual speech translation
Le, H., Pino, J., Wang, C., Gu, J., Schwab, D., and Besacier, L · 2020
Later among the works it cites.
Self-training for end-to-end speech translation
Pino, J., Xu, Q., Ma, X., Dousti, M. J., and Tang, Y · 2020
Later among the works it cites.
Fluent and low-latency simultaneous speech-to-speech translation with self-adaptive training
Zheng, R., Ma, M., Zheng, B., Liu, K., Yuan, J., Church, K., and Huang, L · 2020
Later among the works it cites.
Direct simultaneous speech-to-text translation assisted by synchronized streaming asr
Chen, J., Ma, M., Zheng, R., and Huang, L · 2021
Closest in time.
Consecutive decoding for speech-to-text translation
Dong, Q., Wang, M., Zhou, H., Xu, S., Xu, B., and Li, L · 2021
Closest in time.