2020

Learning Robust and Multilingual Speech Representations

Kawakami, Kazuya, Wang, Luyu, Dyer, Chris et al.

Understand

Unsupervised speech representation learning has shown remarkable success at finding representations that correlate with phonetic structures and improve downstream speech recognition performance.

  • However, most research has been focused on evaluating the representations in terms of their ability to improve the performance of speech recognition systems on read English (e.g.
  • Wall Street Journal and LibriSpeech).
  • This evaluation methodology overlooks two important desiderata that speech representations should have: robustness to domain shifts and transferability to other languages.

Reading the bibliography…