Fetching the paper…
Reading the bibliography…
We propose a simple and effective cross-lingual transfer learning method to adapt monolingual wav2vec-2.0 models for Automatic Speech Recognition (ASR) in resource-scarce languages.
H. Scudder, “Probability of error of some adaptive pattern-recognition machines,” IEEE Trans. Inf. Theory , vol. 11, no. 3, pp. 363–371, 1965
1965
Earlier work this paper cites.
V. I. Levenshtein, “Binary Codes Capable of Correcting Deletions, Insertions and Reversals,” Soviet Physics Doklady , vol. 10, p. 707, Feb. 1966
1966
Earlier work this paper cites.
J. Schmidhuber, “Making the world differentiable: On using self-supervised fully recurrent neural networks for dynamic reinforcement learning and planning in non-stationary environments,” Tech. Rep., 1990
1990
Earlier work this paper cites.
V. R. DeSa, “Learning classification with unlabeled data,” in Proceedings of the 6th International Conference on Neural Information Processing Systems , ser. NIPS’93. San Francisco, CA, USA: Morgan Kaufmann Publishers Inc., 1993, p. 112–119
1993
Earlier work this paper cites.
2012
Earlier work this paper cites.
D. Wang and T. F. Zheng, “Transfer learning for speech and language processing,” 2015
2015
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “LibriSpeech: An ASR corpus based on public domain audio books,” in Proc. ICASSP , Apr. 2015
2015
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” 2019
2019
Earlier work this paper cites.
A. van den Oord, Y. Li, and O. Vinyals, “Representation learning with contrastive predictive coding,” 2019
2019
Earlier work this paper cites.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
M. Ott, S. Edunov, A. Baevski, A. Fan, S. Gross, N. Ng, D. Grangier, and M. Auli, “fairseq: A fast, extensible toolkit for sequence modeling,” 2019
2019
Cited alongside, same era.
A. Ali, P. Bell, J. Glass, Y. Messaoui, H. Mubarak, S. Renals, and Y. Zhang, “The mgb-2 challenge: Arabic multi-dialect broadcast media recognition,” 2019
2019
Cited alongside, same era.
2020
Later among the works it cites.
J. Kahn, M. Rivière, W. Zheng, E. Kharitonov, Q. Xu, P. E. Mazaré, J. Karadayi, V. Liptchinsky, R. Collobert, C. Fuegen, T. Likhomanenko, G. Synnaeve, A. Joulin, A. Mohamed, and E. Dupoux, “Libri-light: A benchmark for asr with limited or no supervision,” in ICASSP 2020 - 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2020, pp. 7669–7673, https://github.com/facebookresearch/libri-light
2020
Later among the works it cites.
S. Watanabe, F. Boyer, X. Chang, P. Guo, T. Hayashi, Y. Higuchi, T. Hori, W.-C. Huang, H. Inaguma, N. Kamo, S. Karita, C. Li, J. Shi, A. S. Subramanian, and W. Zhang, “The 2020 espnet update: new features, broadened applications, performance improvements, and future plans,” 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton, “A simple framework for contrastive learning of visual representations,” 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
A. Conneau, A. Baevski, R. Collobert, A. Mohamed, and M. Auli, “Unsupervised cross-lingual representation learning for speech recognition,” 2020
2020
Cited alongside, same era.
J. Kahn, A. Lee, and A. Hannun, “Self-training for end-to-end speech recognition,” ICASSP 2020 - 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , May 2020. [Online]. Available: http://dx.doi.org/10.1109/ICASSP40776.2020.9054295
2020
Cited alongside, same era.
M. Rivière, A. Joulin, P.-E. Mazaré, and E. Dupoux, “Unsupervised pretraining transfers well across languages,” 2020
2020
Later among the works it cites.
J. Pfeiffer, I. Vulić, I. Gurevych, and S. Ruder, “Mad-x: An adapter-based framework for multi-task cross-lingual transfer,” 2020
2020
Later among the works it cites.
S. Khurana, N. Moritz, T. Hori, and J. L. Roux, “Unsupervised domain adaptation for speech recognition via uncertainty driven self-training,” Proc. ICASSP , 2021
2021
Closest in time.
W.-N. Hsu, A. Sriram, A. Baevski, T. Likhomanenko, Q. Xu, V. Pratap, J. Kahn, A. Lee, R. Collobert, G. Synnaeve, and M. Auli, “Robust wav2vec 2.0: Analyzing domain shift in self-supervised pre-training,” 2021
2021
Closest in time.
S. Kessler, B. Thomas, and S. Karout, “Continual-wav2vec2: an application of continual learning for self-supervised automatic speech recognition,” 2021
2021
Closest in time.
W. Hou, H. Zhu, Y. Wang, J. Wang, T. Qin, R. Xu, and T. Shinozaki, “Exploiting adapters for cross-lingual low-resource speech recognition,” 2021
2021
Closest in time.