Attention is all you need
Original
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Speech2vec: A sequence-to-sequence framework for learning word embeddings from speech
Original
Yu-An Chung and James Glass. 2018 · 2018
Later among the works it cites.
Colorless green recurrent networks dream hierarchically
Kristina Gulordava, Piotr Bojanowski, Edouard Grave, Tal Linzen, and Marco Baroni. 2018 · 2018
Later among the works it cites.
Representation learning with contrastive predictive coding
Original
Aäron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Later among the works it cites.
Can LSTM learn to capture agreement? the case of basque
Shauli Ravfogel, Francis M Tyers, and Yoav Goldberg. 2018 · 2018
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
The zero resource speech challenge 2019: Tts without t
Original
Ewan Dunbar, Robin Algayres, Julien Karadayi, Mathieu Bernard, Juan Benjumea, Xuan-Nga Cao, Lucie Miskic, Charlotte Dugrain, Lucas Ondel, Alan W. Black, Laurent Besacier, Sakriani Sakti, and Emmanuel Dupoux. 2019 · 2019
Later among the works it cites.
Tabula nearly rasa: Probing the linguistic knowledge of character-level neural language models trained on unsegmented text
Original
Michael Hahn and Marco Baroni. 2019 · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
The zero resource speech challenge 2020: Discovering discrete subword and word units
Ewan Dunbar, Julien Karadayi, Mathieu Bernard, Xuan-Nga Cao, Robin Algayres, Lucas Ondel, Laurent Besacier, Sakti Sakriani, and Emmanuel Dupoux. 2020 · 2020
Closest in time.
Libri-light: A benchmark for asr with limited or no supervision
J. Kahn, M. Riviere, W. Zheng, E. Kharitonov, Q. Xu, P.E. Mazare, J. Karadayi, V. Liptchinsky, R. Collobert, C. Fuegen, and et al. 2020 · 2020
Closest in time.
Learning robust and multilingual speech representations
Original
K. Kawakami, L. Wang, C. Dyer, P. Blunsom, and A. van den Oord. 2020 · 2020
Closest in time.
Unsupervised pretraining transfers well across languages
Original
Morgane Rivière, Armand Joulin, Pierre-Emmanuel Mazaré, and Emmanuel Dupoux. 2020 · 2020
Closest in time.
Masked language model scoring
Julian Salazar, Davis Liang, Toan Q. Nguyen, and Katrin Kirchhoff. 2020 · 2020
Closest in time.
Unsupervised pre-training of bidirectional speech encoders via masked reconstruction
Weiran Wang, Qingming Tang, and Karen Livescu. 2020 · 2020
Closest in time.