Fetching the paper…
Reading the bibliography…
Speech representations learned from Self-supervised learning (SSL) models can benefit various speech processing tasks.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An ASR corpus based on public domain audio books,” in ICASSP . IEEE, 2015, pp. 5206–5210
2015
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NIPS , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
P. Yang, X. Sun, W. Li, S. Ma, W. Wu, and H. Wang, “SGM: sequence generation model for multi-label classification,” in COLING . Association for Computational Linguistics, 2018, pp. 3915–3926
2018
Earlier work this paper cites.
G. F. Elsayed, I. J. Goodfellow, and J. Sohl-Dickstein, “Adversarial reprogramming of neural networks,” in ICLR (Poster) . OpenReview.net, 2019
2019
Earlier work this paper cites.
L. Lugosch, M. Ravanelli, P. Ignoto, V. S. Tomar, and Y. Bengio, “Speech model pre-training for end-to-end spoken language understanding,” in INTERSPEECH . ISCA, 2019, pp. 814–818
2019
Earlier work this paper cites.
M. Ott, S. Edunov, A. Baevski, A. Fan, S. Gross, N. Ng, D. Grangier, and M. Auli, “fairseq: A fast, extensible toolkit for sequence modeling,” in Proceedings of NAACL-HLT 2019: Demonstrations , 2019
2019
Earlier work this paper cites.
L. Dong, N. Yang, W. Wang, F. Wei, X. Liu, Y. Wang, J. Gao, M. Zhou, and H. Hon, “Unified language model pre-training for natural language understanding and generation,” in NeurIPS , 2019, pp. 13 042–13 054
2019
Earlier work this paper cites.
A. Baevski, Y. Zhou, A. Mohamed, and M. Auli, “wav2vec 2.0: A framework for self-supervised learning of speech representations,” in NeurIPS , 2020
2020
Earlier work this paper cites.
A. Baevski and A. Mohamed, “Effectiveness of self-supervised pre-training for ASR,” in ICASSP . IEEE, 2020, pp. 7694–7698
2020
Earlier work this paper cites.
T. Shin, Y. Razeghi, R. L. L. IV, E. Wallace, and S. Singh, “Autoprompt: Eliciting knowledge from language models with automatically generated prompts,” in EMNLP (1) . Association for Computational Linguistics, 2020, pp. 4222–4235
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” J. Mach. Learn. Res. , vol. 21, pp. 140:1–140:67, 2020
2020
Earlier work this paper cites.
H. Bao, L. Dong, F. Wei, W. Wang, N. Yang, X. Liu, Y. Wang, J. Gao, S. Piao, M. Zhou et al. , “Unilmv2: Pseudo-masked language models for unified language model pre-training,” in International Conference on Machine Learning . PMLR, 2020, pp. 642–652
2020
Cited alongside, same era.
M. Lewis, Y. Liu, N. Goyal, M. Ghazvininejad, A. Mohamed, O. Levy, V. Stoyanov, and L. Zettlemoyer, “BART: denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension,” in ACL . Association for Computational Linguistics, 2020, pp. 7871–7880
2020
Cited alongside, same era.
W. Hsu, B. Bolte, Y. H. Tsai, K. Lakhotia, R. Salakhutdinov, and A. Mohamed, “Hubert: Self-supervised speech representation learning by masked prediction of hidden units,” IEEE ACM Trans. Audio Speech Lang. Process. , vol. 29, pp. 3451–3460, 2021
2021
Cited alongside, same era.
2021
Later among the works it cites.
B. Lester, R. Al-Rfou, and N. Constant, “The power of scale for parameter-efficient prompt tuning,” in EMNLP (1) . Association for Computational Linguistics, 2021, pp. 3045–3059
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
2021
Cited alongside, same era.
C. Lai, Y. Chuang, H. Lee, S. Li, and J. R. Glass, “Semi-supervised spoken language understanding via self-supervised speech and language model pretraining,” in ICASSP . IEEE, 2021, pp. 7468–7472
2021
Cited alongside, same era.
J. Lin, Y. Y. Lin, C. Chien, and H. Lee, “S2VC: A framework for any-to-any voice conversion with self-supervised pretrained representations,” in Interspeech . ISCA, 2021, pp. 836–840
2021
Cited alongside, same era.
A. Baevski, W.-N. Hsu, A. Conneau, and M. Auli, “Unsupervised speech recognition,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Cited alongside, same era.
S. Yang, P. Chi, Y. Chuang, C. J. Lai, K. Lakhotia, Y. Y. Lin, A. T. Liu, J. Shi, X. Chang, G. Lin, T. Huang, W. Tseng, K. Lee, D. Liu, Z. Huang, S. Dong, S. Li, S. Watanabe, A. Mohamed, and H. Lee, “SUPERB: speech processing universal performance benchmark,” in Interspeech . ISCA, 2021, pp. 1194–1198
2021
Cited alongside, same era.
C.-I. J. Lai, Y. Zhang, A. H. Liu, S. Chang, Y.-L. Liao, Y.-S. Chuang, K. Qian, S. Khurana, D. Cox, and J. Glass, “Parp: Prune, adjust and re-prune for self-supervised speech recognition,” Advances in Neural Information Processing Systems , vol. 34, 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
X. L. Li and P. Liang, “Prefix-tuning: Optimizing continuous prompts for generation,” in ACL/IJCNLP (1) . Association for Computational Linguistics, 2021, pp. 4582–4597
2021
Cited alongside, same era.
T. Schick and H. Schütze, “It’s not just size that matters: Small language models are also few-shot learners,” in NAACL-HLT . Association for Computational Linguistics, 2021, pp. 2339–2352
2021
Later among the works it cites.
C. H. Yang, Y. Tsai, and P. Chen, “Voice2series: Reprogramming acoustic models for time series classification,” in ICML , ser. Proceedings of Machine Learning Research, vol. 139. PMLR, 2021, pp. 11 808–11 819
2021
Later among the works it cites.
2021
Later among the works it cites.
T. Schick and H. Schütze, “Exploiting cloze-questions for few-shot text classification and natural language inference,” in EACL . Association for Computational Linguistics, 2021, pp. 255–269
2021
Later among the works it cites.
T. Gao, A. Fisch, and D. Chen, “Making pre-trained language models better few-shot learners,” in ACL/IJCNLP (1) . Association for Computational Linguistics, 2021, pp. 3816–3830
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
H.-S. Tsai, H.-J. Chang, W.-C. Huang, Z. Huang, K. Lakhotia, S. wen Yang, S. Dong, A. T. Liu, C.-I. J. Lai, J. Shi, X. Chang, P. Hall, H.-J. Chen, S.-W. Li, S. Watanabe, A. rahman Mohamed, and H. yi Lee, “Superb-sg: Enhanced speech processing universal performance benchmark for semantic and generative capabilities,” 2022
2022
Closest in time.
C. Qin and S. Joty, “LFPT5: A unified framework for lifelong few-shot language learning based on prompt tuning of t5,” in International Conference on Learning Representations , 2022
2022
Closest in time.