Fetching the paper…
Reading the bibliography…
Automatic speech recognition (ASR) models are frequently exposed to data distribution shifts in many real-world scenarios, leading to erroneous predictions.
Y. Grandvalet and Y. Bengio, “Semi-supervised learning by entropy minimization,” in NIPS , 2004
2004
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in ICML , 2006
2006
Earlier work this paper cites.
F. Seide, G. Li, X. Chen, and D. Yu, “Feature engineering in context-dependent deep neural networks for conversational speech transcription,” in ASRU , 2011
2011
Earlier work this paper cites.
K. Yao, D. Yu, F. Seide, H. Su, L. Deng, and Y. Gong, “Adaptation of context-dependent deep neural networks for automatic speech recognition,” in SLT , 2012
2012
Earlier work this paper cites.
A. Graves, “Sequence transduction with recurrent neural networks,” in ICML , 2012
2012
Earlier work this paper cites.
D. Yu, K. Yao, H. Su, G. Li, and F. Seide, “Kl-divergence regularized deep neural network adaptation for improved large vocabulary speech recognition,” in ICASSP , 2013
2013
Earlier work this paper cites.
A. Rousseau, P. Deléglise, Y. Esteve et al. , “Enhancing the ted-lium corpus with selected data for language modeling and more ted talks.” in LREC , 2014
2014
Earlier work this paper cites.
Y. Ganin and V. Lempitsky, “Unsupervised domain adaptation by backpropagation,” in ICML , 2015
2015
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in ICASSP , 2015
2015
Earlier work this paper cites.
J. Barker, R. Marxer, E. Vincent, and S. Watanabe, “The third ‘chime’speech separation and recognition challenge: Dataset, task and baselines,” in ASRU , 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” NeurIPS , 2017
2017
Earlier work this paper cites.
E. Tzeng, J. Hoffman, K. Saenko, and T. Darrell, “Adversarial discriminative domain adaptation,” in CVPR , 2017
2017
Earlier work this paper cites.
W.-N. Hsu, Y. Zhang, and J. Glass, “Unsupervised domain adaptation for robust speech recognition via variational autoencoder-based data augmentation,” in ASRU , 2017
2017
Earlier work this paper cites.
S. Sun, B. Zhang, L. Xie, and Y. Zhang, “An unsupervised deep domain adaptation approach for robust speech recognition,” in Neurocomputing , 2017
2017
Earlier work this paper cites.
M. Freitag and Y. Al-Onaizan, “Beam search strategies for neural machine translation,” in ACL , 2017
2017
Earlier work this paper cites.
C. Valentini-Botinhao et al. , “Noisy speech database for training speech enhancement algorithms and tts models,” University of Edinburgh. School of Informatics. Centre for Speech Technology Research (CSTR) , 2017
2017
Cited alongside, same era.
I. Loshchilov and F. Hutter, “Decoupled weight decay regularization,” in ICLR , 2017
2017
Cited alongside, same era.
J. Shen, R. Pang, R. J. Weiss, M. Schuster, N. Jaitly, Z. Yang, Z. Chen, Y. Zhang, Y. Wang, R. Skerrv-Ryan et al. , “Natural tts synthesis by conditioning wavenet on mel spectrogram predictions,” in ICASSP , 2018
2018
Cited alongside, same era.
S. Sun, C.-F. Yeh, M.-Y. Hwang, M. Ostendorf, and L. Xie, “Domain adversarial training for accented speech recognition,” in ICASSP , 2018
2018
Cited alongside, same era.
V. Manohar, P. Ghahremani, D. Povey, and S. Khudanpur, “A teacher-student learning approach for unsupervised domain adaptation of sequence-trained asr models,” in SLT , 2018
W. Hou, J. Wang, X. Tan, T. Qin, and T. Shinozaki, “Cross-domain speech recognition with unsupervised character-level distribution matching,” in INTERSPEECH , 2021
2021
Later among the works it cites.
D. Wang, E. Shelhamer, S. Liu, B. Olshausen, and T. Darrell, “TENT: Fully test-time adaptation by entropy minimization,” in ICLR , 2021
2021
Later among the works it cites.
Y. Liu, P. Kothari, B. van Delft, B. Bellot-Gurlet, T. Mordan, and A. Alahi, “TTT++: When does self-supervised test-time training fail or thrive?” in NeurIPS , 2021
2021
Later among the works it cites.
F. Fleuret et al. , “Test time adaptation through perturbation robustness,” in NeurIPS DistShift Workshop , 2021
2021
Later among the works it cites.
M. Burchi and V. Vielzeuf, “Efficient conformer: Progressive downsampling and grouped attention for automatic speech recognition,” in ASRU , 2021
2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
G. Zhao, S. Sonsaat, A. Silpachai, I. Lucic, E. Chukharev-Hudilainen, J. Levis, and R. Gutierrez-Osuna, “L2-arctic: A non-native english speech corpus,” in INTERSPEECH , 2018
2018
Cited alongside, same era.
C. K. Reddy, E. Beyrami, J. Pool, R. Cutler, S. Srinivasan, and J. Gehrke, “A scalable noisy speech dataset and online subjective test framework,” in INTERSPEECH , 2019
2019
Cited alongside, same era.
N. R. Tom B. Brown, Benjamin Mann and M. S. et al., “Language models are few-shot learners,” in NeurIPS , 2020
2020
Cited alongside, same era.
A. Baevski, Y. Zhou, A. Mohamed, and M. Auli, “wav2vec 2.0: A framework for self-supervised learning of speech representations,” in NeurIPS , 2020
2020
Cited alongside, same era.
Y. Sun, X. Wang, Z. Liu, J. Miller, A. Efros, and M. Hardt, “Test-time training with self-supervision for generalization under distribution shifts,” in ICML , 2020
2020
Cited alongside, same era.
Y. Jin, X. Wang, M. Long, and J. Wang, “Minimum class confusion for versatile domain adaptation,” in ECCV , 2020
2020
Cited alongside, same era.
J. Chen, V. Shah, and A. Kyrillidis, “Negative sampling in semi-supervised learning,” in ICML , 2020
2020
Cited alongside, same era.
Later among the works it cites.
S. Khurana, N. Moritz, T. Hori, and J. Le Roux, “Unsupervised domain adaptation for speech recognition via uncertainty driven self-training,” in ICASSP , 2021
2021
Later among the works it cites.
G.-T. Lin, S.-W. Li, and H.-y. Lee, “Listen, adapt, better wer: Source-free single-utterance test-time adaptation for automatic speech recognition,” in INTERSPEECH , 2022
2022
Later among the works it cites.
M. Zhang, S. Levine, and C. Finn, “MEMO: Test time robustness via adaptation and augmentation,” in NeurIPS , 2022
2022
Later among the works it cites.
D. Chen, D. Wang, T. Darrell, and S. Ebrahimi, “Contrastive test-time adaptation,” in CVPR , 2022
2022
Later among the works it cites.
Q. Wang, O. Fink, L. Van Gool, and D. Dai, “Continual test-time domain adaptation,” in CVPR , 2022
2022
Later among the works it cites.
A. Bartler, A. Bühler, F. Wiewel, M. Döbler, and B. Yang, “MT3: Meta test-time training for self-supervised test-time adaption,” in AISTATS , 2022
2022
Later among the works it cites.
S. Goyal, M. Sun, A. Raghunathan, and Z. Kolter, “Test-time adaptation via conjugate pseudo-labels,” in NeurIPS , 2022
2022
Later among the works it cites.
A. Radford, J. W. Kim, T. Xu, G. Brockman, C. McLeavey, and I. Sutskever, “Robust speech recognition via large-scale weak supervision,” in ICML , 2022
2022
Later among the works it cites.
J. Gao, J. Zhang, X. Liu, T. Darrell, E. Shelhamer, and D. Wang, “Back to the source: Diffusion-driven test-time adaptation,” in ICML DyNN Workshop , 2022
2022
Later among the works it cites.
S. Niu, J. Wu, Y. Zhang, Z. Wen, Y. Chen, P. Zhao, and M. Tan, “Towards stable test-time adaptation in dynamic wild world,” in ICLR , 2023
2023
Closest in time.