Fetching the paper…
Reading the bibliography…
Despite improvements to the generalization performance of automated speech recognition (ASR) models, specializing ASR models for downstream tasks remains a challenging task, primarily due to reduced data availability (necessitating increased data collection), and rapidly shifting data distributions (requiring more frequent model fine-tuning).
T. Pápai, S. Ghosh, and H. Kautz, “Combining subjective probabilities and data in training Markov Logic Networks,” vol. 7523, 09 2012, pp. 90–105
2012
Earlier work this paper cites.
T. Ge, K. He, Q. Ke, and J. Sun, “Optimized product quantization,”
2013
Earlier work this paper cites.
J. Pennington, R. Socher, and C. D. Manning, “Glove: Global vectors for word representation,” in
2014
Earlier work this paper cites.
R. Lin, S. Liu, M. Yang, M. Li, M. Zhou, and S. Li, “Hierarchical recurrent neural network for document modeling,” in
2015
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in
2015
Earlier work this paper cites.
B. Liu and I. Lane, “Dialog context language modeling with recurrent neural networks,” in
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,”
2017
Earlier work this paper cites.
Q. Wu, C. Shen, P. Wang, A. Dick, and A. Van Den Hengel, “Image captioning and visual question answering based on attributes and external knowledge,”
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Gandhe, A. Rastrow, and B. Hoffmeister, “Scalable language model adaptation for spoken dialogue systems,” in
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Jaech and M. Ostendorf, “Personalized language model for query auto-completion,”
2018
Earlier work this paper cites.
S. Kim and F. Metze, “Dialog-context aware end-to-end speech recognition,” in
2018
Earlier work this paper cites.
I. Williams, A. Kannan, P. S. Aleksic, D. Rybach, and T. N. Sainath, “Contextual speech recognition in end-to-end neural network systems using beam search.” in
2018
Earlier work this paper cites.
2018
Cited alongside, same era.
Y. A. Malkov and D. A. Yashunin, “Efficient and robust approximate nearest neighbor search using hierarchical navigable small world graphs,”
2018
Cited alongside, same era.
2018
Cited alongside, same era.
D. Zhao, T. N. Sainath, D. Rybach, P. Rondon, D. Bhatia, B. Li, and R. Pang, “Shallow-fusion end-to-end contextual biasing.” in
2019
Cited alongside, same era.
Z. Chen, M. Jain, Y. Wang, M. L. Seltzer, and C. Fuegen, “Joint grapheme and phoneme embeddings for contextual end-to-end asr.” in
X. Xie, J. Niu, X. Liu, Z. Chen, S. Tang, and S. Yu, “A survey on incorporating domain knowledge into deep learning for medical image analysis,”
2021
Later among the works it cites.
T. Tran, V. Le, H. Le, and T. M. Le, “From deep learning to deep reasoning,” in
2021
Later among the works it cites.
N. Thakur, N. Reimers, J. Daxenberger, and I. Gurevych, “Augmented SBERT: Data augmentation method for improving bi-encoders for pairwise sentence scoring tasks,” in
2021
Later among the works it cites.
S. Dingliwa, A. Shenoy, S. Bodapati, A. Gandhe, R. T. Gadde, and K. Kirchhoff, “Domain prompts: Towards memory and compute efficient domain adaptation of asr systems,” in
2022
Later among the works it cites.
D. Baby, P. D’Alterio, and V. Mendelev, “Incremental learning for rnn-transducer based speech recognition models,” in
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
K. Marino, M. Rastegari, A. Farhadi, and R. Mottaghi, “Ok-vqa: A visual question answering benchmark requiring external knowledge,” in
2019
Cited alongside, same era.
2019
Cited alongside, same era.
J. Johnson, M. Douze, and H. Jégou, “Billion-scale similarity search with gpus,”
2019
Cited alongside, same era.
2020
Cited alongside, same era.
A. Gulati, J. Qin, C.-C. Chiu, N. Parmar, Y. Zhang, J. Yu, W. Han, S. Wang, Z. Zhang, Y. Wu
2020
Cited alongside, same era.
K. Zhao, H. D. Nguyen, A. Jain, N. Susanj, A. Mouchtaris, L. Gupta, and M. Zhao, “Knowledge distillation via module replacing for automatic speech recognition with recurrent neural network transducer,” in
2022
Later among the works it cites.
S. Borgeaud, A. Mensch, J. Hoffmann, T. Cai, E. Rutherford, K. Millican, G. B. Van Den Driessche, J.-B. Lespiau, B. Damoc, A. Clark
2022
Later among the works it cites.
Y. Wu, M. N. Rabe, D. Hutchins, and C. Szegedy, “Memorizing transformers,”
2022
Later among the works it cites.
2022
Later among the works it cites.
T. Munkhdalai, K. C. Sim, A. Chandorkar, F. Gao, M. Chua, T. Strohman, and F. Beaufays, “Fast contextual adaptation with neural associative memory for on-device personalized speech recognition,” in
2022
Later among the works it cites.
K. M. Sathyendra, T. Muniyappa, F.-J. Chang, J. Liu, J. Su, G. P. Strimel, A. Mouchtaris, and S. Kunzmann, “Contextual adapters for personalized speech recognition in neural transducers,” in
2022
Later among the works it cites.
Z. Tang, S. Gu, J. Bao, D. Chen, and F. Wen, “Improved vector quantized diffusion models,”
2022
Later among the works it cites.
A. Goyal, A. Friesen, A. Banino, T. Weber, N. R. Ke, A. P. Badia, A. Guez, M. Mirza, P. C. Humphreys, K. Konyushova
2022
Later among the works it cites.