Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have been applied in the speech domain, often incurring a performance drop due to misaligned between speech and language representations.
1907
Earlier work this paper cites.
1909
Earlier work this paper cites.
1911
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks,” in Proc. ICML . Association for Computing Machinery, 2006
2006
Earlier work this paper cites.
A. Graves, “Sequence transduction with recurrent neural networks,” CoRR , 2012
2012
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Radford, K. Narasimhan, T. Salimans, and I. Sutskever, “Improving language understanding by generative pre-training,” 2018
2018
Earlier work this paper cites.
C.-S. Wu, A. Madotto, E. Hosseini-Asl, C. Xiong, R. Socher, and P. Fung, “Transferable multi-domain state generator for task-oriented dialogue systems,” in Proc. ACL , Jul. 2019, pp. 808–819
2019
Earlier work this paper cites.
K. Guu, K. Lee, Z. Tung, P. Pasupat, and M.-W. Chang, “Retrieval augmented language model pre-training,” in Proc. ICML , vol. 119. PMLR, 2020, pp. 3929–3938
2020
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel, S. Riedel, and D. Kiela, “Retrieval-augmented generation for knowledge-intensive nlp tasks,” in Proc. NIPS , vol. 33, 2020, pp. 9459–9474
2020
Earlier work this paper cites.
A. Rastogi, X. Zang, S. Sunkara, R. Gupta, and P. Khaitan, “Towards scalable multi-domain conversational agents: The schema-guided dialogue dataset,” Proc. AAAI Conference on Artificial Intelligence , 2020
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” The Journal of Machine Learning Research , 2020
2020
Cited alongside, same era.
R. Guo, P. Sun, E. Lindgren, Q. Geng, D. Simcha, F. Chern, and S. Kumar, “Accelerating large-scale inference with anisotropic vector quantization,” in International Conference on Machine Learning . PMLR, 2020, pp. 3887–3896
2020
Cited alongside, same era.
2021
Cited alongside, same era.
A. Jaegle, S. Borgeaud, J. Alayrac, C. Doersch, C. Ionescu, D. Ding, S. Koppula, D. Zoran, A. Brock, E. Shelhamer, O. J. Hénaff, M. M. Botvinick, A. Zisserman, O. Vinyals, and J. Carreira, “Perceiver io: A general architecture for structured inputs & outputs.” arXiv, 2021
Z. Chen, Y. Zhang, A. Rosenberg, B. Ramabhadran, P. J. Moreno, A. Bapna, and H. Zen, “MAESTRO: Matched Speech Text Representations through Modality Matching,” in Proc. Interspeech , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
Z. Chen, Y. Zhang, A. Rosenberg, B. Ramabhadran, P. Moreno, and G. Wang, “Tts4pretrain 2.0: Advancing the use of text and speech in asr pretraining with consistency and contrastive losses,” in Proc. ICASSP , 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
U. Khandelwal, A. Fan, D. Jurafsky, and L. Z. amd M. Lewis, “Nearest neighbor machine translation,” in Proc. ICLR , 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
G. Izacard and E. Grave, “Distilling knowledge from reader to retriever for question answering,” in Proc. ICLR , 2021. [Online]. Available: https://openreview.net/forum?id=NTEz-6wysdb
2021
Cited alongside, same era.
P. Pasupat, Y. Zhang, and K. Guu, “Controllable semantic parsing via retrieval augmentation,” in Proc. EMNLP , Nov. 2021, pp. 7683–7698
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
S. Thomas, B. Kingsbury, G. Saon, and H.-K. J. Kuo, “Integrating text inputs for training and adapting rnn transducer asr models,” in Proc. ICASSP , 2022
2022
Cited alongside, same era.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
R. Gupta, H. Lee, J. Zhao, Y. Cao, A. Rastogi, and Y. Wu, “Show, don’t tell: Demonstrations outperform descriptions for schema-guided task-oriented dialogue,” in Proc. NACCL . ACL, Jul. 2022
2022
Later among the works it cites.
D. Yu, M. Wang, Y. Cao, L. El Shafey, I. Shafran, and H. Soltau, “Knowledge-grounded dialog state tracking,” in Proc. EMNLP , 2022
2022
Later among the works it cites.