Fetching the paper…
Reading the bibliography…
A major focus of recent research in spoken language understanding (SLU) has been on the end-to-end approach where a single model can predict intents directly from speech inputs without intermediate transcripts.
J. Kittler, M. Hatef, R. Duin, and J. Matas, ”On Combining Classifiers,” IEEE Pattern Analysis and Machine Intelligence
1998
Earlier work this paper cites.
V. Goel, H.-K. J. Kuo, S. Deligne, and C. Wu, “Language model estimation for optimizing end-to-end performance of a natural language call routing system,” in Proc. ICASSP
2005
Earlier work this paper cites.
S. Yaman, L. Deng, D. Yu, Y.-Y. Wang, and A. Acero, “An integrative and discriminative technique for spoken utterance classification,” in IEEE Trans. on Audio, Speech, and Language Processing
2008
Earlier work this paper cites.
C. Lee, S. Jung, K. Kim, D. Lee, and G. G. Lee, “Recent approaches to dialog management for spoken dialog systems,” Journal of Computing Science and Engineering
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, G. Stemmer, and K. Vesely, ”The Kaldi Speech Recognition Toolkit,” IEEE 2011 Workshop on Automatic Speech Recognition and Understanding
2011
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “LibriSpeech: An ASR corpus based on public domain audio books,” ICASSP
2015
Earlier work this paper cites.
Y. Qian, R. Ubale, V. Ramanaryanan, P. Lange, D. Suendermann-Oeft, K. Evanini, and E. tsuprun, “Exploring ASR-free end-to-end modeling to improve spoken language understanding in a cloud-based dialog system,” in 2017 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU)
2017
Earlier work this paper cites.
P. Haghani, A. Narayanan, M. Bacchiani, G. Chuang, N. Gaur, P. Moreno, R. Prabhavalkar, Z. Qu, and A. Waters, “From audio to semantics: Approaches to end-to-end spoken language understanding,” in 2018 IEEE Spoken Language Technology Workshop (SLT)
2018
Cited alongside, same era.
D. Serdyuk, Y. Wang, C. Fuegen, A. Kumar, B. Liu, and Y. Bengio, “Towards end-to-end spoken language understanding,” in 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2018
Cited alongside, same era.
S. Ghannay, A. Caubrière, Y. Estève, N. Camelin, E. Simonnet, A. Laurent, and E. Morin, “End-to-end named entity and semantic concept extraction from speech,” in 2018 IEEE Spoken Language Technology Workshop (SLT)
2018
Cited alongside, same era.
Y.-P. Chen, R. Price, and S. Bangalore, “Spoken language understanding without speech recognition,” in 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2018
Cited alongside, same era.
2019
Later among the works it cites.
M. Radfar, A. Mouchtaris, and S. Kunzmann, “End-to-end neural transformer based spoken language understanding,” in Interspeech 2020, Annual Conference of the International Speech Communication Association
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. Huang, H.-K. Kuo, S. Thomas, Z. Kons, K. Audhkhasi, B. Kingsbury, R. Hoory, and M. Picheny, “Leveraging unpaired text data for training end-to-end speech-to-intent systems,” in Proc. ICASSP
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2019
Cited alongside, same era.
2020
Later among the works it cites.
R. Price, M. Mehrabani and S. Bangalore, ”Improved End-To-End Spoken Utterance Classification with a Self-Attention Acoustic Classifier,” 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2020
Later among the works it cites.