Fetching the paper…
Reading the bibliography…
Spoken intent detection has become a popular approach to interface with various smart devices with ease.
1983
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in
2015
Earlier work this paper cites.
O. Vinyals, C. Blundell, T. Lillicrap, and D. Wierstra, “Matching networks for one shot learning,” in
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
Y. Qian, R. Ubale, V. Ramanarayanan, P. L. Lange, D. Suendermann-Oeft, K. Evanini, and E. Tsuprun, “Exploring asr-free end-to-end modeling to improve spoken language understanding in a cloud-based dialog system,”
2017
Earlier work this paper cites.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” in
2017
Earlier work this paper cites.
J. Snell, K. Swersky, and R. Zemel, “Prototypical networks for few-shot learning,” in
2017
Earlier work this paper cites.
J. Bradbury, S. Merity, C. Xiong, and R. Socher, “Quasi-Recurrent Neural Networks,”
2017
Earlier work this paper cites.
P. Warden, “Speech commands: A dataset for limited-vocabulary speech recognition,”
2018
Earlier work this paper cites.
P. Haghani, A. Narayanan, M. Bacchiani, G. Chuang, N. Gaur, P. Moreno, R. Prabhavalkar, Z. Qu, and A. Waters, “From audio to semantics: Approaches to end-to-end spoken language understanding,” in
2018
Earlier work this paper cites.
D. Serdyuk, Y. Wang, C. Fuegen, A. Kumar, B. Liu, and Y. Bengio, “Towards end-to-end spoken language understanding,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Y. Chen, R. Price, and S. Bangalore, “Spoken language understanding without speech recognition,” in
2018
Cited alongside, same era.
S. Ghannay, A. Caubrière, Y. Estève, N. Camelin, E. Simonnet, A. Laurent, and E. Morin, “End-to-end named entity and semantic concept extraction from speech,” in
2018
Cited alongside, same era.
Y. Qian, R. Ubale, P. Lange, K. Evanini, and F. Soong, “From speech signals to semantics — tagging performance at acoustic, phonetic and word levels,” in
2018
L. Lugosch, M. Ravanelli, P. Ignoto, V. S. Tomar, and Y. Bengio, “Speech model pre-training for end-to-end spoken language understanding,”
2019
Later among the works it cites.
S. Bhosale, I. Sheikh, S. H. Dumpala, and S. K. Kopparapu, “End-to-end spoken language understanding: Bootstrapping in low resource scenarios,” in
2019
Later among the works it cites.
S. Pascual, M. Ravanelli, J. Serrà, A. Bonafonte, and Y. Bengio, “Learning problem-agnostic speech representations from multiple self-supervised tasks,”
2019
Later among the works it cites.
A. Rajeswaran, C. Finn, S. M. Kakade, and S. Levine, “Meta-learning with implicit gradients,” in
2019
Later among the works it cites.
K. Lee, S. Maji, A. Ravichandran, and S. Soatto, “Meta-learning with differentiable convex optimization,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
2018
Cited alongside, same era.
S. Gidaris and N. Komodakis, “Dynamic few-shot visual learning without forgetting,” in
2018
Cited alongside, same era.
M. Ravanelli and Y. Bengio, “Speaker recognition from raw waveform with sincnet,” in
2018
Cited alongside, same era.
C.-C. Kao, M. Sun, Y. Gao, S. Vitaladevuni, and C. Wang, “Sub-band convolutional neural networks for small-footprint spoken term classification,”
2019
Cited alongside, same era.
2019
Cited alongside, same era.
L. Bertinetto, J. F. Henriques, P. Torr, and A. Vedaldi, “Meta-learning with differentiable closed-form solvers,” in
2019
Later among the works it cites.
J. Poncelet and H. V. hamme, “Multitask learning with capsule networks for speech-to-intent applications,”
2020
Later among the works it cites.
Y. Huang, H.-K. Kuo, S. Thomas, Z. Kons, K. Audhkhasi, B. Kingsbury, R. Hoory, and M. Picheny, “Leveraging unpaired text data for training end-to-end speech-to-intent systems,” in
2020
Later among the works it cites.
M. Ravanelli, J. Zhong, S. Pascual, P. Swietojanski, J. Monteiro, J. Trmal, and Y. Bengio, “Multi-task self-supervised learning for robust speech recognition,” in
2020
Later among the works it cites.
N. R. Koluguri, M. Kumar, S. H. Kim, C. Lord, and S. Narayanan, “Meta-learning for robust child-adult classification from speech,” in
2020
Later among the works it cites.