Fetching the paper…
Reading the bibliography…
End-to-end (E2E) models are becoming increasingly popular for spoken language understanding (SLU) systems and are beginning to achieve competitive performance to pipeline-based approaches.
A. Deoras, R. Sarikaya, G. Tur, and D. Hakkani-Tür, “Joint decoding for speech recognition and semantic tagging,” in Annual Conference of the International Speech Communication Association (Interspeech) , September 2012. [Online]. Available: https://www.microsoft.com/en-us/research/publication/joint-decoding-for-speech-recognition-and-semantic-tagging/
2012
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,” The journal of machine learning research , vol. 15, no. 1, pp. 1929–1958, 2014
2014
Earlier work this paper cites.
J. Towns, T. Cockerill, M. Dahan, I. Foster, K. Gaither, A. Grimshaw, V. Hazlewood, S. Lathrop, D. Lifka, G. D. Peterson, R. Roskies, J. R. Scott, and N. Wilkins-Diehr, “XSEDE: Accelerating scientific discovery,” Computing in Science & Engineering , vol. 16, no. 5, pp. 62–74, 2014
2014
Earlier work this paper cites.
N. A. Nystrom, M. J. Levine, R. Z. Roskies, and J. R. Scott, “Bridges: a uniquely flexible HPC resource for new communities and data analytics,” in Proc. XSEDE Conference: Scientific Advancements Enabled by Enhanced Cyberinfrastructure , 2015
2015
Earlier work this paper cites.
R. Sarikaya et al. , “An overview of end-to-end language understanding and dialog management for personal digital assistants,” in SLT , 2016, pp. 391–397
2016
Earlier work this paper cites.
Y. Xia, F. Tian, L. Wu, J. Lin, T. Qin, N. Yu, and T.-Y. Liu, “Deliberation networks: Sequence generation beyond one-pass decoding,” Proc. NeurIPS , vol. 30, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Proc. NeurIPS , vol. 30, 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
D. Yu, M. Cohn et al. , “Gunrock: A social bot for complex and engaging long conversations,” in Proc. EMNLP-IJCNLP - System Demonstrations , 2019
2019
Earlier work this paper cites.
J. Devlin, M. Chang, K. Lee, and K. Toutanova, “BERT: pre-training of deep bidirectional transformers for language understanding,” in Proc. NAACL-HLT , 2019, pp. 4171–4186
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
L. Lugosch, M. Ravanelli, P. Ignoto, V. S. Tomar, and Y. Bengio, “Speech model pre-training for end-to-end spoken language understanding,” in Proc. Interspeech , 2019, pp. 814–818
2019
Earlier work this paper cites.
S. Arora, S. Dalmia, P. Denisov, X. Chang, Y. Ueda, Y. Peng, Y. Zhang, S. Kumar, K. Ganesan, B. Yan et al. , “ESPnet-SLU: Advancing spoken language understanding through espnet,” in Proc. Interspeech , 2019
2019
Cited alongside, same era.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga et al. , “Pytorch: An imperative style, high-performance deep learning library,” Proc. NeurIPS , vol. 32, 2019
2019
Cited alongside, same era.
S. Karita, N. Chen, T. Hayashi, T. Hori, H. Inaguma, Z. Jiang, M. Someki, N. E. Y. Soplin, R. Yamamoto, X. Wang, S. Watanabe, T. Yoshimura, and W. Zhang, “A comparative study on transformer vs rnn in speech applications,” in Proc. ASRU , 2019, pp. 449–456
2019
Cited alongside, same era.
D. S. Park, W. Chan, Y. Zhang, C. Chiu, B. Zoph, E. D. Cubuk, and Q. V. Le, “Specaugment: A simple data augmentation method for automatic speech recognition,” in Interspeech , 2019, pp. 2613–2617
2019
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue et al. , “Transformers: State-of-the-art natural language processing,” in Proc. EMNLP: System Demonstrations , Online, Oct. 2020, pp. 38–45
2020
Later among the works it cites.
M. Saxon, S. Choudhary, J. P. McKenna, and A. Mouchtaris, “End-to-End Spoken Language Understanding for Generalized Voice Assistants,” in Proc. Interspeech , 2021, pp. 4738–4742
2021
Later among the works it cites.
J. Ganhotra, S. Thomas, H.-K. J. Kuo, S. Joshi, G. Saon, Z. Tüske, and B. Kingsbury, “Integrating Dialog History into End-to-End Spoken Language Understanding Systems,” in Proc. Interspeech , 2021, pp. 1254–1258
2021
Later among the works it cites.
S. Arora, A. Ostapenko, V. Viswanathan, S. Dalmia, F. Metze, S. Watanabe, and A. W. Black, “Rethinking end-to-end evaluation of decomposable tasks: A case study on spoken language understanding,” in Proc. Interspeech , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
R. Müller, S. Kornblith, and G. E. Hinton, “When does label smoothing help?” Advances in neural information processing systems , vol. 32, 2019
2019
Cited alongside, same era.
E. Bastianelli, A. Vanzo, P. Swietojanski, and V. Rieser, “SLURP: A spoken language understanding resource package,” in Proc. EMNLP , 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
K. Song, X. Tan, T. Qin, J. Lu, and T. Liu, “MPNet: Masked and permuted pre-training for language understanding,” in Proc. NeurIPS , 2020
2020
Cited alongside, same era.
Y. Huang, H.-K. Kuo, S. Thomas, Z. Kons, K. Audhkhasi, B. Kingsbury, R. Hoory, and M. Picheny, “Leveraging unpaired text data for training end-to-end speech-to-intent systems,” in Proc. ICASSP . IEEE, 2020, pp. 7984–7988
2020
Cited alongside, same era.
K. Hu, T. N. Sainath, R. Pang, and R. Prabhavalkar, “Deliberation model based two-pass end-to-end speech recognition,” in Proc. ICASSP . IEEE, 2020, pp. 7799–7803
2020
Cited alongside, same era.
A. Gulati, J. Qin, C. Chiu, N. Parmar, Y. Zhang, J. Yu, W. Han, S. Wang, Z. Zhang, Y. Wu, and R. Pang, “Conformer: Convolution-augmented transformer for speech recognition,” in Proc. Interspeech , 2020, pp. 5036–5040
2020
Cited alongside, same era.
Y.-S. Chuang, C.-L. Liu, H. yi Lee, and L. shan Lee, “SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering,” in Proc. Interspeech , 2021
2021
Later among the works it cites.
C.-I. Lai, Y.-S. Chuang, H.-Y. Lee, S.-W. Li, and J. Glass, “Semi-supervised spoken language understanding via self-supervised speech and language model pretraining,” in Proc. ICASSP . IEEE, 2021, pp. 7468–7472
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
P. G. Shivakumar, N. Kumar, P. Georgiou, and S. Narayanan, “Rnn based incremental online spoken language understanding,” in Proc. SLT . IEEE, 2021, pp. 989–996
2021
Later among the works it cites.
P. Guo, F. Boyer, X. Chang, T. Hayashi, Y. Higuchi, H. Inaguma, N. Kamo, C. Li et al. , “Recent developments on espnet toolkit boosted by conformer,” in Proc. ICASSP , 2021, pp. 5874–5878
2021
Later among the works it cites.
2021
Later among the works it cites.