Fetching the paper…
Reading the bibliography…
Recent years have witnessed significant improvement in ASR systems to recognize spoken utterances.
1907
Earlier work this paper cites.
X. Li and D. Roth, “Learning question classifiers,” in 19th International Conference on Computational Linguistics, COLING 2002, Taipei, Taiwan , 2002. [Online]. Available: https://www.aclweb.org/anthology/C02-1150/
2002
Earlier work this paper cites.
2005
Earlier work this paper cites.
2005
Earlier work this paper cites.
G. Tür, D. Hakkani-Tür, and L. P. Heck, “What is left to be understood in atis?” in 2010 IEEE Spoken Language Technology Workshop, SLT 2010, California, USA . IEEE, 2010, pp. 19–24. [Online]. Available: https://doi.org/10.1109/SLT.2010.5700816
2010
Earlier work this paper cites.
R. Socher, A. Perelygin, J. Wu, J. Chuang, C. D. Manning, A. Y. Ng, and C. Potts, “Recursive deep models for semantic compositionality over a sentiment treebank,” in Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing, EMNLP, Seattle, Washington, USA . ACL, 2013, pp. 1631–1642. [Online]. Available: https://www.aclweb.org/anthology/D13-1170/
2013
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An ASR corpus based on public domain audio books,” in IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP, Queensland, Australia . IEEE, 2015, pp. 5206–5210. [Online]. Available: https://doi.org/10.1109/ICASSP.2015.7178964
2015
Cited alongside, same era.
F. Ladhak, A. Gandhe, M. Dreyer, L. Mathias, A. Rastrow, and B. Hoffmeister, “Latticernn: Recurrent neural networks over lattices,” in Interspeech . ISCA, 2016, pp. 695–699. [Online]. Available: https://doi.org/10.21437/Interspeech.2016-1583
2016
Cited alongside, same era.
R. He and J. J. McAuley, “Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering,” in Proceedings of the 25th International Conference on World Wide Web, WWW 2016, Montreal, Canada . ACM, 2016, pp. 507–517. [Online]. Available: https://doi.org/10.1145/2872427.2883037
2016
Cited alongside, same era.
P. Yenigalla, A. Kumar, S. Tripathi, C. Singh, S. Kar, and J. Vepa, “Speech emotion recognition using spectrogram & phoneme embedding,” in Interspeech 2018 . ISCA, 2018, pp. 3688–3692. [Online]. Available: https://doi.org/10.21437/Interspeech.2018-1811
2018
Later among the works it cites.
T. Desot, F. Portet, and M. Vacher, “SLU for voice command in smart home: Comparison of pipeline and end-to-end approaches,” in IEEE Automatic Speech Recognition and Understanding Workshop, ASRU 2019, Singapore . IEEE, 2019, pp. 822–829. [Online]. Available: https://doi.org/10.1109/ASRU46091.2019.9003891
2019
Later among the works it cites.
K. Irie, R. Prabhavalkar, A. Kannan, A. Bruguier, D. Rybach, and P. Nguyen, “On the choice of modeling unit for sequence-to-sequence speech recognition,” in Interspeech 2019, 20th Annual Conference of the International Speech Communication Association, Graz, Austria, 15-19 September 2019 , G. Kubin and Z. Kacic, Eds. ISCA, 2019, pp. 3800–3804. [Online]. Available: https://doi.org/10.21437/Interspeech.2019-2277
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang, “Squad: 100, 000+ questions for machine comprehension of text,” in Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing, EMNLP 2016, Texas, USA , 2016, pp. 2383–2392. [Online]. Available: https://doi.org/10.18653/v1/d16-1264
2016
Cited alongside, same era.
D. Serdyuk, Y. Wang, C. Fuegen, A. Kumar, B. Liu, and Y. Bengio, “Towards end-to-end spoken language understanding,” in 2018 IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2018, Calgary, AB, Canada . IEEE, 2018, pp. 5754–5758. [Online]. Available: https://doi.org/10.1109/ICASSP.2018.8461785
2018
Cited alongside, same era.
S. Ghannay, A. Caubrière, Y. Estève, N. Camelin, E. Simonnet, A. Laurent, and E. Morin, “End-to-end named entity and semantic concept extraction from speech,” in 2018 IEEE Spoken Language Technology Workshop, SLT 2018, Athens, Greece . IEEE, 2018, pp. 692–699. [Online]. Available: https://doi.org/10.1109/SLT.2018.8639513
2018
Cited alongside, same era.
Y. Weng, S. S. Miryala, C. Khatri, R. Wang, H. Zheng, P. Molino, M. Namazifar, A. Papangelis, H. Williams, F. Bell, and G. Tür, “Joint contextual modeling for ASR correction and language understanding,” in 2020 IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2020, Barcelona, Spain . IEEE, 2020, pp. 6349–6353. [Online]. Available: https://doi.org/10.1109/ICASSP40776.2020.9053213
2020
Later among the works it cites.
A. Mani, S. Palaskar, and S. Konam, “Towards understanding asr error correction for medical conversations,” in Proceedings of the First Workshop on Natural Language Processing for Medical Conversations , 2020, pp. 7–11
2020
Later among the works it cites.
A. Fang, S. Filice, N. Limsopatham, and O. Rokhlenko, “Using phoneme representations to build predictive models robust to ASR errors,” in Proceedings of the 43rd International ACM SIGIR conference on research and development in Information Retrieval, SIGIR 2020, Virtual Event, China . ACM, 2020, pp. 699–708. [Online]. Available: https://doi.org/10.1145/3397271.3401050
2020
Later among the works it cites.