Fetching the paper…
Reading the bibliography…
Spoken conversational question answering (SCQA) requires machines to model complex dialogue flow given the speech utterances and text corpora.
L.-s. Lee, J. Glass, H.-y. Lee, and C.-a. Chan, “Spoken content retrieval—beyond cascading speech recognition with text retrieval,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 23, no. 9, pp. 1389–1420, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
C. Guo, G. Pleiss, Y. Sun, and K. Q. Weinberger, “On calibration of modern neural networks,” in International Conference on Machine Learning , 2017, pp. 1321–1330
2017
Earlier work this paper cites.
M. Honnibal and I. Montani, “spaCy 2: Natural language understanding with Bloom embeddings, convolutional neural networks and incremental parsing,” 2017, to appear
2017
Earlier work this paper cites.
C. You, Q. Yang, H. Shan, L. Gjesteby, G. Li, S. Ju, Z. Zhang, Z. Zhao, Y. Zhang, W. Cong et al. , “Structurally-sensitive multi-scale deep neural network for low-dose ct denoising,” IEEE Access , 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
C.-H. Lee, S.-M. Wang, H.-C. Chang, and H.-Y. Lee, “ODSQA: Open-domain spoken question answering dataset,” in 2018 IEEE Spoken Language Technology Workshop (SLT) . IEEE, 2018, pp. 949–956
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Cited alongside, same era.
D. Serdyuk, Y. Wang, C. Fuegen, A. Kumar, B. Liu, and Y. Bengio, “Towards end-to-end spoken language understanding,” in 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2018, pp. 5754–5758
2018
Cited alongside, same era.
C. You, G. Li, Y. Zhang, X. Zhang, H. Shan, M. Li, S. Ju, Z. Zhao, Z. Zhang, W. Cong et al. , “CT super-resolution GAN constrained by the identical, residual, and cycle learning ensemble (gan-circle),” IEEE Transactions on Medical Imaging , 2019
2019
Cited alongside, same era.
C. You, L. Yang, Y. Zhang, and G. Wang, “Low-dose ct via deep cnn with skip connection and network-in-network,” in Developments in X-Ray Tomography XII , 2019
2019
Cited alongside, same era.
2020
Closest in time.
D. Su and P. Fung, “Improving spoken question answering using contextualized word representation,” in 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2020, pp. 8004–8008
2020
Closest in time.
N. Chen, F. Liu, C. You, P. Zhou, and Y. Zou, “Adaptive bi-directional attention: Exploring multi-granularity representations for machine reading comprehension,” in IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , 2020
2020
Closest in time.
Z. Sun, P. K. Sarma, W. A. Sethares, and Y. Liang, “Learning relationships between text, audio, and video via deep canonical correlation for multimodal language analysis,” in Proceedings of the AAAI Conference on Artificial Intelligence , 2020, pp. 8992–8999
2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S.-B. Luo, H.-S. Lee, K.-Y. Chen, and H.-M. Wang, “Spoken multiple-choice question answering using multimodal convolutional neural networks,” in 2019 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU) . IEEE, 2019, pp. 772–778
2019
Cited alongside, same era.
C.-H. Lee, Y.-N. Chen, and H.-Y. Lee, “Mitigating the impact of speech recognition errors on spoken question answering by adversarial domain adaptation,” in 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2019, pp. 7300–7304
2019
Cited alongside, same era.
Y.-H. H. Tsai, S. Bai, P. P. Liang, J. Z. Kolter, L.-P. Morency, and R. Salakhutdinov, “Multimodal transformer for unaligned multimodal language sequences,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Florence, Italy: Association for Computational Linguistics, 7 2019
2019
Cited alongside, same era.
D. Peskov, J. Barrow, P. Rodriguez, G. Neubig, and J. Boyd-Graber, “Mitigating noisy inputs for question answering,” in Annual Conference of the International Speech Communication Association (INTERSPEECH) , 2019
2019
Cited alongside, same era.
S. Reddy, D. Chen, and C. D. Manning, “Coqa: A conversational question answering challenge,” Trans. Assoc. Comput. Linguistics , vol. 7, pp. 249–266, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
C. You, J. Yang, J. Chapiro, and J. S. Duncan, “Unsupervised wasserstein distance guided domain adaptation for 3d multi-domain liver segmentation,” in Interpretable and Annotation-Efficient Learning for Medical Image Computing , 2020
2020
Cited alongside, same era.
Closest in time.
S. Siriwardhana, A. Reis, R. Weerasekera, and S. Nanayakkara, “Jointly Fine-Tuning “BERT-like” Self Supervised Models to Improve Multimodal Speech Emotion Recognition,” in INTERSPEECH , 2020
2020
Closest in time.
Z. Lan, M. Chen, S. Goodman, K. Gimpel, P. Sharma, and R. Soricut, “ALBERT: A Lite BERT for Self-supervised Learning of Language Representations,” in International Conference on Learning Representations , 2020
2020
Closest in time.
2021
Closest in time.
C. You, N. Chen, and Y. Zou, “Knowledge distillation for improved accuracy in spoken question answering,” in IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , 2021
2021
Closest in time.
M. Ünlü and E. Arisoy, “Uncertainty-aware representations for spoken question answering,” in IEEE Spoken Language Technology Workshop (SLT) , 2021
2021
Closest in time.
2021
Closest in time.
E. Arısoy and M. Ünlü, “Uncertainty-aware representations for spoken question answering,” 2021
2021
Closest in time.