Fetching the paper…
Reading the bibliography…
In this work, we define barge-in verification as a supervised learning task where audio-only information is used to classify user spoken dialogue into true and false barge-ins.
M. Sugiyama, H. Sawai, and A. H. Waibel, “Review of TDNN (time delay neural network) architectures for speech recognition,” in 1991., IEEE International Sympoisum on Circuits and Systems , 1991, pp. 582–585
1991
Earlier work this paper cites.
I. Guyon, J. Makhoul, R. Schwartz, and V. Vapnik, “What size test set gives good error rate estimates?” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 20, no. 1, pp. 52–64, 1998
1998
Earlier work this paper cites.
N. Ström and S. Seneff, “Intelligent barge-in in conversational systems.” in INTERSPEECH , 2000, pp. 652–655
2000
Earlier work this paper cites.
R. C. Rose and H. K. Kim, “A hybrid barge-in procedure for more reliable turn-taking in human-machine dialog systems,” in 2003 IEEE Workshop on Automatic Speech Recognition and Understanding (IEEE Cat. No. 03EX721) . IEEE, 2003, pp. 198–203
2003
Earlier work this paper cites.
R. Pieraccini, D. Suendermann, K. Dayanidhi, and J. Liscombe, “Are we there yet? research in commercial spoken dialog systems,” in International Conference on Text, Speech and Dialogue . Springer, 2009, pp. 3–13
2009
Earlier work this paper cites.
K. Komatani and H. G. Okuno, “Online error detection of barge-in utterances by using individual users’ utterance histories in spoken dialogue system,” in Proceedings of the SIGDIAL 2010 Conference , 2010, pp. 289–296
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, G. Stemmer, and K. Vesely, “The Kaldi Speech Recognition Toolkit,” in Workshop on Automatic Speech Recognition and Understanding , 2011
2011
Earlier work this paper cites.
E. Shriberg, A. Stolcke, D. Hakkani-Tür, and L. Heck, “Learning when to listen: Detecting system-addressed speech in human-human-computer dialog,” in Proc. Interspeech , 2012, pp. 334–337
2012
Earlier work this paper cites.
E. Selfridge, I. Arizmendi, P. A. Heeman, and J. D. Williams, “Continuously predicting and processing barge-in during a live spoken dialogue task,” in Proceedings of the SIGDIAL 2013 Conference , 2013, pp. 384–393
2013
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an ASR corpus based on public domain audio books,” in 2015 IEEE international conference on acoustics, speech and signal processing (ICASSP) . IEEE, 2015, pp. 5206–5210
2015
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, J. Davison, S. Shleifer, P. von Platen, C. Ma, Y. Jernite, J. Plu, C. Xu, T. L. Scao, S. Gugger, M. Drame, Q. Lhoest, and A. M. Rush, “Transformers: State-of-the-art natural language processing,” in Proceedings of EMNLP: System Demonstrations , 2020, pp. 38–45
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
A. Norouzian, B. Mazoure, D. Connolly, and D. Willett, “Exploring attention mechanism for acoustic-based classification of speech utterances into system-directed and non-system-directed,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2019, pp. 7310–7314
2019
Cited alongside, same era.
2019
Cited alongside, same era.
E. Hosseini-Asl, B. McCann, C.-S. Wu, S. Yavuz, and R. Socher, “A simple language model for task-oriented dialogue,” Advances in Neural Information Processing Systems , vol. 33, pp. 20 179–20 191, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.