Fetching the paper…
Reading the bibliography…
Speaker diarization (SD) is typically used with an automatic speech recognition (ASR) system to ascribe speaker labels to recognized words.
D. G. Canavan, Alexandra and G. Zipperlen, “Callhome american english speech ldc97s42,” Web Download. Philadelphia: Linguistic Data Consortium
1997
Earlier work this paper cites.
C. Cieri et al
2004
Earlier work this paper cites.
C. Cieri et al
2005
Earlier work this paper cites.
J. G. Fiscus, J. Ajot, N. Radde, C. Laprun, et al
2006
Earlier work this paper cites.
J. G. Fiscus et al
2007
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, and J. Weston, “Curriculum learning,” in Proceedings of the 26th annual international conference on machine learning
2009
Earlier work this paper cites.
M. T. Knox, N. Mirghafori, and G. Friedland, “Where did i go wrong?: Identifying troublesome segments for speaker diarization systems,” in Thirteenth Annual Conference of the International Speech Communication Association
2012
Earlier work this paper cites.
D. Garcia-Romero, D. Snyder, G. Sell, D. Povey, and A. McCree, “Speaker diarization using deep neural network embeddings,” in 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
P. Cyrta, T. Trzciński, and W. Stokowiec, “Speaker diarization using deep recurrent convolutional neural networks for speaker embeddings,” in Information Systems Architecture and Technology: Proceedings of 38th International Conference on Information Systems Architecture and Technology–ISAT 2017: Part I
2017
Earlier work this paper cites.
Q. Wang, C. Downey, et al
2018
Earlier work this paper cites.
R. Yin, H. Bredin, and C. Barras, “Neural speech turn segmentation and affinity propagation for speaker diarization,” in Annual Conference of the International Speech Communication Association
2018
Earlier work this paper cites.
2018
Cited alongside, same era.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-vectors: Robust dnn embeddings for speaker recognition,” in 2018 IEEE international conference on acoustics, speech and signal processing (ICASSP)
2018
Cited alongside, same era.
2018
Cited alongside, same era.
A. Zhang, Q. Wang, Z. Zhu, J. Paisley, and C. Wang, “Fully supervised speaker diarization,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2019
Cited alongside, same era.
W. Zhou, W. Michel, K. Irie, M. Kitza, R. Schlüter, and H. Ney, “The rwth asr system for ted-lium release 2: Improving hybrid hmm with specaugment,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2020
Later among the works it cites.
Y. Wang, A. Mohamed, D. Le, C. Liu, A. Xiao, J. Mahadeokar, H. Huang, A. Tjandra, X. Zhang, F. Zhang, C. Fuegen, G. Zweig, and M. L. Seltzer, “Transformer-based acoustic modeling for hybrid speech recognition,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2020
Later among the works it cites.
A. Gulati, J. Qin, C.-C. Chiu, N. Parmar, Y. Zhang, J. Yu, W. Han, S. Wang, Z. Zhang, Y. Wu, et al
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. R. Salakhutdinov, and Q. V. Le, “Xlnet: Generalized autoregressive pretraining for language understanding,” Advances in neural information processing systems
2019
Cited alongside, same era.
2020
Cited alongside, same era.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Y. Higuchi, M. Suzuki, and G. Kurata, “Speaker embeddings incorporating acoustic conditions for diarization,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2020
Cited alongside, same era.
2020
Later among the works it cites.
2021
Later among the works it cites.
T. J. Park, N. Kanda, D. Dimitriadis, K. J. Han, S. Watanabe, and S. Narayanan, “A review of speaker diarization: Recent advances with deep learning,” Computer Speech & Language
2022
Later among the works it cites.
W. Xia, H. Lu, Q. Wang, A. Tripathi, Y. Huang, I. L. Moreno, and H. Sak, “Turn-to-diarize: Online speaker diarization constrained by transformer transducer speaker turn detection,” in ICASSP 2022-2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
M. India, J. Hernando, and J. A. Fonollosa, “Language modelling for speaker diarization in telephonic interviews,” Computer Speech & Language
2023
Closest in time.
M. Chen, A. Papangelis, C. Tao, S. Kim, A. Rosenbaum, Y. Liu, Z. Yu, and D. Hakkani-Tur, “Places: Prompting language models for social conversation synthesis,” in Findings of the Association for Computational Linguistics: EACL 2023
2023
Closest in time.