Fetching the paper…
Reading the bibliography…
More and more neural network approaches have achieved considerable improvement upon submodules of speaker diarization system, including speaker change detection and segment-wise speaker embedding extraction.
K. C. Gowda and G. Krishna, “Agglomerative clustering using the concept of mutual nearest neighbourhood,”
1978
Earlier work this paper cites.
S. E. Tranter and D. A. Reynolds, “An overview of automatic speaker diarization systems,”
2006
Earlier work this paper cites.
S. J. D. Prince and J. H. Elder, “Probabilistic linear discriminant analysis for inferences about identity,” in
2007
Earlier work this paper cites.
U. von Luxburg, “A tutorial on spectral clustering,”
2007
Earlier work this paper cites.
S. Meignier and T. Merlin, “Lium spkdiarization: an open source toolkit for diarization,” in
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, G. Stemmer, and K. Vesely, “The kaldi speech recognition toolkit,” in
2011
Earlier work this paper cites.
X. Anguera, S. Bozonnet, N. Evans, C. Fredouille, G. Friedland, and O. Vinyals, “Speaker diarization: A review of recent research,”
2012
Earlier work this paper cites.
S. H. Shum, N. Dehak, R. Dehak, and J. R. Glass, “Unsupervised methods for speaker diarization: An integrated and iterative approach,”
2013
Earlier work this paper cites.
P. Kenny, T. Stafylakis, P. Ouellet, M. J. Alam, and P. Dumouchel, “Plda for speaker verification with utterances of arbitrary duration,” in
2013
Cited alongside, same era.
G. Sell and D. Garcia-Romero, “Speaker diarization with plda i-vector scoring and unsupervised calibration,” in
2014
Cited alongside, same era.
G. Sell and D. Garcia-Romero, “Diarization resegmentation in the factor analysis subspace,” in
2015
Cited alongside, same era.
M. Hrúz and Z. Zajíc, “Convolutional neural network for speaker change detection in telephone speaker diarization system,” in
2017
Cited alongside, same era.
R. Yin, H. Bredin, and C. Barras, “Speaker change detection in broadcast tv using bidirectional long short-term memory networks,” in
2017
Cited alongside, same era.
Q. Wang, C. Downey, L. Wan, P. A. Mansfield, and I. L. Moreno, “Speaker diarization with lstm,” in
2018
Later among the works it cites.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-vectors: Robust dnn embeddings for speaker recognition,” in
2018
Later among the works it cites.
W. Cai, J. Chen, and M. Li, “Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,” in
2018
Later among the works it cites.
Weicheng Cai, Jinkun Chen and Ming Li, “Analysis of length normalization in end-to-end speaker verification system,” in
2018
Later among the works it cites.
R. Yin, H. Bredin, and C. Barras, “Neural speech turn segmentation and affinity propagation for speaker diarization,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Garcia-Romero, D. Snyder, G. Sell, D. Povey, and A. McCree, “Speaker diarization using deep neural network embeddings,” in
2017
Cited alongside, same era.
G. Wisniewksi, H. Bredin, G. Gelly, and C. Barras, “Combining speaker turn embedding and incremental structure prediction for low-latency speaker diarization,” in
2017
Cited alongside, same era.
M. Price, J. Glass, and A. P. Chandrakasan, “A low-power speech recognizer and voice activity detector using deep neural networks,”
2018
Cited alongside, same era.
G. Sell, D. Snyder, A. McCree, D. Garcia-Romero, J. Villalba, M. Maciejewski, V. Manohar, N. Dehak, D. Povey, S. Watanabe, and S. Khudanpur, “Diarization is hard: Some experiences and lessons learned for the jhu team in the inaugural dihard challenge,” in
2018
Later among the works it cites.
A. Zhang, Q. Wang, Z. Zhu, J. Paisley, and C. Wang, “Fully supervised speaker diarization,” in
2019
Closest in time.