Fetching the paper…
Reading the bibliography…
This paper introduces the second DIHARD challenge, the second in a series of speaker diarization challenges intended to improve the robustness of diarization systems to variation in recording equipment, noise conditions, and conversational domain.
L. Hamers
1989
Earlier work this paper cites.
J. S. Garofolo
1993
Earlier work this paper cites.
R. Real and J. M. Vargas, “The probabilistic basis of Jaccard’s index of similarity,”
1996
Earlier work this paper cites.
C. Cieri, D. Miller, and K. Walker, “From Switchboard to Fisher: Telephone collection protocols, their uses and yields,” in
2003
Earlier work this paper cites.
Y. L. Qian, S. X. Lin, Y. D. Zhang, Y. Liu, H. Liu, and Q. Liu, “An introduction to corpora resources of 863 program for Chinese language processing and human-machine interaction,”
2004
Earlier work this paper cites.
J. G. Fiscus, J. Ajot, M. Michel, and J. S. Garofolo, “The Rich Transcription 2006 Spring Meeting Recognition Evaluation,” in
2006
Earlier work this paper cites.
S. Srinivasan, N. Roman, and D. Wang, “Binary and ratio time-frequency masks for robust speech recognition,”
2006
Earlier work this paper cites.
X. Anguera, C. Wooters, and J. Hernando, “Acoustic beamforming for speaker diarization of meetings,”
2007
Earlier work this paper cites.
S. J. Prince and J. H. Elder, “Probabilistic linear discriminant analysis for inferences about identity,” in
2007
Earlier work this paper cites.
K. J. Han, S. Kim, and S. S. Narayanan, “Strategies to improve the robustness of agglomerative hierarchical clustering under data source variation for speaker diarization,”
2008
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Earlier work this paper cites.
D. Garcia-Romero and C. Y. Espy-Wilson, “Analysis of i-vector length normalization in speaker recognition systems,” in
2011
Earlier work this paper cites.
M. Rouvier, G. Dupuy, P. Gay, E. Khoury, T. Merlin, and S. Meignier, “An open-source state-of-the-art toolbox for broadcast news diarization,” in
2013
Earlier work this paper cites.
S. H. Yella and H. Bourlard, “Improved overlap speech diarization of meeting recordings using long-term conversational features,” in
2013
Earlier work this paper cites.
G. Sell and D. Garcia-Romero, “Speaker diarization with PLDA i-vector scoring and unsupervised calibration,” in
2014
Cited alongside, same era.
S. H. Yella, A. Stolcke, and M. Slaney, “Artificial neural network features for speaker diarization,” in
2014
Cited alongside, same era.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “LibriSpeech: an ASR corpus based on public domain audio books,” in
2015
Cited alongside, same era.
D. Snyder, G. Chen, and D. Povey, “MUSAN: A music, speech, and noise corpus,”
2015
Cited alongside, same era.
W. Zhu and J. Pelecanos, “Online speaker diarization using adapted i-vector transforms,” in
2016
Cited alongside, same era.
N. Ryant, E. Bergelson, K. Church, A. Cristia, J. Du, S. Ganapathy, S. Khudanpur, D. Kowalski, M. Krishnamoorthy, R. Kulshreshta
2018
Later among the works it cites.
N. Ryant, K. Church, C. Cieri, A. Cristia, J. Du, S. Ganapathy, and M. Liberman, “First DIHARD challenge evaluation plan,” Tech. Rep., 2018. [Online]. Available:
2018
Later among the works it cites.
G. Sell, D. Snyder, A. McCree, D. Garcia-Romero, J. Villalba, M. Maciejewski, V. Manohar, N. Dehak, D. Povey, S. Watanabe
2018
Later among the works it cites.
M. Diez, F. Landini, L. Burget, J. Rohdin, A. Silnova, K. Zmolıková, O. Novotnỳ, K. Veselỳ, O. Glembek, O. Plchot
2018
Later among the works it cites.
J. Barker, S. Watanabe, E. Vincent, and J. Trmal, “The Fifth ‘CHiME’ Speech Separation and Recognition Challenge: Dataset, task and baselines,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Milner and T. Hain, “Segment-oriented evaluation of speaker diarisation performance,” in
2016
Cited alongside, same era.
E. Bergelson, “Bergelson Seedlings HomeBank Corpus,” 2016, doi:10.21415/T5PK6D
2016
Cited alongside, same era.
D. Snyder, P. Ghahremani, D. Povey, D. Garcia-Romero, Y. Carmiel, and S. Khudanpur, “Deep neural network-based speaker embeddings for end-to-end speaker verification,” in
2016
Cited alongside, same era.
D. Garcia-Romero, D. Snyder, G. Sell, D. Povey, and A. McCree, “Speaker diarization using deep neural network embeddings,” in
2017
Cited alongside, same era.
I. Viñals, A. Ortega, J. A. V. López, A. Miguel, and E. Lleida, “Domain adaptation of PLDA models in broadcast diarization by means of unsupervised speaker clustering.” in
2017
Cited alongside, same era.
L. Sun, J. Du, L.-R. Dai, and C.-H. Lee, “Multiple-target deep learning for LSTM-RNN based speech enhancement,” in
2017
Cited alongside, same era.
A. Nagrani, J. S. Chung, and A. Zisserman, “VoxCeleb: a large-scale speaker identification dataset,”
2017
Cited alongside, same era.
J. Tracey and S. Strassel, “VAST: A corpus of video annotation for speech technologies,” in
2018
Later among the works it cites.
T. Gao, J. Du, L.-R. Dai, and C.-H. Lee, “Densely connected progressive learning for LSTM-based speech enhancement,” in
2018
Later among the works it cites.
L. Sun, J. Du, C. Jiang, X. Zhang, S. He, B. Yin, and C.-H. Lee, “Speaker diarization with enhancing speech for the First DIHARD Challenge,”
2018
Later among the works it cites.
L. Sun, J. Du, T. Gao, Y.-D. Lu, Y. Tsao, C.-H. Lee, and N. Ryant, “A novel LSTM-based speech preprocessor for speaker diarization in realistic mismatch conditions,” in
2018
Later among the works it cites.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-vectors: Robust DNN embeddings for speaker recognition,” in
2018
Later among the works it cites.
M. Diez, L. Burget, and P. Matejka, “Speaker diarization based on Bayesian HMM with eigenvoice priors,” in
2018
Later among the works it cites.
J. S. Chung, A. Nagrani, and A. Zisserman, “VoxCeleb2: Deep speaker recognition,”
2018
Later among the works it cites.
A. Zhang, Q. Wang, Z. Zhu, J. Paisley, and C. Wang, “Fully supervised speaker diarization,”
2019
Closest in time.
N. Ryant, K. Church, C. Cieri, A. Cristia, J. Du, S. Ganapathy, and M. Liberman, “Second DIHARD challenge evaluation plan,” Tech. Rep., 2019. [Online]. Available:
2019
Closest in time.