Fetching the paper…
Reading the bibliography…
In this paper, we present the submitted system for the second DIHARD Speech Diarization Challenge from the DKULENOVO team.
“Agglomerative clustering using the concept of mutual nearest neighbourhood,”
K Chidananda Gowda and G Krishna, · 1978
Earlier work this paper cites.
“The icsi meeting corpus,”
Adam Janin, Don Baron, Jane Edwards, Dan Ellis, David Gelbart, Nelson Morgan, Barbara Peskin, Thilo Pfau, Elizabeth Shriberg, Andreas Stolcke, et al., · 2003
Earlier work this paper cites.
“Towards robust speaker segmentation: The icsi-sri fall 2004 diarization system,”
Chuck Wooters, James Fung, Barbara Peskin, and Xavier Anguera, · 2004
Earlier work this paper cites.
“The ami meeting corpus: A pre-announcement,”
Jean Carletta, Simone Ashby, Sebastien Bourban, Mike Flynn, Mael Guillemot, Thomas Hain, Jaroslav Kadlec, Vasilis Karaiskos, Wessel Kraaij, Melissa Kronenthal, et al., · 2005
Earlier work this paper cites.
“An overview of automatic speaker diarization systems,”
Sue E Tranter and Douglas A Reynolds, · 2006
Earlier work this paper cites.
“The 2006 athens information technology speech activity detection and speaker diarization systems,”
Elias Rentzeperis, Andreas Stergiou, Christos Boukis, Aristodemos Pnevmatikakis, and Lazaros C Polymenakos, · 2006
Earlier work this paper cites.
“Enhanced svm training for robust speech activity detection,”
Andrey Temko, Dusan Macho, and Climent Nadeu, · 2007
Earlier work this paper cites.
“Probabilistic linear discriminant analysis for inferences about identity,”
Simon JD Prince and James H Elder, · 2007
Earlier work this paper cites.
“A tutorial on spectral clustering,”
Ulrike Von Luxburg, · 2007
Earlier work this paper cites.
“Unsupervised speaker adaptation based on the cosine similarity for text-independent speaker verification.,”
Stephen Shum, Najim Dehak, Reda Dehak, and James R Glass, · 2010
Earlier work this paper cites.
[Online]. Available: https://webrtc.org/
“The webrtc project,” 2011, · 2011
Earlier work this paper cites.
“The kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al., · 2011
Earlier work this paper cites.
“Speaker diarization: A review of recent research,”
Xavier Anguera, Simon Bozonnet, Nicholas Evans, Corinne Fredouille, Gerald Friedland, and Oriol Vinyals, · 2012
Earlier work this paper cites.
“Real-life voice activity detection with lstm recurrent neural networks and an application to hollywood movies,”
Florian Eyben, Felix Weninger, Stefano Squartini, and Björn Schuller, · 2013
Cited alongside, same era.
“Plda for speaker verification with utterances of arbitrary duration,”
Patrick Kenny, Themos Stafylakis, Pierre Ouellet, Md Jahangir Alam, and Pierre Dumouchel, · 2013
Cited alongside, same era.
“Speaker diarization with plda i-vector scoring and unsupervised calibration,”
Gregory Sell and Daniel Garcia-Romero, · 2014
Cited alongside, same era.
“Diarization resegmentation in the factor analysis subspace,”
Gregory Sell and Daniel Garcia-Romero, · 2015
Cited alongside, same era.
“Musan: A music, speech, and noise corpus,”
David Snyder, Guoguo Chen, and Daniel Povey, · 2015
Cited alongside, same era.
“Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,”
Weicheng Cai, Jinkun Chen, and Ming Li, · 2018
Later among the works it cites.
“Analysis of length normalization in end-to-end speaker verification system,”
Weicheng Cai, Jinkun Chen, and Ming Li, · 2018
Later among the works it cites.
“Speaker diarization with lstm,”
Quan Wang, Carlton Downey, Li Wan, Philip Andrew Mansfield, and Ignacio Lopz Moreno, · 2018
Later among the works it cites.
“Diarization is hard: Some experiences and lessons learned for the jhu team in the inaugural dihard challenge.,”
Gregory Sell, David Snyder, Alan McCree, Daniel Garcia-Romero, Jesús Villalba, Matthew Maciejewski, Vimal Manohar, Najim Dehak, Daniel Povey, Shinji Watanabe, et al., · 2018
Later among the works it cites.
“Speaker diarization based on bayesian hmm with eigenvoice priors.,”
Mireia Diez, Lukás Burget, and Pavel Matejka, · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Feature learning with raw-waveform cldnns for voice activity detection,”
Rubén Zazo Candil, Tara N Sainath, Gabor Simko, and Carolina Parada, · 2016
Cited alongside, same era.
“Bergelson seedlings homebank corpus,” 2016
E Bergelson, · 2016
Cited alongside, same era.
“Convolutional neural network for speaker change detection in telephone speaker diarization system,”
Marek Hrúz and Zbyněk Zajíc, · 2017
Cited alongside, same era.
“Speaker change detection in broadcast tv using bidirectional long short-term memory networks,”
Ruiqing Yin, Hervé Bredin, and Claude Barras, · 2017
Cited alongside, same era.
“Voxceleb: a large-scale speaker identification dataset,”
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman, · 2017
Cited alongside, same era.
“A study on data augmentation of reverberant speech for robust speech recognition,”
Tom Ko, Vijayaditya Peddinti, Daniel Povey, Michael L Seltzer, and Sanjeev Khudanpur, · 2017
Cited alongside, same era.
“Temporal modeling using dilated convolution and gating for voice-activity-detection,”
Shuo-Yiin Chang, Bo Li, Gabor Simko, Tara N Sainath, Anshuman Tripathi, Aäron van den Oord, and Oriol Vinyals, · 2018
Cited alongside, same era.
“But system for dihard speech diarization challenge 2018,”
Ondrej Novotnỳ, Karel Veselỳ, Ondrej Glembek, Oldrich Plchot, Ladislav Mošner, and Pavel Matejka, · 2018
Later among the works it cites.
“Enhancement and analysis of conversational speech: Jsalt 2017,”
Neville Ryanta, Elika Bergelson, Kenneth Church, Alejandrina Cristia, Jun Du, Sriram Ganapathy, Sanjeev Khudanpur, Diana Kowalski, Mahesh Krishnamoorthy, Rajat Kulshreshta, et al., · 2018
Later among the works it cites.
“Lstm based similarity measurement with spectral clustering for speaker diarization,”
Qingjian Lin, Ruiqing Yin, Ming Li, Hervé Bredin, and Claude Barras, · 2019
Later among the works it cites.
“Fully supervised speaker diarization,”
Aonan Zhang, Quan Wang, Zhenyao Zhu, John Paisley, and Chong Wang, · 2019
Later among the works it cites.
“Discriminative neural clustering for speaker diarisation,”
Qiujia Li, Florian L Kreyssig, Chao Zhang, and Philip C Woodland, · 2019
Later among the works it cites.
“The second dihard diarization challenge: Dataset, task, and baselines,”
Neville Ryant, Kenneth Church, Christopher Cieri, Alejandrina Cristia, Jun Du, Sriram Ganapathy, and Mark Liberman, · 2019
Later among the works it cites.
“Dihard corpus,”
Ryant et al, · 2019
Later among the works it cites.