Fetching the paper…
Reading the bibliography…
This paper describes a method for overlap-aware speaker diarization.
“Algebraic connectivity of graphs,”
Miroslav Fiedler, · 1973
Earlier work this paper cites.
“Matrix perturbation theory,”
G. W. Stewart and Ji-Guang Sun, · 1990
Earlier work this paper cites.
“On spectral clustering: Analysis and an algorithm,”
Andrew Y. Ng, Michael I. Jordan, and Yair Weiss, · 2001
Earlier work this paper cites.
“Multiclass spectral clustering,”
Stella X. Yu and Jianbo Shi, · 2003
Earlier work this paper cites.
“The ami meeting corpus: A pre-announcement,”
Jean Carletta, Simone Ashby, Sebastien Bourban, Mike Flynn, Maël Guillemot, Thomas Hain, Jaroslav Kadlec, Vasilis Karaiskos, Wessel Kraaij, Melissa Kronenthal, Guillaume Lathoud, Mike Lincoln, Agnes Lisowska Masson, Iain McCowan, Wilfried Post, Dennis Reidsma, and Pierre Wellner, · 2005
Earlier work this paper cites.
“An overview of automatic speaker diarization systems,”
Sue Tranter and Douglas A. Reynolds, · 2006
Earlier work this paper cites.
“A spectral clustering approach to speaker diarization,”
Huazhong Ning, Ming Liu, Hao Tang, and Thomas S. Huang, · 2006
Earlier work this paper cites.
“A tutorial on spectral clustering,”
Ulrike von Luxburg, · 2007
Earlier work this paper cites.
“Overlapped speech detection for improved speaker diarization in multiparty meetings,”
Kofi Boakye, Beatriz Trueba-Hornero, Oriol Vinyals, and Gerald Friedland, · 2008
Earlier work this paper cites.
“Speech overlap detection in a two-pass speaker diarization system,”
Marijn Huijbregts, David A. van Leeuwen, and Franciska de Jong, · 2009
Earlier work this paper cites.
“Speaker diarization exploiting the eigengap criterion and cluster ensembles,”
Nikoletta Bassiou, Vassiliki Moschou, and Constantine Kotropoulos, · 2010
Earlier work this paper cites.
“Front-end factor analysis for speaker verification,”
Najim Dehak, Patrick Kenny, Réda Dehak, Pierre Dumouchel, and Pierre Ouellet, · 2011
Earlier work this paper cites.
“The Kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al., · 2011
Earlier work this paper cites.
“Scikit-learn: Machine learning in Python,”
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay, · 2011
Earlier work this paper cites.
“Speaker diarization: A review of recent research,”
Xavier Anguera Miró, Simon Bozonnet, Nicholas W. D. Evans, Corinne Fredouille, Gerald Friedland, and Oriol Vinyals, · 2012
Earlier work this paper cites.
“Speaker diarization of overlapping speech based on silence distribution in meeting recordings,”
Sree Harsha Yella and Fabio Valente, · 2012
Cited alongside, same era.
“On the use of agglomerative and spectral clustering in speaker diarization of meetings,”
Jordi Luque and Javier Hernando, · 2012
Cited alongside, same era.
“On the use of spectral and iterative methods for speaker diarization,”
Stephen Shum, Najim Dehak, and Jim Glass, · 2012
Cited alongside, same era.
“Detecting overlapping speech with long short-term memory recurrent neural networks,”
Jürgen T. Geiger, Florian Eyben, Björn W. Schuller, and Gerhard Rigoll, · 2013
Cited alongside, same era.
“Deep neural networks for small footprint text-dependent speaker verification,”
Ehsan Variani, Xin Lei, Erik McDermott, Ignacio Lopez-Moreno, and Javier Gonzalez-Dominguez, · 2014
Cited alongside, same era.
“Speaker diarization with enhancing speech for the first dihard challenge,”
Lei Sun, Jun Du, Chao Jiang, Xueyang Zhang, Shan He, Bing Yin, and Chin-Hui Lee, · 2018
Later among the works it cites.
“X-vectors: Robust DNN embeddings for speaker recognition,”
David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, and Sanjeev Khudanpur, · 2018
Later among the works it cites.
“Speaker diarization based on Bayesian HMM with eigenvoice priors,”
Mireia Díez, Lukás Burget, and Pavel Matejka, · 2018
Later among the works it cites.
“Diarization is hard: Some experiences and lessons learned for the jhu team in the inaugural dihard challenge,”
Gregory Sell, David Snyder, Alan McCree, Daniel Garcia-Romero, Jesús Villalba, Matthew Maciejewski, Vimal Manohar, Najim Dehak, Daniel Povey, Shinji Watanabe, and Sanjeev Khudanpur, · 2018
Later among the works it cites.
“Detection of overlapping speech for the purposes of speaker diarization,”
Marie Kunesová, Marek Hrúz, Zbynek Zajíc, and Vlasta Radová, · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gregory Sell and Daniel Garcia-Romero, · 2015
Cited alongside, same era.
“A time delay neural network architecture for efficient modeling of long temporal contexts,”
Vijayaditya Peddinti, Daniel Povey, and Sanjeev Khudanpur, · 2015
Cited alongside, same era.
“Librispeech: An asr corpus based on public domain audio books,”
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur, · 2015
Cited alongside, same era.
“Musan: A music, speech, and noise corpus,”
David Snyder, Guoguo Chen, and Daniel Povey, · 2015
Cited alongside, same era.
“A comparative study of spectral clustering for i-vector-based speaker clustering under noisy conditions,”
Naohiro Tawara, Tetsuji Ogawa, and Tetsunori Kobayashi, · 2015
Cited alongside, same era.
“Speaker diarization using deep neural network embeddings,”
Daniel Garcia-Romero, David Snyder, Gregory Sell, Daniel Povey, and Alan McCree, · 2017
Cited alongside, same era.
“Detecting overlapped speech on short timeframes using deep learning,”
Valentin Andrei, Horia Cucu, and Corneliu Burileanu, · 2017
Cited alongside, same era.
Latané Bullock, Hervé Bredin, and L. Paola García-Perera, · 2019
Later among the works it cites.
“Bayesian hmm based x-vector clustering for speaker diarization,”
Mireia Diez, Lukás Burget, Shuai Wang, Johan Rohdin, and Jan Cernocký, · 2019
Later among the works it cites.
“Lstm based similarity measurement with spectral clustering for speaker diarization,”
Qingjian Lin, Ruiqing Yin, Ming Li, Hervé Bredin, and Claude Barras, · 2019
Later among the works it cites.
Yusuke Fujita, Shinji Watanabe, Shota Horiguchi, Yawen Xue, and Kenji Nagamatsu, · 2020
Closest in time.
“Speaker diarization with region proposal network,”
Zili Huang, Shinji Watanabe, Yusuke Fujita, Paola García, Yiwen Shao, Daniel Povey, and Sanjeev Khudanpur, · 2020
Closest in time.
“Auto-tuning spectral clustering for speaker diarization using normalized maximum eigengap,”
Tae Jin Park, Kyu J. Han, Manoj Kumar, and Shrikanth S. Narayanan, · 2020
Closest in time.
Ivan Medennikov, Maxim Korenevsky, Tatiana Prisyach, Yuri Y. Khokhlov, Mariya Korenevskaya, Ivan Sorokin, Tatiana V. Timofeeva, Anton Mitrofanov, Andrei Andrusenko, Ivan Podluzhny, Aleksandr Laptev, and Aleksei Romanenko, · 2020
Closest in time.
“Chime-6 challenge: Tackling multispeaker speech recognition for unsegmented recordings,”
Shinji Watanabe, Michael Mandel, Jon Barker, and Emmanuel Vincent, · 2020
Closest in time.
“Continuous speech separation: Dataset and analysis,”
Zhuo Chen, Takuya Yoshioka, Liang Lu, Tianyan Zhou, Zhong Meng, Yi Luo, J. Wu, and Jinyu Li, · 2020
Closest in time.
“Optimizing bayesian hmm based x-vector clustering for the second dihard speech diarization challenge,”
Mireia Díez, Lukás Burget, Federico Landini, Shuai Wang, and Jan Cernocký, · 2020
Closest in time.