Fetching the paper…
Reading the bibliography…
We propose a new speaker diarization system based on a recently introduced unsupervised clustering technique namely, generative adversarial network mixture model (GANMM).
Adam Janin, Don Baron, Jane Edwards, Dan Ellis, David Gelbart, Nelson Morgan, Barbara Peskin, Thilo Pfau, Elizabeth Shriberg, Andreas Stolcke, et al., · 2003
Earlier work this paper cites.
“The Rich Transcription 2006 spring meeting recognition evaluation,”
Jonathan G Fiscus, Jerome Ajot, Martial Michel, and John S Garofolo, · 2006
Earlier work this paper cites.
Segmentation, diarization and speech transcription: surprise data unraveled,
Marijn Huijbregts, · 2008
Earlier work this paper cites.
“Stream-based speaker segmentation using speaker factors and eigenvoices,”
Fabio Castaldo, Daniele Colibro, Emanuele Dalmasso, Pietro Laface, and Claudio Vair, · 2008
Earlier work this paper cites.
“An information theoretic approach to speaker diarization of meeting data,”
Deepu Vijayasenan, Fabio Valente, and Hervé Bourlard, · 2009
Earlier work this paper cites.
“BIC-based speaker segmentation using divide-and-conquer strategies with application to speaker diarization,”
Shih-Sian Cheng, Hsin-Min Wang, and Hsin-Chia Fu, · 2010
Earlier work this paper cites.
“The Kaldi speech recognition toolkit,”
Daniel Povey et al., · 2011
Earlier work this paper cites.
“A global optimization framework for speaker diarization,”
Mickael Rouvier and Sylvain Meignier, · 2012
Earlier work this paper cites.
“Speaker-adaptive speech recognition using speaker diarization for improved transcription of large spoken archives,”
Petr Cerva, Jan Silovsky, Jindrich Zdansky, Jan Nouza, and Ladislav Seps, · 2013
Earlier work this paper cites.
“Unsupervised methods for speaker diarization: An integrated and iterative approach,”
Stephen H Shum, Najim Dehak, Réda Dehak, and James R Glass, · 2013
Earlier work this paper cites.
“A study of the cosine distance-based mean shift for telephone speech diarization,”
Mohammed Senoussaoui, Patrick Kenny, Themos Stafylakis, and Pierre Dumouchel, · 2014
Earlier work this paper cites.
“Speaker diarization with PLDA i-vector scoring and unsupervised calibration,”
Gregory Sell and Daniel Garcia-Romero, · 2014
Cited alongside, same era.
“Generative adversarial nets,”
Ian Goodfellow et al., · 2014
Cited alongside, same era.
“Speaker diarization with i-vectors from DNN senone posteriors,”
Gregory Sell, Daniel Garcia-Romero, and Alan McCree, · 2015
Cited alongside, same era.
“Investigating various diarization algorithms for speaker in the wild (SITW) speaker recognition challenge,”
Yi Liu, Yao Tian, Liang He, and Jia Liu, · 2016
Cited alongside, same era.
“Unrolled generative adversarial networks,”
Luke Metz, Ben Poole, David Pfau, and Jascha Sohl-Dickstein, · 2016
Cited alongside, same era.
“Speaker diarization using deep neural network embeddings,”
“Multi-generator generative adversarial nets,”
Quan Hoang, Tu Dinh Nguyen, Trung Le, and Dinh Phung, · 2017
Later among the works it cites.
“Generalization and equilibrium in generative adversarial nets (GANs),”
Sanjeev Arora, Rong Ge, Yingyu Liang, Tengyu Ma, and Yi Zhang, · 2017
Later among the works it cites.
“Diarization is hard: Some experiences and lessons learned for the JHU team in the inaugural DIHARD challenge,”
Gregory Sell et al., · 2018
Later among the works it cites.
“Speaker diarization with LSTM,”
Quan Wang, Carlton Downey, Li Wan, Philip Andrew Mansfield, and Ignacio Lopz Moreno, · 2018
Later among the works it cites.
“Fully supervised speaker diarization,”
Aonan Zhang, Quan Wang, Zhenyao Zhu, John Paisley, and Chong Wang, · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Daniel Garcia-Romero, David Snyder, Gregory Sell, Daniel Povey, and Alan McCree, · 2017
Cited alongside, same era.
“Speaker diarization using deep recurrent convolutional neural networks for speaker embeddings,”
Pawel Cyrta, Tomasz Trzciński, and Wojciech Stokowiec, · 2017
Cited alongside, same era.
“Tristounet: triplet loss for speaker turn embedding,”
Hervé Bredin, · 2017
Cited alongside, same era.
“Speaker diarization using convolutional neural network for statistics accumulation refinement,”
Zbynek Zajíc, Marek Hrúz, and Ludek Müller, · 2017
Cited alongside, same era.
“Developing on-line speaker diarization system,”
Dimitrios Dimitriadis and Petr Fousek, · 2017
Cited alongside, same era.
“Wasserstein generative adversarial networks,”
Martin Arjovsky, Soumith Chintala, and Léon Bottou, · 2017
Cited alongside, same era.
“X-vectors: Robust DNN embeddings for speaker recognition,”
David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, and Sanjeev Khudanpur, · 2018
Later among the works it cites.
“Links: A high-dimensional online clustering method,”
Philip Andrew Mansfield, Quan Wang, Carlton Downey, Li Wan, and Ignacio Lopez Moreno, · 2018
Later among the works it cites.
“Neural speech turn segmentation and affinity propagation for speaker diarization,”
Ruiqing Yin, Hervé Bredin, and Claude Barras, · 2018
Later among the works it cites.
“Mixture of GANs for clustering,”
Yang Yu and Wen-Ji Zhou, · 2018
Later among the works it cites.
“Fully supervised speaker diarization,”
Aonan Zhang, Quan Wang, Zhenyao Zhu, John Paisley, and Chong Wang, · 2019
Closest in time.