Fetching the paper…
Reading the bibliography…
We address the problem of effectively handling overlapping speech in a diarization system.
“Efficient use of overlap information in speaker diarization,”
Scott Otterson and Mari Ostendorf, · 2007
Earlier work this paper cites.
“Unleashing the killer corpus: experiences in creating the multi-everything AMI Meeting Corpus,”
Jean Carletta, · 2007
Earlier work this paper cites.
“Overlapped speech detection for improved speaker diarization in multiparty meetings,”
Kofi Boakye, Beatriz Trueba-Hornero, Oriol Vinyals, and Gerald Friedland, · 2008
Earlier work this paper cites.
“Speech overlap detection in a two-pass speaker diarization system,”
Marijn Huijbregts, David A. van Leeuwen, and Franciska de Jong, · 2009
Earlier work this paper cites.
“Speaker diarization of overlapping speech based on silence distribution in meeting recordings,”
Sree Harsha Yella and Fabio Valente, · 2012
Earlier work this paper cites.
“The ETAPE Corpus for the Evaluation of Speech-based TV Content Processing in the French Language,”
Guillaume Gravier, Gilles Adda, Niklas Paulson, Matthieu Carré, Aude Giraudel, and Olivier Galibert, · 2012
Earlier work this paper cites.
“Detecting overlapping speech with long short-term memory recurrent neural networks,”
Jürgen T Geiger, Florian Eyben, Björn Schuller, and Gerhard Rigoll, · 2013
Earlier work this paper cites.
“Impact of overlapping speech detection on speaker diarization for broadcast news and debates,”
D. Charlet, C. Barras, and J. Liénard, · 2013
Earlier work this paper cites.
“Diarization resegmentation in the factor analysis subspace,”
Gregory Sell and Daniel Garcia-Romero, · 2015
Cited alongside, same era.
“Detecting Overlapped Speech on Short Timeframes Using Deep Learning,”
Valentin Andrei, Horia Cucu, and Corneliu Burileanu, · 2017
Cited alongside, same era.
“Enhancing lstm rnn-based speech overlap detection by artificially mixed data,”
Gerhard Hagerer, Vedhas Pandit, Florian Eyben, and Björn Schuller, · 2017
Cited alongside, same era.
“pyannote.metrics: a toolkit for reproducible evaluation, diagnostic, and error analysis of speaker diarization systems,”
Hervé Bredin, · 2017
Cited alongside, same era.
“First dihard challenge evaluation plan,” 2018
Neville Ryant, Kenneth Church, Christopher Cieri, Alejandrina Cristia, Jun Du, Sriram Ganapathy, and Mark Liberman, · 2018
Cited alongside, same era.
“Neural Speech Turn Segmentation and Affinity Propagation for Speaker Diarization,”
“Speaker Diarization based on Bayesian HMM with Eigenvoice Priors,”
Mireia Diez, Lukas Burget, and Pavel Matejka, · 2018
Later among the works it cites.
“Detection of overlapping speech for the purposes of speaker diarization,”
Marie Kunešová, Marek Hrúz, Zbyněk Zajíc, and Vlasta Radová, · 2019
Closest in time.
“Second dihard challenge evaluation plan,”
Neville Ryant, Kenneth Church, Christopher Cieri, Alejandrina Cristia, Jun Du, Sriram Ganapathy, and Mark Liberman, · 2019
Closest in time.
“The Second DIHARD Diarization Challenge: Dataset, Task, and Baselines,”
Neville Ryant, Kenneth Church, Christopher Cieri, Alejandrina Cristia, Jun Du, Sriram Ganapathy, and Mark Liberman, · 2019
Closest in time.
“Personal vad: Speaker-conditioned voice activity detection,”
Shaojin Ding, Quan Wang, Shuo-yiin Chang, Li Wan, and Ignacio Lopez Moreno, · 2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ruiqing Yin, Hervé Bredin, and Claude Barras, · 2018
Cited alongside, same era.
“Speaker recognition from raw waveform with sincnet,”
Mirco Ravanelli and Yoshua Bengio, · 2018
Cited alongside, same era.
“Countnet: Estimating the number of concurrent speakers using supervised learning,”
Fabian-Robert Stöter, Soumitro Chakrabarty, Bernd Andreas Edler, and Emanuël A. P. Habets, · 2019
Closest in time.
“pyannote.audio: Neural Building Blocks for Speaker Diarization,”
pyannote.audio contributors, · 2020
Closest in time.