Fetching the paper…
Reading the bibliography…
This report introduces a new corpus of music, speech, and noise.
1996 english broadcast news speech (hub4)
J. F. W. F. David Graff, John Garofolo and D. Pallett · 1997
Earlier work this paper cites.
http://www.itl.nist.gov/iad/mig/tests/sre/2010/ , 2010
The NIST year 2010 speaker recognition evaluation plan · 2010
Earlier work this paper cites.
The million song dataset
T. Bertin-Mahieux, D. P. Ellis, B. Whitman, and P. Lamere · 2011
Earlier work this paper cites.
The Kaldi speech recognition toolkit
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlíček, Y. Qian, P. Schwarz, et al · 2011
Cited alongside, same era.
Supervised/unsupervised voice activity detectors for text-dependent speaker recognition on the rsr2015 corpus
J. Alam, P. Kenny, P. Ouellet, T. Stafylakis, and P. Dumouchel · 2014
Cited alongside, same era.
Music tonality features for speech/music discrimination
G. Sell and P. Clark · 2014
Cited alongside, same era.
Time delay deep neural network-based universal background models for speaker recognition
D. P. David Snyder, Daniel Garcia-Romero · 2015
Closest in time.
Gtzan music/speech
G. Tzanetakis · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…