Fetching the paper…
Reading the bibliography…
Phoneme boundary detection plays an essential first step for a variety of speech processing applications such as speaker diarization, speech science, keyword spotting, etc.
“Timit acoustic phonetic continuous speech corpus,”
J. S Garofolo, · 1993
Earlier work this paper cites.
“Transcribing radio news,”
Francis Kubala, Tasos Anastasakos, Hubert Jin, Long Nguyen, and Richard Schwartz, · 1996
Earlier work this paper cites.
“A new text-independent method for phoneme segmentation,”
Guido Aversano, Anna Esposito, and M Marinaro, · 2001
Earlier work this paper cites.
“Real-time pitch determination of one or more voices by nonnegative matrix factorization,”
Fei Sha and Lawrence K Saul, · 2005
Earlier work this paper cites.
“The buckeye corpus of conversational speech: Labeling conventions and a test of transcriber reliability,”
M. A Pitt et al., · 2005
Earlier work this paper cites.
“Phoneme alignment based on discriminative learning,”
Joseph Keshet, Shai Shalev-Shwartz, Yoram Singer, and Dan Chazan, · 2005
Earlier work this paper cites.
“On the relation between maximum spectral transition positions and phone boundaries,”
Sorin Dusan and Lawrence Rabiner, · 2006
Earlier work this paper cites.
“A tutorial on energy-based learning,”
Yann LeCun, Sumit Chopra, Raia Hadsell, M Ranzato, and F Huang, · 2006
Earlier work this paper cites.
“Finding maximum margin segments in speech,”
Yago Pereiro Estevan, Vincent Wan, and Odette Scharenborg, · 2007
Earlier work this paper cites.
“Phonemic segmentation using the generalised gamma distribution and small sample bayesian information criterion,”
George Almpanidis and Constantine Kotropoulos, · 2008
Earlier work this paper cites.
“Discriminative keyword spotting,”
Joseph Keshet, David Grangier, and Samy Bengio, · 2009
Cited alongside, same era.
“Audio segmentation for speech recognition using segment features,”
David Rybach, Christian Gollan, Ralf Schluter, and Hermann Ney, · 2009
Cited alongside, same era.
“An improved speech segmentation quality measure: the r-value,”
O. J. Räsänen, U. K. Laine, and T. Altosaar, · 2009
Cited alongside, same era.
“Blind segmentation of speech using non-linear filtering methods,”
Okko Räsänen, Unto K Laine, and Toomas Altosaar, · 2011
Cited alongside, same era.
“A review on speaker diarization systems and approaches,”
Mohammad H Moattar and Mohammad M Homayounpour, · 2012
Cited alongside, same era.
“Accurate speech segmentation by mimicking human auditory processing,”
Sarah King and Mark Hasegawa-Johnson, · 2013
Cited alongside, same era.
“Automatic measurement of vowel duration via structured prediction,”
Yossi Adi, Joseph Keshet, Emily Cibelli, Erin Gustafson, Cynthia Clopper, and Matthew Goldrick, · 2016
Later among the works it cites.
“Blind phoneme segmentation with temporal prediction errors,”
Paul Michel, Okko Räsänen, Roland Thiolliere, and Emmanuel Dupoux, · 2016
Later among the works it cites.
“Text-independent phoneme segmentation combining egg and speech data,”
Lijiang Chen, Xia Mao, and Hong Yan, · 2016
Later among the works it cites.
“Simple and accurate dependency parsing using bidirectional lstm feature representations,”
Eliyahu Kiperwasser and Yoav Goldberg, · 2016
Later among the works it cites.
“Phoneme boundary detection using deep bidirectional lstms,”
Joerg Franke, Markus Mueller, Fatima Hamlaoui, Sebastian Stueker, and Alex Waibel, · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Automatic tools for analyzing spoken hebrew,”
A. Ben-Shalom, J. Keshet, D. Modan, and A. Laufer, · 2014
Cited alongside, same era.
“Vowel duration measurement using deep neural networks,”
Yossi Adi, Joseph Keshet, and Matthew Goldrick, · 2015
Cited alongside, same era.
“Blind phone segmentation based on spectral change detection using legendre polynomial approximation,”
Dac-Thang Hoang and Hsiao-Chuan Wang, · 2015
Cited alongside, same era.
Liang Lu, Lingpeng Kong, Chris Dyer, Noah A Smith, and Steve Renals, · 2016
Later among the works it cites.
“Structed: risk minimization in structured prediction,”
Yossi Adi and Joseph Keshet, · 2016
Later among the works it cites.
“Sequence segmentation using joint rnn and structured prediction models,”
Yossi Adi, Joseph Keshet, Emily Cibelli, and Matthew Goldrick, · 2017
Later among the works it cites.
“Montreal forced aligner: Trainable text-speech alignment using kaldi.,”
Michael McAuliffe, Michaela Socolof, Sarah Mihuc, Michael Wagner, and Morgan Sonderegger, · 2017
Later among the works it cites.