Fetching the paper…
Reading the bibliography…
In this paper we introduce a new natural language processing dataset and benchmark for predicting prosodic prominence from written text.
LibriTTS: A corpus derived from LibriSpeech for text-to-speech
Heiga Zen, Viet Dang, Rob Clark, Yu Zhang, Ron J Weiss, Ye Jia, Zhifeng Chen, and Yonghui Wu. 2019 · 1904
Earlier work this paper cites.
Some acoustic correlates of word stress in american english
Philip Lieberman. 1960 · 1960
Earlier work this paper cites.
The sound pattern of english
Noam Chomsky and Morris Halle. 1968 · 1968
Earlier work this paper cites.
Accent is predictable (if you’re a mind-reader)
Dwight Bolinger. 1972 · 1972
Earlier work this paper cites.
Sentence stress and syntactic transformations
Joan W Bresnan. 1973 · 1973
Earlier work this paper cites.
Pitch accent in context predicting intonational prominence from text
Julia Hirschberg. 1993 · 1993
Earlier work this paper cites.
The Boston University radio news corpus
Mari Ostendorf, Patti J Price, and Stefanie Shattuck-Hufnagel. 1995 · 1995
Earlier work this paper cites.
A probabilistic model of lexical and syntactic access and disambiguation
Daniel Jurafsky. 1996 · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
The perception of prosodic prominence
Jacques Terken and Dik Hermes. 2000 · 2000
Earlier work this paper cites.
Probabilistic relations between words: Evidence from reduction in lexical production
Daniel Jurafsky, Alan Bell, Michelle Gregory, and William D Raymond. 2001 · 2001
Earlier work this paper cites.
Learning to predict pitch accents and prosodic boundaries in dutch
Erwin Marsi, Martin Reynaert, Antal van den Bosch, Walter Daelemans, and Veronique Hoste. 2003 · 2003
Earlier work this paper cites.
Using conditional random fields to predict pitch accents in conversational speech
Michelle L Gregory and Yasemin Altun. 2004 · 2004
Cited alongside, same era.
Intertranscriber reliability of prosodic labeling on telephone conversation using tobi
Tae-Jin Yoon, Sandra Chavarria, Jennifer Cole, and Mark Hasegawa-Johnson. 2004 · 2004
Cited alongside, same era.
Pitch accent prediction: Effects of genre and speaker
Jiahong Yuan, Jason M Brenier, and Daniel Jurafsky. 2005 · 2005
Cited alongside, same era.
To memorize or to predict: Prominence labeling in conversational speech
Ani Nenkova, Jason Brenier, Anubha Kothari, Sasha Calhoun, Laura Whitton, David Beaver, and Dan Jurafsky. 2007 · 2007
Cited alongside, same era.
An acoustic measure for word prominence in spontaneous speech
Dagen Wang and Shrikanth Narayanan. 2007 · 2007
Cited alongside, same era.
Automatic prosodic labeling with conditional random fields and rich acoustic features
Efficient higher-order CRFs for morphological tagging
Thomas Mueller, Helmut Schmid, and Hinrich Schütze. 2013 · 2013
Later among the works it cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Later among the works it cites.
LibriSpeech: an ASR corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur. 2015 · 2015
Later among the works it cites.
Simple semi-supervised POS tagging
Karl Stratos and Michael Collins. 2015 · 2015
Later among the works it cites.
Analyzing the contribution of top-down lexical and bottom-up acoustic cues in the detection of sentence prominence
Sofoklis Kakouros, Joris Pelemans, Lyan Verwimp, Patrick Wambacq, and Okko Räsänen. 2016 · 2016
Later among the works it cites.
3pro–an unsupervised method for the automatic detection of sentence prominence in speech
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gina-Anne Levow. 2008 · 2008
Cited alongside, same era.
Naïve listeners’ prominence and boundary perception
Yoonsook Mo, Jennifer Cole, and Eun-Kyung Lee. 2008 · 2008
Cited alongside, same era.
Automatic detection and classification of prosodic events
Andrew Rosenberg. 2009 · 2009
Cited alongside, same era.
Experimental and theoretical advances in prosody: A review
Michael Wagner and Duane G Watson. 2010 · 2010
Cited alongside, same era.
Unsupervised Learning for Text-to-Speech Synthesis
Oliver Watts. 2012 · 2012
Cited alongside, same era.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, Greg S. Corrado, and Jeffrey Dean. 2013 · 2013
Cited alongside, same era.
Sofoklis Kakouros and Okko Räsänen. 2016 · 2016
Later among the works it cites.
Using continuous lexical embeddings to improve symbolic-prosody prediction in a text-to-speech front-end
Asaf Rendel, Raul Fernandez, Ron Hoory, and Bhuvana Ramabhadran. 2016 · 2016
Later among the works it cites.
Montreal Forced Aligner: Trainable text-speech alignment using kaldi
Michael McAuliffe, Michaela Socolof, Sarah Mihuc, Michael Wagner, and Morgan Sonderegger. 2017 · 2017
Later among the works it cites.
Hierarchical representation and estimation of prosody using continuous wavelet transform
Antti Suni, Juraj Šimko, Daniel Aalto, and Martti Vainio. 2017 · 2017
Later among the works it cites.
Effects of word embeddings on neural network-based pitch accent detection
Sabrina Stehwien, Ngoc Thang Vu, and Antje Schweitzer. 2018 · 2018
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Closest in time.