Fetching the paper…
Reading the bibliography…
In this paper, we present an improved model for voicing silent speech, where audio is synthesized from facial electromyography (EMG) signals.
Fundamentals of speech recognition
Lawrence Rabiner and Biing-Hwang Juang. 1993 · 1993
Earlier work this paper cites.
Towards continuous speech recognition using surface electromyography
Szu-Chen Jou, Tanja Schultz, Matthias Walliczek, Florian Kraft, and Alex Waibel. 2006 · 2006
Earlier work this paper cites.
Synthesizing speech from electromyography using voice transformation techniques
Arthur R. Toth, Michael Wand, and Tanja Schultz. 2009 · 2009
Earlier work this paper cites.
Velum movement detection based on surface electromyography for speech interface
João Freitas, António JS Teixeira, Samuel S Silva, Catarina Oliveira, and Miguel Sales Dias. 2014 · 2014
Earlier work this paper cites.
Deep speech: Scaling up end-to-end speech recognition
Awni Hannun, Carl Case, Jared Casper, Bryan Catanzaro, Greg Diamos, Erich Elsen, Ryan Prenger, Sanjeev Satheesh, Shubho Sengupta, Adam Coates, et al. 2014 · 2014
Earlier work this paper cites.
Direct conversion from facial myoelectric signals to speech using deep neural networks
Lorenz Diener, Matthias Janke, and Tanja Schultz. 2015 · 2015
Earlier work this paper cites.
Wav2letter: an end-to-end convnet-based speech recognition system
Ronan Collobert, Christian Puhrsch, and Gabriel Synnaeve. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
WaveNet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew W. Senior, and Koray Kavukcuoglu. 2016 · 2016
Cited alongside, same era.
EMG-to-speech: Direct generation of speech from facial electromyographic signals
Matthias Janke and Lorenz Diener. 2017 · 2017
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2017 · 2017
Cited alongside, same era.
Montreal forced aligner: Trainable text-speech alignment using kaldi
Michael McAuliffe, Michaela Socolof, Sarah Mihuc, Michael Wagner, and Morgan Sonderegger. 2017 · 2017
Cited alongside, same era.
Development of sEMG sensors and algorithms for silent speech recognition
Geoffrey S Meltzner, James T Heaton, Yunbin Deng, Gianluca De Luca, Serge H Roy, and Joshua C Kline. 2018 · 2018
Later among the works it cites.
Self-attention with relative position representations
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani. 2018 · 2018
Later among the works it cites.
Multi-task learning of speech recognition and speech synthesis parameters for ultrasound-based silent speech interfaces
László Tóth, Gábor Gosztolya, Tamás Grósz, Alexandra Markó, and Tamás Gábor Csapó. 2018 · 2018
Later among the works it cites.
Speech synthesis from neural decoding of spoken sentences
Gopala K Anumanchipalli, Josh Chartier, and Edward F Chang. 2019 · 2019
Later among the works it cites.
wav2vec: Unsupervised pre-training for speech recognition
Steffen Schneider, Alexei Baevski, Ronan Collobert, and Michael Auli. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
David Gaddy and Dan Klein. 2020 · 2020
Later among the works it cites.