Fetching the paper…
Reading the bibliography…
The emergence of commercial tools for real-time performance-based 2D animation has enabled 2D characters to appear on live broadcasts and streaming platforms.
Modeling coarticulation in synthetic visual speech
Michael M Cohen, Dominic W Massaro, and others. 1993 · 1993
Earlier work this paper cites.
TIMIT acoustic-phonetic continuous speech corpus
John S Garofolo, Lori F Lamel, William M Fisher, Jonathan G Fiscus, David S Pallett, Nancy L Dahlgren, and Victor Zue. 1993 · 1993
Earlier work this paper cites.
Using dynamic time warping to find patterns in time series.. In KDD workshop
Donald J Berndt and James Clifford. 1994 · 1994
Earlier work this paper cites.
The illusion of life: Disney animation
Frank Thomas, Ollie Johnston, and Frank. Thomas. 1995 · 1995
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Miketalk: A talking facial display based on morphing visemes. In Computer Animation 98. Proceedings
Tony Ezzat and Tomaso Poggio. 1998 · 1998
Earlier work this paper cites.
Voice puppetry. In Proceedings of the 26th annual conference on Computer graphics and interactive techniques
Matthew Brand. 1999 · 1999
Earlier work this paper cites.
Picture my voice: Audio to visual speech synthesis using artificial neural networks. In AVSP’99-International Conference on Auditory-Visual Speech Processing
Dominic W Massaro, Jonas Beskow, Michael M Cohen, Christopher L Fry, and Tony Rodgriguez. 1999 · 1999
Earlier work this paper cites.
Audio-to-visual conversion using hidden markov models
Soonkyu Lee and DongSuk Yook. 2002 · 2002
Earlier work this paper cites.
Context dependent viseme models for voice driven animation. In Video/Image Processing and Multimedia Communications, 2003. 4th EURASIP Conference focused on
Xie Lei, Jiang Dongmei, Ilse Ravyse, Wemer Verhelst, Hichem Sahli, Velina Slavova, and Zhao Rongchun. 2003 · 2003
Earlier work this paper cites.
Sphinx-4: A flexible open source framework for speech recognition
Willie Walker, Paul Lamere, Philip Kwok, Bhiksha Raj, Rita Singh, Evandro Gouvea, Peter Wolf, and Joe Woelfel. 2004 · 2004
Earlier work this paper cites.
Expressive speech-driven facial animation
Yong Cao, Wen C Tien, Petros Faloutsos, and Frédéric Pighin. 2005 · 2005
Cited alongside, same era.
Framewise phoneme classification with bidirectional LSTM and other neural network architectures
Alex Graves and Jürgen Schmidhuber. 2005 · 2005
Cited alongside, same era.
Audio-based context recognition
Antti J Eronen, Vesa T Peltonen, Juha T Tuomi, Anssi P Klapuri, Seppo Fagerlund, Timo Sorsa, Gaëtan Lorho, and Jyri Huopaniemi. 2006 · 2006
Cited alongside, same era.
Torch7: A matlab-like environment for machine learning. In BigLearn, NIPS Workshop
Ronan Collobert, Koray Kavukcuoglu, and Clément Farabet. 2011 · 2011
Cited alongside, same era.
Phoneme-to-viseme Mapping for Visual Speech Recognition.. In ICPRAM (2)
Luca Cappelletta and Naomi Harte. 2012 · 2012
Cited alongside, same era.
Dynamic units of visual speech. In Proceedings of the 11th ACM SIGGRAPH/Eurographics conference on Computer Animation
Automatic speech recognition: A deep learning approach
Dong Yu and Li Deng. 2014 · 2014
Later among the works it cites.
Photo-real talking head with deep bidirectional LSTM. In Acoustics, Speech and Signal Processing (ICASSP), 2015 IEEE International Conference on
Bo Fan, Lijuan Wang, Frank K Soong, and Lei Xie. 2015 · 2015
Later among the works it cites.
A decision tree framework for spatiotemporal sequence prediction. In Proceedings of the 21th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining
Taehwan Kim, Yisong Yue, Sarah Taylor, and Iain Matthews. 2015 · 2015
Later among the works it cites.
Audiovisual speech synthesis: An overview of the state-of-the-art
Wesley Mattheyses and Werner Verhelst. 2015 · 2015
Later among the works it cites.
JALI: an animator-centric viseme model for expressive lip synchronization
Pif Edwards, Chris Landreth, Eugene Fiume, and Karan Singh. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sarah L Taylor, Moshe Mahler, Barry-John Theobald, and Iain Matthews. 2012 · 2012
Cited alongside, same era.
Hybrid speech recognition with deep bidirectional LSTM. In Automatic Speech Recognition and Understanding (ASRU), 2013 IEEE Workshop on
Alex Graves, Navdeep Jaitly, and Abdel-rahman Mohamed. 2013a · 2013
Cited alongside, same era.
Speech recognition with deep recurrent neural networks. In Acoustics, speech and signal processing (icassp), 2013 ieee international conference on
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton. 2013b · 2013
Cited alongside, same era.
A practical and configurable lip sync method for games. In Proceedings of Motion on Games
Yuyu Xu, Andrew W Feng, Stacy Marsella, and Ari Shapiro. 2013 · 2013
Cited alongside, same era.
Towards end-to-end speech recognition with recurrent neural networks. In Proceedings of the 31st International Conference on Machine Learning (ICML-14)
Alex Graves and Navdeep Jaitly. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
How ’The Simpsons’ Used Adobe Character Animator To Create A Live Episode
Ian Failes. 2016 · 2016
Later among the works it cites.
Voice Animator: Automatic Lip-Synching in Limited Animation by Audio. In International Conference on Advances in Computer Entertainment
Shoichi Furukawa, Tsukasa Fukusato, Shugo Yamaguchi, and Shigeo Morishima. 2017 · 2017
Later among the works it cites.
Audio-driven facial animation by joint end-to-end learning of pose and emotion
Tero Karras, Timo Aila, Samuli Laine, Antti Herva, and Jaakko Lehtinen. 2017 · 2017
Later among the works it cites.
Synthesizing obama: learning lip sync from audio
Supasorn Suwajanakorn, Steven M Seitz, and Ira Kemelmacher-Shlizerman. 2017 · 2017
Later among the works it cites.
A deep learning approach for generalized speech animation
Sarah Taylor, Taehwan Kim, Yisong Yue, Moshe Mahler, James Krahe, Anastasio Garcia Rodriguez, Jessica Hodgins, and Iain Matthews. 2017 · 2017
Later among the works it cites.
Speech Recognition — A comparison of popular services in EN and NL
Bjorn Vuylsteker. 2017 · 2017
Later among the works it cites.