Fetching the paper…
Reading the bibliography…
This work presents a scalable solution to open-vocabulary visual speech recognition.
Automatic lipreading by optical-flow analysis
Kenji Mase and Alex Pentland · 1991
Earlier work this paper cites.
Computer-coding the IPA: A proposed extension of SAMPA
John C Wells · 1995
Earlier work this paper cites.
Continuous automatic speech recognition by lipreading
Alan J Goldschen, Oscar N Garcia, and Eric D Petajan · 1997
Earlier work this paper cites.
Dynamic features for visual speechreading: A systematic comparison
Michael S Gray, Javier R Movellan, and Terrence J Sejnowski · 1997
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Speaker independent audio-visual database for bimodal asr
Gerasimos Potamianos, Eric Cosatto, Hans Peter Graf, and David B Roe · 1997
Earlier work this paper cites.
Discriminative training of hmm stream exponents for audio-visual speech recognition
Gerasimos Potamianos and Hans Peter Graf · 1998
Earlier work this paper cites.
An image transform approach for hmm based automatic lipreading
Gerasimos Potamianos, Hans Peter Graf, and Eric Cosatto · 1998
Earlier work this paper cites.
Bimodal speech recognition using coupled hidden Markov models
Stephen M Chu and Thomas S Huang · 2000
Earlier work this paper cites.
Audio visual speech recognition
Chalapathy Neti, Gerasimos Potamianos, Juergen Luettin, Iain Matthews, Herve Glotin, Dimitra Vergyri, June Sison, and Azad Mashari · 2000
Earlier work this paper cites.
Diatom autofocusing in brightfield microscopy: A comparative study
José Luis Pech-Pacheco, Gabriel Cristóbal, Jesús Chamorro-Martinez, and Joaquín Fernández-Valdivia · 2000
Earlier work this paper cites.
Extraction of visual features for lipreading
Iain Matthews, Timothy F Cootes, J Andrew Bangham, Stephen Cox, and Richard Harvey · 2002
Earlier work this paper cites.
Weighted finite-state transducers in speech recognition
Mehryar Mohri, Fernando Pereira, and Michael Riley · 2002
Earlier work this paper cites.
Video shot boundary detection based on color histogram
Jordi Mas and Gabriel Fernandez · 2003
Earlier work this paper cites.
Audio-visual speech recognition using lip movement extracted from side-face images
Tomoaki Yoshinaga, Satoshi Tamura, Koji Iwano, and Sadaoki Furui · 2003
Earlier work this paper cites.
Audio-visual automatic speech recognition: An overview
Gerasimos Potamianos, Chalapathy Neti, Juergen Luettin, and Iain Matthews · 2004
Earlier work this paper cites.
Multi-modal speech recognition using optical-flow analysis for lip images
Satoshi Tamura, Koji Iwano, and Sadaoki Furui · 2004
Earlier work this paper cites.
An audio-visual corpus for speech perception and automatic speech recognition
Martin Cooke, Jon Barker, Stuart Cunningham, and Xu Shao · 2006
Earlier work this paper cites.
Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber · 2006
Earlier work this paper cites.
Patch-based representation of visual speech
Patrick Lucey and Sridha Sridharan · 2006
Earlier work this paper cites.
Adaptive multimodal fusion by uncertainty compensation
Vassilis Pitsikalis, Athanassios Katsamanis, George Papandreou, and Petros Maragos · 2006
Earlier work this paper cites.
Multimodal fusion and learning with uncertain features applied to audiovisual speech recognition
George Papandreou, Athanassios Katsamanis, Vassilis Pitsikalis, and Petros Maragos · 2007
Earlier work this paper cites.
An automatic lipreading system for spoken digits with limited training data
Shi-Lin Wang, Alan Wee-Chung Liew, Wai H Lau, and Shu Hung Leung · 2008
Earlier work this paper cites.
Information theoretic feature extraction for audio-visual speech recognition
Mihai Gurban and Jean-Philippe Thiran · 2009
Earlier work this paper cites.
Adaptive multimodal fusion by uncertainty compensation with application to audiovisual speech recognition
George Papandreou, Athanassios Katsamanis, Vassilis Pitsikalis, and Petros Maragos · 2009
Earlier work this paper cites.
Lipreading with local spatiotemporal descriptors
Guoying Zhao, Mark Barnard, and Matti Pietikainen · 2009
Earlier work this paper cites.
Flumejava: easy, efficient data-parallel pipelines
Craig Chambers, Ashish Raniwala, Frances Perry, Stephen Adams, Robert R Henry, Robert Bradshaw, and Nathan Weizenbaum · 2010
Earlier work this paper cites.
A study of influence of word lip reading by change of frame rate
Takeshi Saitoh and Ryosuke Konishi · 2010
Cited alongside, same era.
Lip reading using optical flow and support vector machines
Ayaz A Shaikh, Dinesh K Kumar, Wai C Yau, MZ Che Azemin, and Jayavardhana Gubbi · 2010
Cited alongside, same era.
Visual speech recognition
Ahmad BA Hassanat · 2011
Cited alongside, same era.
Multimodal deep learning
Jiquan Ngiam, Aditya Khosla, Mingyu Kim, Juhan Nam, Honglak Lee, and Andrew Y Ng · 2011
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Geoffrey Hinton, Li Deng, Dong Yu, George E Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara N Sainath, et al · 2012
Cited alongside, same era.
Large scale deep neural network acoustic modeling with semi-supervised training data for YouTube video transcription
Audio-visual speech recognition using bimodal-trained bottleneck features for a person with severe hearing loss
Yuki Takashima, Ryo Aihara, Tetsuya Takiguchi, Yasuo Ariki, Nobuyuki Mitani, Kiyohiro Omori, and Kaoru Nakazono · 2016
Later among the works it cites.
Lipreading with long short-term memory
Michael Wand, Jan Koutnik, and J urgen Schmidhuber · 2016
Later among the works it cites.
Lip2AudSpec: Speech reconstruction from silent lip movements video
Hassan Akbari, Himani Arora, Liangliang Cao, and Nima Mesgarani · 2017
Later among the works it cites.
LipNet: End-to-end sentence-level lipreading
Yannis M Assael, Brendan Shillingford, Shimon Whiteson, and Nando de Freitas · 2017
Later among the works it cites.
Phoneme-to-viseme mappings: the good, the bad, and the ugly
Helen L Bear and Richard Harvey · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hank Liao, Erik McDermott, and Andrew Senior · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Silence in the EHR: Infrequent documentation of aphonia in the electronic health record
Megan A. Morris and Abel N. Kho · 2014
Cited alongside, same era.
Lipreading using convolutional neural network
Kuniaki Noda, Yuki Yamaguchi, Kazuhiro Nakadai, Hiroshi G Okuno, and Tetsuya Ogata · 2014
Cited alongside, same era.
Striving for simplicity: The all convolutional net
Jost Tobias Springenberg, Alexey Dosovitskiy, Thomas Brox, and Martin Riedmiller · 2014
Cited alongside, same era.
The effect of speaking rate on audio and visual speech
Sarah Taylor, Barry-John Theobald, and Iain Matthews · 2014
Cited alongside, same era.
A review of recent advances in visual speech decoding
Z Zhou, G Zhao, X Hong, and M Pietikäinen · 2014
Cited alongside, same era.
Joon Son Chung and Andrew Zisserman · 2017
Later among the works it cites.
Lip reading sentences in the wild
Joon Son Chung, Andrew Senior, Oriol Vinyals, and Andrew Zisserman · 2017
Later among the works it cites.
Vid2speech: Speech reconstruction from silent video
Ariel Ephrat and Shmuel Peleg · 2017
Later among the works it cites.
Visual speech enhancement using noise-invariant training
Aviv Gabbay, Asaph Shamir, and Shmuel Peleg · 2017
Later among the works it cites.
Exploring ROI size in deep learning based lipreading
Alexandros Koumparoulis, Gerasimos Potamianos, Youssef Mroueh, and Steven J Rennie · 2017
Later among the works it cites.
Generating intelligible audio speech from visual speech
Thomas Le Cornu and Ben Milner · 2017
Later among the works it cites.
Utilizing lipreading in large vocabulary continuous speech recognition
Karel Paleček · 2017
Later among the works it cites.
End-to-end multi-view lipreading
Stavros Petridis, Yujiang Wang, Zuwei Li, and Maja Pantic · 2017
Later among the works it cites.
A comparison of sequence-to-sequence models for speech recognition
Rohit Prabhavalkar, Kanishka Rao, Tara Sainath, Bo Li, Leif Johnson, and Navdeep Jaitly · 2017
Later among the works it cites.
Exploring architectures, data and units for streaming end-to-end speech recognition with RNN-transducer
Kanishka Rao, Haşim Sak, and Rohit Prabhavalkar · 2017
Later among the works it cites.
Neural speech recognizer: Acoustic-to-word LSTM model for large vocabulary speech recognition
Hagen Soltau, Hank Liao, and Hasim Sak · 2017
Later among the works it cites.
Combining residual networks with LSTMs for lipreading
Themos Stafylakis and Georgios Tzimiropoulos · 2017
Later among the works it cites.
Audio visual speech recognition using deep recurrent neural networks
Abhinav Thanda and Shankar M. Venkatesan · 2017
Later among the works it cites.
3d convolutional neural networks for cross audio-visual matching recognition
Amirsina Torfi, Seyed Mehdi Iranmanesh, Nasser Nasrabadi, and Jeremy Dawson · 2017
Later among the works it cites.
Improving speaker-independent lipreading with domain-adversarial training
Michael Wand and Jürgen Schmidhuber · 2017
Later among the works it cites.
Ariel Ephrat, Inbar Mosseri, Oran Lang, Tali Dekel, Kevin Wilson, Avinatan Hassidim, William T Freeman, and Michael Rubinstein · 2018
Closest in time.
Hospital inpatient national statistics
HCUPnet · 2018
Closest in time.
Compact language detector v3
Alex Salcianu, Andy Golding, Anton Bakalov, Chris Alberti, Daniel Andor, David Weiss, Emily Pitler, Greg Coppola, Jason Riesa, Kuzman Ganchev, Michael Ringgaard, Nan Hua, Ryan McDonald, Slav Petrov, Stefan Istrate, and Terry Koo · 2018
Closest in time.
New safety collaborative will improve outcomes for patients with tracheostomies
The Health Foundation · 2018
Closest in time.
Yuxin Wu and Kaiming He · 2018
Closest in time.
LCANet: End-to-end lipreading with cascaded attention-ctc
Kai Xu, Dawei Li, Nick Cassimatis, and Xiaolong Wang · 2018
Closest in time.