Fetching the paper…
Reading the bibliography…
We present an introspection of an audiovisual speech enhancement model.
“Audio visual speech recognition,”
Chalapathy Neti, Gerasimos Potamianos, Juergen Luettin, Iain Matthews, Herve Glotin, Dimitra Vergyri, June Sison, and Azad Mashari, · 2000
Earlier work this paper cites.
“Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,”
Antony W Rix, John G Beerends, Michael P Hollier, and Andries P Hekstra, · 2001
Earlier work this paper cites.
“Dlib-ml: A machine learning toolkit,”
Davis E King, · 2009
Earlier work this paper cites.
“Twin-hmm-based audio-visual speech enhancement,”
Ahmed Hussen Abdelaziz, Steffen Zeiler, and Dorothea Kolossa, · 2013
Earlier work this paper cites.
Speech enhancement: theory and practice
Philipos C Loizou, · 2013
Earlier work this paper cites.
“Robust asr using neural network based speech enhancement and feature simulation,”
Sunit Sivasankaran, Aditya Arie Nugraha, Emmanuel Vincent, Juan A Morales-Cordovilla, Siddharth Dalmia, Irina Illina, and Antoine Liutkus, · 2015
Earlier work this paper cites.
“Introducing the turbo-twin-hmm for audio-visual speech enhancement,”
Steffen Zeiler, Hendrik Meutzner, Ahmed Hussen Abdelaziz, and Dorothea Kolossa, · 2016
Cited alongside, same era.
“Out of time: automated lip sync in the wild,”
Joon Son Chung and Andrew Zisserman, · 2016
Cited alongside, same era.
“Colorful image colorization,”
Richard Zhang, Phillip Isola, and Alexei A Efros, · 2016
Cited alongside, same era.
“Speech enhancement for robust automatic speech recognition: Evaluation using a baseline system and instrumental measures,”
Alastair H Moore, P Peso Parada, and Patrick A Naylor, · 2017
Cited alongside, same era.
“Supervised speech separation based on deep learning: An overview,”
DeLiang Wang and Jitong Chen, · 2018
Cited alongside, same era.
“Visual speech enhancement,”
Aviv Gabbay, Asaph Shamir, and Shmuel Peleg, · 2018
Later among the works it cites.
“The conversation: Deep audio-visual speech enhancement,”
Triantafyllos Afouras, Joon Son Chung, and Andrew Zisserman, · 2018
Later among the works it cites.
“Looking to listen at the cocktail party: a speaker-independent audio-visual model for speech separation,”
Ariel Ephrat, Inbar Mosseri, Oran Lang, Tali Dekel, Kevin Wilson, Avinatan Hassidim, William T Freeman, and Michael Rubinstein, · 2018
Later among the works it cites.
“VoiceID loss: Speech enhancement for speaker verification,”
Suwon Shon, Hao Tang, and James Glass, · 2019
Later among the works it cites.
“Convolutional neural network-based speech enhancement for cochlear implant recipients,”
Nursadul Mamun, Soheil Khorram, and John HL Hansen, · 2019
Later among the works it cites.
“My lips are concealed: Audio-visual speech enhancement through obstructions,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ke Tan and DeLiang Wang, · 2018
Cited alongside, same era.
Triantafyllos Afouras, Joon Son Chung, and Andrew Zisserman, · 2019
Later among the works it cites.