Fetching the paper…
Reading the bibliography…
This paper presents, a first of its kind, audio-visual (AV) speech enhacement challenge in real-noisy settings.
S. Boll, “A spectral subtraction algorithm for suppression of acoustic noise in speech,” in
1979
Earlier work this paper cites.
1985
Earlier work this paper cites.
I. Recommendation, “1534-1,“method for the subjective assessment of intermediate sound quality (mushra)”,”
2001
Earlier work this paper cites.
C. Sanderson, “The VidTIMIT database,” IDIAP, Tech. Rep., 2002
2002
Earlier work this paper cites.
E. Bailly-Bailliére, S. Bengio, F. Bimbot, M. Hamouz, J. Kittler, J. Mariéthoz, J. Matas, K. Messer, V. Popovici, F. Porée
2003
Cited alongside, same era.
B. Lee, M. Hasegawa-Johnson, C. Goudeseune, S. Kamdar, S. Borys, M. Liu, and T. S. Huang, “AVICAR: audio-visual speech corpus in a car environment.” in
2004
Cited alongside, same era.
M. Cooke, J. Barker, S. Cunningham, and X. Shao, “An audio-visual corpus for speech perception and automatic speech recognition,”
2006
Cited alongside, same era.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Cited alongside, same era.
J. Barker, R. Marxer, E. Vincent, and S. Watanabe, “The third chime speech separation and recognition challenge: Dataset, task and baselines,” in
2015
Later among the works it cites.
S. Pascual, A. Bonafonte, and J. Serra, “Segan: Speech enhancement generative adversarial network,”
2017
Later among the works it cites.
2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…