Fetching the paper…
Reading the bibliography…
Deep Neural Networks (DNN) have been successful in en- hancing noisy speech signals.
“Hearing lips and seeing voices,”
Harry McGurk and John MacDonald, · 1976
Earlier work this paper cites.
“Backpropagation through time: what it does and how to do it,”
Paul J Werbos, · 1990
Earlier work this paper cites.
“Lipreading and audio-visual speech perception,”
Quentin Summerfield, · 1992
Earlier work this paper cites.
“Stacked generalization,”
David H Wolpert, · 1992
Earlier work this paper cites.
“Long short-term memory,”
Sepp Hochreiter and Jürgen Schmidhuber, · 1997
Earlier work this paper cites.
“Long short-term memory,”
Sepp Hochreiter and Jürgen Schmidhuber, · 1997
Earlier work this paper cites.
Antony W Rix, John G Beerends, Michael P Hollier, and Andries P Hekstra, · 2001
Earlier work this paper cites.
“Audio-visual automatic speech recognition: An overview,”
Gerasimos Potamianos, Chalapathy Neti, Juergen Luettin, and Iain Matthews, · 2004
Earlier work this paper cites.
“Framewise phoneme classification with bidirectional lstm and other neural network architectures,”
Alex Graves and Jürgen Schmidhuber, · 2005
Earlier work this paper cites.
“Reducing the dimensionality of data with neural networks,”
Geoffrey E Hinton and Ruslan R Salakhutdinov, · 2006
Cited alongside, same era.
“Audiovisual database of spoken American English LDC2009V01,” Philadelphia: Linguistic Data Consortium, 2009,
Carolyn Richie, Sarah Warburton, and Megan Carter, · 2009
Cited alongside, same era.
“Theano: a CPU and GPU math expression compiler,”
James Bergstra, Olivier Breuleux, Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, Guillaume Desjardins, Joseph Turian, David Warde-Farley, and Yoshua Bengio, · 2010
Cited alongside, same era.
“Multimodal deep learning,”
Jiquan Ngiam, Aditya Khosla, Mingyu Kim, Juhan Nam, Honglak Lee, and Andrew Y Ng, · 2011
Cited alongside, same era.
“Scikit-learn: Machine learning in Python,”
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay, · 2011
Cited alongside, same era.
“Speech recognition with deep recurrent neural networks,”
Alan Graves, Abdel-rahman Mohamed, and Geoffrey Hinton, · 2013
Later among the works it cites.
“Ensemble deep learning for speech recognition,”
Li Deng and John C Platt, · 2014
Later among the works it cites.
“Adam: A method for stochastic optimization,”
Diederik Kingma and Jimmy Ba, · 2014
Later among the works it cites.
“Convolutional, long short-term memory, fully connected deep neural networks,”
Tara N Sainath, Oriol Vinyals, Andrew Senior, and Hasim Sak, · 2015
Later among the works it cites.
“A regression approach to speech enhancement based on deep neural networks,”
Yong Xu, Jun Du, Li-Rong Dai, and Chin-Hui Lee, · 2015
Later among the works it cites.
“librosa: 0.4.1,” Oct. 2015
Brian McFee, Matt McVicar, Colin Raffel, Dawen Liang, Oriol Nieto, Eric Battenberg, Josh Moore, Dan Ellis, Ryuichi Yamamoto, Rachel Bittner, Douglas Repetto, Petr Viktorin, João Felipe Santos, and Adrian Holovaty, · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nitish Srivastava and Ruslan R Salakhutdinov, · 2012
Cited alongside, same era.
“Imagenet classification with deep convolutional neural networks,”
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton, · 2012
Cited alongside, same era.
“Recurrent neural networks for noise reduction in robust asr.,”
Andrew L Maas, Quoc V Le, Tyler M O’Neil, Oriol Vinyals, Patrick Nguyen, and Andrew Y Ng, · 2012
Cited alongside, same era.
Later among the works it cites.
“Batch normalization: Accelerating deep network training by reducing internal covariate shift,”
Sergey Ioffe and Christian Szegedy, · 2015
Later among the works it cites.
“100 nonspeech sounds,”
Guoning Hu, · 2016
Closest in time.