Fetching the paper…
Reading the bibliography…
Multi-channel speech enhancement with ad-hoc sensors has been a challenging task.
Extrapolation, interpolation, and smoothing of stationary time series
Norbert Wiener, · 1949
Earlier work this paper cites.
“Image method for efficiently simulating small-room acoustics,”
Jont B Allen and David A Berkley, · 1979
Earlier work this paper cites.
“An alternative approach to linearly constrained adaptive beamforming,”
Lloyd Griffiths and CW Jim, · 1982
Earlier work this paper cites.
“Improved mvdr beamforming using single-channel mask prediction networks.,”
Hakan Erdogan, John R Hershey, Shinji Watanabe, Michael I Mandel, and Jonathan Le Roux, · 1985
Earlier work this paper cites.
“DARPA TIMIT acoustic-phonetic continous speech corpus CD-ROM. NIST speech disc 1-1.1,”
John S Garofolo, Lori F Lamel, William M Fisher, Jonathon G Fiscus, and David S Pallett, · 1993
Earlier work this paper cites.
“Speech dereverberation via maximum-kurtosis subband adaptive filtering,”
Bradford W Gillespie, Henrique S Malvar, and Dinei AF Florêncio, · 2001
Earlier work this paper cites.
“Blind source separation exploiting higher-order frequency dependencies,”
Taesu Kim, Hagai T Attias, Soo-Young Lee, and Te-Won Lee, · 2007
Earlier work this paper cites.
“Beamforming with a maximum negentropy criterion,”
Kenichi Kumatani, John McDonough, Barbara Rauch, Dietrich Klakow, Philip N Garner, and Weifeng Li, · 2009
Earlier work this paper cites.
“Diffuse reverberation model for efficient image-source simulation of room impulse responses,”
Eric A Lehmann and Anders M Johansson, · 2010
Earlier work this paper cites.
“CrowdMOS: An approach for crowdsourcing mean opinion score studies,”
Flávio Ribeiro, Dinei Florêncio, Cha Zhang, and Michael Seltzer, · 2011
Earlier work this paper cites.
Microphone arrays: signal processing techniques and applications
Michael Brandstein and Darren Ward, · 2013
Cited alongside, same era.
“Deep learning for monaural speech separation,”
Po-Sen Huang, Minje Kim, Mark Hasegawa-Johnson, and Paris Smaragdis, · 2014
Cited alongside, same era.
“Discriminatively trained recurrent neural networks for single-channel speech separation,”
Felix Weninger, John R Hershey, Jonathan Le Roux, and Björn Schuller, · 2014
Cited alongside, same era.
“Optimal distributed minimum-variance beamforming approaches for speech enhancement in wireless acoustic sensor networks,”
Shmulik Markovich-Golan, Alexander Bertrand, Marc Moonen, and Sharon Gannot, · 2015
Cited alongside, same era.
“Freesound,”
2015
Cited alongside, same era.
“100 nonspeech sounds,”
Guoning Hu, · 2015
Cited alongside, same era.
“Speech enhancement in multiple-noise conditions using deep neural networks,”
Anurag Kumar and Dinei Florêncio, · 2016
Later among the works it cites.
“Glottal model based speech beamforming for ad-hoc microphone arrays,”
Yang Zhang, Dinei Florêncio, and Mark Hasegawa-Johnson, · 2017
Later among the works it cites.
“Speech enhancement using Bayesian Wavenet,”
Kaizhi Qian, Yang Zhang, Shiyu Chang, Xuesong Yang, Dinei Florêncio, and Mark Hasegawa-Johnson, · 2017
Later among the works it cites.
“A Wavenet for speech denoising,”
Dario Rethage, Jordi Pons, and Xavier Serra, · 2017
Later among the works it cites.
“SEGAN: Speech enhancement generative adversarial network,”
Santiago Pascual, Antonio Bonafonte, and Joan Serrà, · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Determined blind source separation unifying independent vector analysis and nonnegative matrix factorization,”
Daichi Kitamura, Nobutaka Ono, Hiroshi Sawada, Hirokazu Kameoka, and Hiroshi Saruwatari, · 2016
Cited alongside, same era.
“Long short-term memory for speaker generalization in supervised speech separation,”
Jitong Chen and Deliang Wang, · 2016
Cited alongside, same era.
“Neural network based spectral mask estimation for acoustic beamforming,”
Jahn Heymann, Lukas Drude, and Reinhold Haeb-Umbach, · 2016
Cited alongside, same era.
“WaveNet: A generative model for raw audio,”
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu, · 2016
Cited alongside, same era.
“English multi-speaker corpus for CSTR voice cloning toolkit,”
Junichi Yamagishi,
Cited in the paper.
“On time-frequency mask estimation for mvdr beamforming with application in robust speech recognition,”
Xiong Xiao, Shengkui Zhao, Douglas L Jones, Eng Siong Chng, and Haizhou Li, · 2017
Later among the works it cites.
“A speech enhancement algorithm by iterating single-and multi-microphone processing and its application to robust asr,”
Xueliang Zhang, Zhong-Qiu Wang, and DeLiang Wang, · 2017
Later among the works it cites.
“Dnn-based speech mask estimation for eigenvector beamforming,”
Lukas Pfeifenberger, Matthias Zöhrer, and Franz Pernkopf, · 2017
Later among the works it cites.
“FreeSFX,”
2017
Later among the works it cites.