Fetching the paper…
Reading the bibliography…
We present aTENNuate, a simple deep state-space autoencoder configured for efficient online raw speech enhancement in an end-to-end fashion.
J. Chen, J. Benesty, Y. Huang, and S. Doclo, “New insights into the noise reduction wiener filter,” IEEE Transactions on audio, speech, and language processing , vol. 14, no. 4, pp. 1218–1234, 2006
2006
Earlier work this paper cites.
S. V. Vaseghi, Advanced digital signal processing and noise reduction . John Wiley & Sons, 2008
2008
Earlier work this paper cites.
R. Girshick, “Fast r-cnn,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 1440–1448
2015
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
Z. Zhao, H. Liu, and T. Fingscheidt, “Convolutional neural networks to enhance coded speech,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 27, no. 4, pp. 663–678, 2018
2018
Earlier work this paper cites.
H.-S. Choi, J.-H. Kim, J. Huh, A. Kim, J.-W. Ha, and K. Lee, “Phase-aware speech enhancement with deep complex u-net,” in International Conference on Learning Representations , 2018
2018
Earlier work this paper cites.
D. Rethage, J. Pons, and X. Serra, “A wavenet for speech denoising,” in 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2018, pp. 5069–5073
2018
Earlier work this paper cites.
J.-M. Valin, “A hybrid dsp/deep learning approach to real-time full-band speech enhancement,” in 2018 IEEE 20th international workshop on multimedia signal processing (MMSP) . IEEE, 2018, pp. 1–5
2018
Earlier work this paper cites.
R. Giri, U. Isik, and A. Krishnaswamy, “Attention wave-u-net for speech enhancement,” in 2019 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) . IEEE, 2019, pp. 249–253
2019
Earlier work this paper cites.
V. Srinivasarao and U. Ghanekar, “Speech enhancement-an enhanced principal component analysis (epca) filter approach,” Computers & Electrical Engineering , vol. 85, p. 106657, 2020
2020
Earlier work this paper cites.
C. Yu, R. E. Zezario, S.-S. Wang, J. Sherman, Y.-Y. Hsieh, X. Lu, H.-M. Wang, and Y. Tsao, “Speech enhancement based on denoising autoencoder with multi-branched encoders,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 28, pp. 2756–2769, 2020
2020
Earlier work this paper cites.
A. Azarang and N. Kehtarnavaz, “A review of multi-objective deep learning speech denoising methods,” Speech Communication , vol. 122, pp. 1–10, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2022
Later among the works it cites.
C. Zheng, H. Zhang, W. Liu, X. Luo, A. Li, X. Li, and B. C. Moore, “Sixty years of frequency-domain monaural speech enhancement: From traditional to deep learning methods,” Trends in Hearing , vol. 27, p. 23312165231209913, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
S. Zhao, B. Ma, K. N. Watcharasupat, and W.-S. Gan, “Frcrn: Boosting feature representation using frequency recurrence for monaural speech enhancement,” in ICASSP 2022-2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2022, pp. 9281–9285
2022
Cited alongside, same era.
Z. Kong, W. Ping, A. Dantrey, and B. Catanzaro, “Speech denoising in the waveform domain with self-attention,” in ICASSP 2022-2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2022, pp. 7867–7871
2022
Cited alongside, same era.
K. Goel, A. Gu, C. Donahue, and C. Ré, “It’s raw! audio generation with state-space models,” in International Conference on Machine Learning . PMLR, 2022, pp. 7616–7633
2022
Cited alongside, same era.
A. Gu, K. Goel, A. Gupta, and C. Ré, “On the parameterization and initialization of diagonal state space models,” Advances in Neural Information Processing Systems , vol. 35, pp. 35 971–35 983, 2022
2022
Cited alongside, same era.
Y. R. Pei, S. Brüers, S. Crouzet, D. McLelland, and O. Coenen, “A Lightweight Spatiotemporal Network for Online Eye Tracking with Event Camera,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops , 2024
2024
Closest in time.
2024
Closest in time.
Z. Wang, X. Zhu, Z. Zhang, Y. Lv, N. Jiang, G. Zhao, and L. Xie, “Selm: Speech enhancement using discrete tokens and language models,” in ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2024, pp. 11 561–11 565
2024
Closest in time.
Y. Du, X. Liu, and Y. Chua, “Spiking structured state space model for monaural speech enhancement,” in ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2024, pp. 766–770
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.