Fetching the paper…
Reading the bibliography…
Personalized speech enhancement (PSE) is a real-time SE approach utilizing a speaker embedding of a target person to remove background noise, reverberation, and interfering voices.
J. Kearns, “LibriVox: Free Public Domain Audiobooks,” Reference Reviews , 2014
2014
Earlier work this paper cites.
V. Christophe, Y. Junichi, and M. Kirsten, “CSTR VCTK Corpus: English Multi-speaker Corpus for CSTR Voice Cloning Toolkit,” The Centre for Speech Technology Research (CSTR) , 2016
2016
Earlier work this paper cites.
K. Žmolíková, M. Delcroix, K. Kinoshita, T. Higuchi, A. Ogawa, and T. Nakatani, “Speaker-Aware Neural Network Based Beamformer for Speaker Extraction in Speech Mixtures,” in Proc. INTERSPEECH , 2017, pp. 2655–2659
2017
Earlier work this paper cites.
J. F. Gemmeke, D. P. W. Ellis, D. Freedman, A. Jansen, W. Lawrence, R. C. Moore, M. Plakal, and M. Ritter, “Audio Set: An Ontology and Human-labeled Dataset for Audio Events,” in Proc. ICASSP , 2017, pp. 776–780
2017
Earlier work this paper cites.
E. Fonseca, J. Pons, X. Favory, F. Font, D. Bogdanov, A. Ferraro, S. Oramas, A. Porter, and X. Serra, “Freesound Datasets: A Platform for the Creation of Open Audio Datasets,” in Proc. ISMIR , 2017, pp. 486–493
2017
Earlier work this paper cites.
J. Wang, J. Chen, D. Su, L. Chen, M. Yu, Y. Qian, and D. Yu, “Deep Extractor Network for Target Speaker Recovery from Single Channel Speech Mixtures,” in Proc. INTERSPEECH , 2018, pp. 307–311
2018
Earlier work this paper cites.
R. Scheibler, E. Bezzam, and I. Dokmanić, “Pyroomacoustics: A Python Package for Audio Room Simulation and Array Processing Algorithms,” in Proc. ICASSP . IEEE, 2018, pp. 351–355
2018
Earlier work this paper cites.
Q. Wang, H. Muckenhirn, K. Wilson, P. Sridhar, Z. Wu, J. R. Hershey, R. A. Saurous, R. J. Weiss, Y. Jia, and I. L. Moreno, “VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking,” in Proc. INTERSPEECH , 2019, pp. 2728–2732
2019
Earlier work this paper cites.
Q. Wang, I. L. Moreno, M. Saglam, K. Wilson, A. Chiao, R. Liu, Y. He, W. Li, J. Pelecanos, M. Nika, and A. Gruenstein, “VoiceFilter-Lite: Streaming Targeted Voice Separation for On-Device Speech Recognition,” in Proc. INTERSPEECH , 2020, pp. 2677–2681
2020
Earlier work this paper cites.
M. Delcroix, T. Ochiai, K. Zmolikova, K. Kinoshita, N. Tawara, T. Nakatani, and S. Araki, “Improving Speaker Discrimination of Target Speech Extraction with Time-domain SpeakerBeam,” in Proc. ICASSP , 2020, pp. 691–695
2020
Earlier work this paper cites.
J.-M. Valin, U. Isik, N. Phansalkar, R. Giri, K. Helwani, and A. Krishnaswamy, “A Perceptually-Motivated Approach for Low-Complexity, Real-Time Enhancement of Fullband Speech,” in Proc. INTERSPEECH , 2020, pp. 2482–2486
2020
Cited alongside, same era.
2020
Cited alongside, same era.
R. Giri, S. Venkataramani, J.-M. Valin, U. Isik, and A. Krishnaswamy, “Personalized PercepNet: Real-Time, Low-Complexity Target Voice Separation and Enhancement,” in Proc. INTERSPEECH , 2021, pp. 1124–1128
2021
Cited alongside, same era.
K. Sridhar, R. Cutler, A. Saabas, T. Parnamaa, M. Loide, H. Gamper, S. Braun, R. Aichner, and S. Srinivasan, “ICASSP 2021 Acoustic Echo Cancellation Challenge: Datasets, Testing Framework, and Results,” in Proc. ICASSP , 2021, pp. 151–155
2021
Cited alongside, same era.
M. Thakker, S. E. Eskimez, T. Yoshioka, and H. Wang, “Fast Real-time Personalized Speech Enhancement: End-to-End Enhancement Network (E3Net) and Knowledge Distillation,” in Proc. INTERSPEECH , 2022, pp. 991–995
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Cutler, A. Saabas, T. Parnamaa, M. Loide, S. Sootla, M. Purin, H. Gamper, S. Braun, K. Sorensen, R. Aichner, and S. Srinivasan, “INTERSPEECH 2021 Acoustic Echo Cancellation Challenge,” in Proc. INTERSPEECH , 2021, pp. 4748–4752
2021
Cited alongside, same era.
J. Casebeer, N. J. Bryan, and P. Smaragdis, “Auto-DSP: Learning to Optimize Acoustic Echo Cancellers,” in Proc. WASPAA , 2021, pp. 291–295
2021
Cited alongside, same era.
2021
Cited alongside, same era.
T. Zhou, Y. Zhao, and J. Wu, “ResNeXt and Res2Net Structures for Speaker Verification,” in Proc. SLT , 2021, pp. 301–307
2021
Cited alongside, same era.
S. E. Eskimez, T. Yoshioka, H. Wang, X. Wang, Z. Chen, and X. Huang, “Personalized speech enhancement: New models and comprehensive evaluation,” in Proc. ICASSP , 2022, pp. 356–360
2022
Cited alongside, same era.
R. Cutler, A. Saabas, T. Pärnamaa, M. Purin, H. Gamper, S. Braun, K. Sørensen, and R. Aichner, “ICASSP 2022 Acoustic Echo Cancellation Challenge,” in Proc. ICASSP , 2022, pp. 9107–9111
2022
Closest in time.
S. Braun and M. L. Valero, “Task splitting for dnn-based acoustic echo and noise removal,” in Proc. IWAENC , 2022, pp. 1–5
2022
Closest in time.
H. Dubey, V. Gopal, R. Cutler, A. Aazami, S. Matusevych, S. Braun, S. E. Eskimez, M. Thakker, T. Yoshioka, H. Gamper et al. , “ICASSP 2022 Deep Noise Suppression Challenge,” in Proc. ICASSP , 2022, pp. 9271–9275
2022
Closest in time.
M. Purin, S. Sootla, M. Sponza, A. Saabas, and R. Cutler, “AECMOS: A Speech Quality Assessment Metric for Echo Impairment,” in Proc. ICASSP , 2022, pp. 901–905
2022
Closest in time.