Fetching the paper…
Reading the bibliography…
This paper presents the objective, dataset, baseline, and metrics of Task 3 of the DCASE2025 Challenge on sound event localization and detection (SELD).
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proc. of IEEE CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
M. Defferrard, K. Benzi, P. Vandergheynst, and X. Bresson, “Fma: A dataset for music analysis,” in Proc. of ISMIR , 2017
2017
Earlier work this paper cites.
S. Adavanne, A. Politis, J. Nikunen, and T. Virtanen, “Sound event localization and detection of overlapping sources using convolutional recurrent neural networks,” IEEE Journal of Selected Topics in Signal Processing , vol. 13, no. 1, pp. 34–48, 2019
2019
Earlier work this paper cites.
S. Adavanne, A. Politis, and T. Virtanen, “A multi-room reverberant dataset for sound event localization and detection,” in Proc. of DCASE Workshop , 2019
2019
Earlier work this paper cites.
L. Mazzon, Y. Koizumi, M. Yasuda, and N. Harada, “First order Ambisonics domain spatial augumentation for DNN-based direction of arrival estimation,” in Proc. of DCASE Workshop , 2019
2019
Earlier work this paper cites.
S. Kapka and M. Lewandowski, “Sound source detection, localization and classification using consecutive ensemble of CRNN models,” in Proc. of DCASE Workshop , 2019
2019
Earlier work this paper cites.
Y. Cao, Q. Kong, T. Iqbal, F. An, W. Wang, and M. D. Plumbley, “Polyphonic sound event detection and localization using a two-stage strategy,” in Proc. of DCASE Workshop , 2019
2019
Earlier work this paper cites.
A. Politis, S. Adavanne, and T. Virtanen, “A dataset of reverberant spatial sound scenes with moving sources for sound event localization and detection,” in Proc. of DCASE Workshop , 2020
2020
Earlier work this paper cites.
T. N. Tho Nguyen, D. L. Jones, and W.-S. Gan, “A sequence matching network for polyphonic sound event localization and detection,” in Proc. of IEEE ICASSP , 2020, pp. 71–75
2020
Earlier work this paper cites.
A. Politis, S. Adavanne, D. Krause, A. Deleforge, P. Srivastava, and T. Virtanen, “A dataset of dynamic reverberant sound scenes with directional interferers for sound event localization and detection,” in Proc. of DCASE Workshop , 2021
2021
Earlier work this paper cites.
A. Politis, A. Mesaros, S. Adavanne, T. Heittola, and T. Virtanen, “Overview and evaluation of sound event localization and detection in dcase 2019,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 29, pp. 684–698, 2021
2021
Earlier work this paper cites.
Y. Cao, T. Iqbal, Q. Kong, F. An, W. Wang, and M. D. Plumbley, “An improved event-independent network for polyphonic sound event localization and detection,” in Proc. of IEEE ICASSP , 2021
2021
Cited alongside, same era.
K. Shimada, Y. Koyama, N. Takahashi, S. Takahashi, and Y. Mitsufuji, “Accdoa: Activity-coupled cartesian direction of arrival representation for sound event localization and detection,” in Proc. of IEEE ICASSP , 2021, pp. 915–919
2021
Cited alongside, same era.
P. Sudarsanam, A. Politis, and K. Drossos, “Assessment of self-attention on learned features for sound event localization and detection,” in Proc. of DCASE Workshop , 2021
2021
Cited alongside, same era.
E. Fonseca, X. Favory, J. Pons, F. Font, and X. Serra, “Fsd50k: an open dataset of human-labeled sound events,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 30, pp. 829–852, 2021
2021
Cited alongside, same era.
Q. Wang, Y. Jiang, S. Cheng, M. Hu, Z. Nian, P. Hu, Z. Liu, Y. Dong, M. Cai, J. Du, and C.-H. Lee, “The nerc-slip system for sound event localization and detection of dcase2023 challenge,” DCASE2023 Challenge, Tech. Rep., June 2023
2023
Later among the works it cites.
D. A. Krause, G. García-Barrios, A. Politis, and A. Mesaros, “Binaural sound source distance estimation and localization for a moving listener,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 32, pp. 996–1011, 2023
2023
Later among the works it cites.
S. S. Kushwaha, I. R. Roman, M. Fuentes, and J. P. Bello, “Sound source distance estimation in diverse and dynamic acoustic conditions,” in Proc. of IEEE WASPAA , 2023, pp. 1–5
2023
Later among the works it cites.
B. S. Liang, A. S. Liang, I. Roman, T. Weiss, B. Duinkharjav, J. P. Bello, and Q. Sun, “Reconstructing room scales with a single sound for augmented reality displays,” Journal of Information Display , vol. 24, no. 1, pp. 1–12, 2023
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. N. T. Nguyen, D. L. Jones, K. N. Watcharasupat, H. Phan, and W.-S. Gan, “SALSA-Lite: A fast and effective feature for polyphonic sound event localization and detection with microphone arrays,” in Proc. of IEEE ICASSP , 2022, pp. 716–720
2022
Cited alongside, same era.
Y. Koyama, K. Shigemi, M. Takahashi, K. Shimada, N. Takahashi, E. Tsunoo, S. Takahashi, and Y. Mitsufuji, “Spatial data augmentation with simulated room impulse responses for sound event localization and detection,” in Proc. of IEEE ICASSP , 2022, pp. 8872–8876
2022
Cited alongside, same era.
A. Politis, K. Shimada, P. Sudarsanam, S. Adavanne, D. Krause, Y. Koyama, N. Takahashi, S. Takahashi, Y. Mitsufuji, and T. Virtanen, “Starss22: A dataset of spatial recordings of real scenes with spatiotemporal annotations of sound events,” in Proc. of DCASE Workshop , 2022
2022
Cited alongside, same era.
Q. Wang, L. Chai, H. Wu, Z. Nian, S. Niu, S. Zheng, Y. Wang, L. Sun, Y. Fang, J. Pan, J. Du, and C.-H. Lee, “The nerc-slip system for sound event localization and detection of dcase2022 challenge,” DCASE2022 Challenge, Tech. Rep., June 2022
2022
Cited alongside, same era.
J. Hu, Y. Cao, M. Wu, Q. Kong, F. Yang, M. D. Plumbley, and J. Yang, “Sound event localization and detection for real spatial sound scenes: Event-independent network and data augmentation chains,” DCASE2022 Challenge, Tech. Rep., June 2022
2022
Cited alongside, same era.
K. Shimada, Y. Koyama, S. Takahashi, N. Takahashi, E. Tsunoo, and Y. Mitsufuji, “Multi-accdoa: Localizing and detecting overlapping sounds from the same class with auxiliary duplicating permutation invariant training,” in Proc. of IEEE ICASSP , 2022, pp. 316–320
2022
Cited alongside, same era.
Q. Wang, J. Du, H.-X. Wu, J. Pan, F. Ma, and C.-H. Lee, “A four-stage data augmentation approach to resnet-conformer based acoustic modeling for sound event localization and detection,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 31, pp. 1251–1264, 2023
2023
Cited alongside, same era.
K. Shimada, A. Politis, P. Sudarsanam, D. A. Krause, K. Uchida, S. Adavanne, A. Hakala, Y. Koyama, N. Takahashi, S. Takahashi, T. Virtanen, and Y. Mitsufuji, “Starss23: An audio-visual dataset of spatial recordings of real scenes with spatiotemporal annotations of sound events,” Advances in neural information processing systems , vol. 36, pp. 72 931–72 957, 2023
2023
Cited alongside, same era.
Later among the works it cites.
J. Wilkins, M. Fuentes, L. Bondi, S. Ghaffarzadegan, A. Abavisani, and J. P. Bello, “Two vs. four-channel sound event localization and detection,” in Proc. of DCASE Workshop , 2023
2023
Later among the works it cites.
D. D.-G. Aparicio, A. Politis, P. A. Sudarsanam, K. Shimada, D. Krause, K. Uchida, Y. Koyama, N. Takahashi, S. Takahashi, T. Shibuya, Y. Mitsufuji, and T. Virtanen, “Baseline models and evaluation of sound event localization and detection with distance estimation in dcase 2024 challenge,” in Proc. of DCASE Workshop , 2024, pp. 41–45
2024
Later among the works it cites.
D. Berghi, P. Wu, J. Zhao, W. Wang, and P. J. B. Jackson, “Fusion of audio and visual embeddings for sound event localization and detection,” in Proc. of IEEE ICASSP , 2024, pp. 8816–8820
2024
Later among the works it cites.
Q. Wang, Y. Dong, H. Hong, R. Wei, M. Hu, S. Cheng, Y. Jiang, M. Cai, X. Fang, and J. Du, “The nerc-slip system for sound event localization and detection with source distance estimation of dcase 2024 challenge,” DCASE2024 Challenge, Tech. Rep., June 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
D. A. Krause, A. Politis, and A. Mesaros, “Sound event detection and localization with distance estimation,” in Proc. of EUSIPCO , 2024, pp. 286–290
2024
Later among the works it cites.
I. R. Roman, C. Ick, S. Ding, A. S. Roman, B. McFee, and J. P. Bello, “Spatial scaper: a library to simulate and augment soundscapes for sound event localization and detection in realistic rooms,” in Proc. of IEEE ICASSP , 2024, pp. 1221–1225
2024
Later among the works it cites.