Fetching the paper…
Reading the bibliography…
Sound event detection (SED) and localization refer to recognizing sound events and estimating their spatial and temporal locations.
C. Knapp and G. Carter, “The generalized correlation method for estimation of time delay,”
1976
Earlier work this paper cites.
M. A. Gerzon, “Ambisonics in multichannel broadcasting and video,”
1985
Earlier work this paper cites.
J. Scheuing and B. Yang, “Correlation-based TDOA-estimation for multiple sources in reverberant environments,” in
2008
Earlier work this paper cites.
M. Brandstein and D. Ward,
2013
Earlier work this paper cites.
E. Çakır, T. Heittola, H. Huttunen, and T. Virtanen, “Polyphonic sound event detection using multi label deep neural networks,” in
2015
Earlier work this paper cites.
K. J. Piczak, “Environmental sound classification with convolutional neural networks,” in
2015
Earlier work this paper cites.
H. Zhang, I. McLoughlin, and Y. Song, “Robust sound event recognition using convolutional neural networks,” in
2015
Earlier work this paper cites.
X. Xiao, S. Zhao, X. Zhong, D. L. Jones, E. S. Chng, and H. Li, “A learning-based approach to direction of arrival estimation in noisy and reverberant environments,” in
2015
Earlier work this paper cites.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in
2015
Earlier work this paper cites.
K. Choi, G. Fazekas, and M. Sandler, “Automatic tagging using deep convolutional neural networks,” in
2016
Earlier work this paper cites.
G. Parascandolo, H. Huttunen, and T. Virtanen, “Recurrent neural networks for polyphonic sound event detection in real life recordings,” in
2016
Cited alongside, same era.
F. Vesperini, P. Vecchiotti, E. Principi, S. Squartini, and F. Piazza, “A neural network based algorithm for speaker localization in a multi-room environment,” in
2016
Cited alongside, same era.
A. Mesaros, T. Heittola, and T. Virtanen, “Metrics for polyphonic sound event detection,”
2016
Cited alongside, same era.
Y. Xu, Q. Kong, Q. Huang, W. Wang, and M. D. Plumbley, “Convolutional gated recurrent neural network incorporating spatial features for audio tagging,” in
2017
Cited alongside, same era.
J. Salamon and J. P. Bello, “Deep convolutional neural networks and data augmentation for environmental sound classification,”
2017
Cited alongside, same era.
T. Virtanen, M. D. Plumbley, and D. Ellis,
2018
Later among the works it cites.
A. Mesaros, T. Heittola, E. Benetos, P. Foster, M. Lagrange, T. Virtanen, and M. D. Plumbley, “Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,”
2018
Later among the works it cites.
E. Fonseca, M. Plakal, F. Font, D. P. W. Ellis, X. Favory, J. Pons, and X. Serra, “General-purpose tagging of Freesound audio with AudioSet labels: task description, dataset, and baseline,” in
2018
Later among the works it cites.
W. He, P. Motlicek, and J. Odobez, “Deep neural networks for multiple speaker detection and localization,” in
2018
Later among the works it cites.
E. L. Ferguson, S. B. Williams, and C. T. Jin, “Sound source localization in a multipath environment using convolutional neural networks,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Çakır, G. Parascandolo, T. Heittola, H. Huttunen, and T. Virtanen, “Convolutional recurrent neural networks for polyphonic sound event detection,”
2017
Cited alongside, same era.
N. Ma, T. May, and G. J. Brown, “Exploiting deep neural networks and head movements for robust binaural localization of multiple sources in reverberant environments,”
2017
Cited alongside, same era.
S. Chakrabarty and E. A. P. Habets, “Broadband DOA estimation using convolutional neural networks trained with noise signals,” in
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Y. Sun, J. Chen, C. Yuen, and S. Rahardja, “Indoor sound source localization with probabilistic neural network,”
2018
Later among the works it cites.
S. Adavanne, A. Politis, and T. Virtanen, “Direction of arrival estimation for multiple sound sources using convolutional recurrent neural network,” in
2018
Later among the works it cites.
2019
Closest in time.
S. Adavanne, A. Politis, J. Nikunen, and T. Virtanen, “Sound event localization and detection of overlapping sources using convolutional recurrent neural networks,”
2019
Closest in time.