Fetching the paper…
Reading the bibliography…
This article is a survey on deep learning methods for single and multiple sound source localization.
Multi-speaker DoA estimation using deep convolutional networks trained with noise signals
S. Chakrabarty and E. A. P. Habets · 1941
Earlier work this paper cites.
The generalized correlation method for estimation of time delay
C. Knapp and G. Carter · 1976
Earlier work this paper cites.
Image method for efficiently simulating small-room acoustics
J. B. Allen and D. A. Berkley · 1979
Earlier work this paper cites.
Multiple emitter location and signal parameter estimation
R. Schmidt · 1986
Earlier work this paper cites.
Array signal processing with interconnected neuron-like elements
R. Rastogi, P. Gupta, and R. Kumaresan · 1987
Earlier work this paper cites.
Neural networks for narrowband and wideband direction finding
D. Goryn and M. Kaveh · 1988
Earlier work this paper cites.
Bearing estimation using neural networks
S. Jha, R. Chapman, and T. Durrani · 1988
Earlier work this paper cites.
Beamforming: A versatile approach to spatial filtering
B. D. Van Veen and K. M. Buckley · 1988
Earlier work this paper cites.
Bearing estimation using neural optimisation methods
S. Jha and T. Durrani · 1989
Earlier work this paper cites.
ESPRIT: Estimation of signal parameters via rotational invariance techniques
R. Roy and T. Kailath · 1989
Earlier work this paper cites.
Phoneme recognition using time-delay neural networks
A. Waibel, T. Hanazawa, G. Hinton, K. Shikano, and K. Lang · 1989
Earlier work this paper cites.
Direction of arrival estimation using artificial neural networks
S. Jha and T. Durrani · 1991
Earlier work this paper cites.
BREF, a large vocabulary spoken corpus for French
L. Lamel, J.-L. Gauvain, and M. Eskenazi · 1991
Earlier work this paper cites.
General metatheory of auditory localisation
M. A. Gerzon · 1992
Earlier work this paper cites.
The ML bearing estimation by using neural networks
L. Falong, J. Hongbing, and Z. Xiaopeng · 1993
Earlier work this paper cites.
Finding the direction of a sound source using a vector sound-intensity probe
R. Hickling, W. Wei, and R. Raspet · 1993
Earlier work this paper cites.
The reactive intensity of general time-harmonic structure-borne sound fields
W. Maysenhölder · 1993
Earlier work this paper cites.
Acoustic vector-sensor array processing
A. Nehorai and E. Paldi · 1994
Earlier work this paper cites.
Complex-valued neural network for direction of arrival estimation
W.-H. Yang, K.-K. Chan, and P.-R. Chang · 1994
Earlier work this paper cites.
Direction finding in phased arrays with a neural network beamformer
H. Southall, J. Simmers, and T. O’Donnell · 1995
Earlier work this paper cites.
An artificial neural network for sound localization using binaural cues
M. S. Datum, F. Palmieri, and A. Moiseff · 1996
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
A neural network-based smart antenna for multiple source tracking
A. El Zooghby, C. Christodoulou, and M. Georgiopoulos · 2000
Earlier work this paper cites.
The use of computer modeling in room acoustics
J. H. Rindel · 2000
Earlier work this paper cites.
Microphone Arrays: Signal Processing Techniques and Applications
M. Brandstein and D. Ward · 2001
Earlier work this paper cites.
Robust localization in reverberant rooms
J. H. DiBiase, H. F. Silverman, and M. S. Brandstein · 2001
Earlier work this paper cites.
DDSP: Differentiable digital signal processing, 2020
J. Engel, L. Hantrakul, C. Gu, and A. Roberts · 2001
Earlier work this paper cites.
Sound source localization using sound intensity measured by a three dimensional PU-probe
R. Raangs and E. Druyvesteyn · 2002
Earlier work this paper cites.
On the approximate W-disjoint orthogonality of speech
S. Rickard · 2002
Earlier work this paper cites.
Computational modelling and simulation of acoustic spaces
P. Svensson and U. R. Kristiansen · 2002
Earlier work this paper cites.
An overview of microflown technologies
H.-E. de Bree · 2003
Earlier work this paper cites.
Direction of arrival estimation for multiple source signals using independent component analysis
H. Sawada, R. Mukai, and S. Makino · 2003
Earlier work this paper cites.
Relative transfer function identification using speech signals
I. Cohen · 2004
Earlier work this paper cites.
AV16.3: an audio-visual corpus for speaker localization and tracking
G. Lathoud, J.-M. Odobez, and D. Gatica-Perez · 2004
Earlier work this paper cites.
A large set of audio features for sound description (similarity and classification) in the CUIDADO project
G. Peeters · 2004
Earlier work this paper cites.
A Matlab simulation of shoebox room acoustics for use in research and teaching
D. Campbell, K. Palomaki, and G. Brown · 2005
Earlier work this paper cites.
Time difference of arrival estimation of speech source in a noisy and reverberant environment
T. G. Dvorkind and S. Gannot · 2005
Earlier work this paper cites.
Real time acoustic rendering of complex environments including diffraction and curved surfaces
K. Bouatouch, O. Deille, J. Maillard, J. Martin, and N. Noé · 2006
Earlier work this paper cites.
Stable signal recovery from incomplete and inaccurate measurements
E. J. Candes, J. K. Romberg, and T. Tao · 2006
Earlier work this paper cites.
Room impulse response generator
E. A. P. Habets · 2006
Earlier work this paper cites.
Simple unsupervised multi-object tracking, 2020
S. Karthik, A. Prabhu, and V. Gandhi · 2006
Earlier work this paper cites.
Analysis, synthesis, and perception of spatial sound: binaural localization modeling and multichannel loudspeaker reproduction
J. Merimaa · 2006
Earlier work this paper cites.
An overview of automatic speaker diarization systems
S. E. Tranter and D. A. Reynolds · 2006
Earlier work this paper cites.
Broadband MUSIC: Opportunities and challenges for multiple source localization
J. P. Dmochowski, J. Benesty, and S. Affes · 2007
Earlier work this paper cites.
Springer Handbook of Acoustics
T. D. Rossing · 2007
Earlier work this paper cites.
The CLEAR 2007 evaluation
R. Stiefelhagen, K. Bernardin, R. Bowers, R. T. Rose, M. Michel, and J. Garofolo · 2007
Earlier work this paper cites.
Acoustic eyes: a novel sound source localization and monitoring technique with 3D sound probes
T. Basten, H. de Bree, and S. Sadasivan · 2008
Earlier work this paper cites.
Microphone Array Signal Processing
J. Benesty, J. Chen, and Y. Huang · 2008
Earlier work this paper cites.
Binaural tracking of multiple moving sources
N. Roman and D. Wang · 2008
Earlier work this paper cites.
Roomsimove
E. Vincent and D. R. Campbell · 2008
Earlier work this paper cites.
A robust method to count and locate audio sources in a multichannel underdetermined mixture
S. Arberet, R. Gribonval, and F. Bimbot · 2009
Earlier work this paper cites.
Model-based expectation-maximization source separation and localization
M. I. Mandel, R. J. Weiss, and D. P. Ellis · 2009
Earlier work this paper cites.
Direction estimation based on sound intensity vectors
S. Tervo · 2009
Earlier work this paper cites.
WOZ acoustic data collection for interactive TV
A. Brutti, L. Cristoforetti, W. Kellermann, L. Marquardt, and M. Omologo · 2010
Earlier work this paper cites.
Under-determined reverberant audio source separation using a full-rank spatial covariance model
N. Q. Duong, E. Vincent, and R. Gribonval · 2010
Earlier work this paper cites.
3D source localization in the spherical harmonic domain using a pseudointensity vector
D. P. Jarrett, E. A. Habets, and P. A. Naylor · 2010
Earlier work this paper cites.
The cone of silence: speech separation by localization, 2020
T. Jenrungrot, V. Jayaram, S. Seitz, and I. Kemelmacher-Shlizerman · 2010
Earlier work this paper cites.
Diffuse reverberation model for efficient image-source simulation of room impulse responses
E. A. Lehmann and A. M. Johansson · 2010
Earlier work this paper cites.
Nested arrays: A novel approach to array processing with enhanced degrees of freedom
P. Pal and P. P. Vaidyanathan · 2010
Earlier work this paper cites.
Rays or waves? understanding the strengths and weaknesses of computational room acoustics modeling techniques
S. Siltanen, T. Lokki, and L. Savioja · 2010
Earlier work this paper cites.
Sparse sensing with co-prime samplers and arrays
P. P. Vaidyanathan and P. Pal · 2010
Earlier work this paper cites.
Room acoustics simulation for multichannel microphone arrays
A. Wabnitz, N. Epain, C. Jin, and A. Van Schaik · 2010
Earlier work this paper cites.
Direction of arrival estimation of humans with a small sensor array using an artificial neural network
Y. Kim and H. Ling · 2011
Earlier work this paper cites.
Comparison of subspace-based and steered beamformer-based reflection localization methods
E. Mabande, H. Sun, K. Kowalczyk, and W. Kellermann · 2011
Earlier work this paper cites.
A probabilistic model for robust localization based on a binaural auditory front-end
T. May, S. Van De Par, and A. Kohlrausch · 2011
Earlier work this paper cites.
Speaker diarization: A review of recent research
X. Anguera, S. Bozonnet, N. Evans, C. Fredouille, G. Friedland, and O. Vinyals · 2012
Earlier work this paper cites.
Deep learning of representations for unsupervised and transfer learning
Y. Bengio · 2012
Earlier work this paper cites.
Multi-source TDOA estimation in reverberant audio using angular spectra and clustering
C. Blandin, A. Ozerov, and E. Vincent · 2012
Earlier work this paper cites.
Narrowband source localization in an unknown reverberant environment using wavefield sparse decomposition
G. Chardon and L. Daudet · 2012
Earlier work this paper cites.
2D sound-source localization on the binaural manifold
A. Deleforge and R. Horaud · 2012
Earlier work this paper cites.
Rigid sphere room impulse response simulation: Algorithm and applications
D. Jarrett, E. Habets, M. Thomas, and P. Naylor · 2012
Earlier work this paper cites.
An efficient maximum likelihood method for direction-of-arrival estimation via sparse Bayesian learning
Z.-M. Liu, Z.-T. Huang, and Y.-Y. Zhou · 2012
Earlier work this paper cites.
Model-based deep learning, 2020
N. Shlezinger, J. Whang, Y. C. Eldar, and A. G. Dimakis · 2012
Earlier work this paper cites.
Transtrack: Multiple object tracking with transformer, 2020
P. Sun, J. Cao, Y. Jiang, R. Zhang, E. Xie, Z. Yuan, C. Wang, and P. Luo · 2012
Earlier work this paper cites.
Binaural localization of multiple sources in reverberant and noisy environments
J. Woodruff and D. Wang · 2012
Earlier work this paper cites.
Representation learning: A review and new perspectives
Y. Bengio, A. Courville, and P. Vincent · 2013
Earlier work this paper cites.
Variational EM for binaural sound-source separation and localization
A. Deleforge, F. Forbes, and R. Horaud · 2013
Earlier work this paper cites.
An invitation to compressive sensing
S. Foucart and H. Rauhut · 2013
Earlier work this paper cites.
Fundamentals of General Linear Acoustics
F. Jacobsen and P. M. Juhl · 2013
Earlier work this paper cites.
The cosparse analysis model and algorithms
S. Nam, M. E. Davies, M. Elad, and R. Gribonval · 2013
Earlier work this paper cites.
Direction of arrival estimation for spherical microphone arrays by combination of independent component analysis and sparse recovery
T. Noohi, N. Epain, and C. T. Jin · 2013
Earlier work this paper cites.
Speaker tracking using recursive EM algorithms
O. Schwartz and S. Gannot · 2013
Earlier work this paper cites.
An approach for sound source localization by complex-valued neural network
H. Tsuzuki, M. Kugler, S. Kuroyanagi, and A. Iwata · 2013
Earlier work this paper cites.
High-accuracy TDOA-based localization without time synchronization
B. Xu, G. Sun, R. Yu, and Z. Yang · 2013
Earlier work this paper cites.
A learning-based approach to robust binaural sound localization
K. Youssef, S. Argentieri, and J. Zarader · 2013
Earlier work this paper cites.
K. Cho, B. van Merrienboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
The DIRHA simulated corpus
L. Cristoforetti, M. Ravanelli, M. Omologo, A. Sosi, and A. Abad · 2014
Earlier work this paper cites.
A Bayesian direction-of-arrival model for an undetermined number of sources using a two-microphone array
J. Escolano, N. Xiang, J. M. Perez-Lorenzo, M. Cobos, and J. J. Lopez · 2014
Earlier work this paper cites.
Multiple source localisation in the spherical harmonic domain
C. Evers, A. H. Moore, and P. A. Naylor · 2014
Earlier work this paper cites.
Single-snapshot DOA estimation by using compressed sensing
S. Fortunati, R. Grasso, F. Gini, M. S. Greco, and K. LePage · 2014
Earlier work this paper cites.
Generative Adversarial Nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Multichannel audio database in various acoustic environments
E. Hadad, F. Heese, P. Vary, and S. Gannot · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification, 2014
Y. Kim · 2014
Earlier work this paper cites.
Auto-encoding variational Bayes
D. P. Kingma and M. Welling · 2014
Earlier work this paper cites.
Hearing behind walls: localizing sources in the room next door with cosparsity
S. Kitić, N. Bertin, and R. Gribonval · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
D. J. Rezende, S. Mohamed, and D. Wierstra · 2014
Earlier work this paper cites.
Compressive beamforming
A. Xenaki, P. Gerstoft, and K. Mosegaard · 2014
Earlier work this paper cites.
Off-grid DOA estimation using array covariance matrix and block-sparse Bayesian learning
Y. Zhang, Z. Ye, X. Xu, and N. Hu · 2014
Earlier work this paper cites.
A survey on sound source localization in robotics: From binaural to array processing methods
S. Argentieri, P. Danes, and P. Souères · 2015
Earlier work this paper cites.
Co-localization of audio sources in images using binaural features and locally-linear regression
A. Deleforge, R. Horaud, Y. Y. Schechner, and L. Girin · 2015
Earlier work this paper cites.
Tree-based recursive expectation-maximization algorithm for localization of acoustic sources
Y. Dorfan and S. Gannot · 2015
Earlier work this paper cites.
The ACE challenge — Corpus description and performance evaluation
J. Eaton, N. D. Gaubitch, A. H. Moore, and P. A. Naylor · 2015
Earlier work this paper cites.
Classification of spatial audio location and content using convolutional neural networks
T. Hirvonen · 2015
Earlier work this paper cites.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Earlier work this paper cites.
Estimation of relative transfer function in the presence of stationary noise based on segmental power spectral density matrix subtraction
X. Li, L. Girin, R. Horaud, and S. Gannot · 2015
Earlier work this paper cites.
Exploiting deep neural networks and head movements for binaural localisation of multiple speakers in reverberant conditions
N. Ma, G. Brown, and T. May · 2015
Earlier work this paper cites.
Performance analysis of the covariance subtraction method for relative transfer function estimation and comparison to the covariance whitening method
S. Markovich-Golan and S. Gannot · 2015
Earlier work this paper cites.
3D localization of multiple sound sources with intensity vector estimates in single source zones
D. Pavlidi, S. Delikaris-Manias, V. Pulkki, and A. Mouchtaris · 2015
Earlier work this paper cites.
On sound source localization of speech signals using deep neural networks
R. Roden, N. Moritz, S. Gerlach, S. Weinzierl, and S. Goetze · 2015
Earlier work this paper cites.
U-Net: convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Cited alongside, same era.
Multiple model high-spatial resolution HRTF measurements
J. Thiemann and S. Van De Par · 2015
Cited alongside, same era.
Multitarget tracking
B.-n. Vo, M. Mallick, Y. Bar-shalom, S. Coraluppi, R. Osborne, R. Mahler, and B.-t. Vo · 2015
Cited alongside, same era.
Grid-free compressive beamforming
A. Xenaki and P. Gerstoft · 2015
Cited alongside, same era.
A learning-based approach to direction of arrival estimation in noisy and reverberant environments
X. Xiao, S. Zhao, X. Zhong, D. L. Jones, E. S. Chng, and H. Li · 2015
Cited alongside, same era.
Enhancing sparsity and resolution via reweighted atomic norm minimization
Z. Yang and L. Xie · 2015
Cited alongside, same era.
Multitask learning of time-frequency CNN for sound source localization
C. Pang, H. Liu, and X. Li · 2019
Later among the works it cites.
Overview and evaluation of sound event localization and detection in DCASE 2019
A. Politis, A. Mesaros, S. Adavanne, T. Heittola, and T. Virtanen · 2019
Later among the works it cites.
Sound event localization and detection using CRNN architecture with Mixup for model generalization
P. Pratik, W. J. Jee, S. Nagisetty, R. Mars, and C. Lim · 2019
Later among the works it cites.
Source localization in reverberant rooms using deep learning and microphone arrays
H. Pujol, E. Bavu, and A. Garcia · 2019
Later among the works it cites.
Deep learning for audio signal processing
H. Purwins, B. Li, T. Virtanen, J. Schlüter, S.-Y. Chang, and T. Sainath · 2019
Later among the works it cites.
Fundamentals of Spherical Array Processing
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural machine translation by jointly learning to align and translate, May 2016
D. Bahdanau, K. Cho, and Y. Bengio · 2016
Cited alongside, same era.
Microphone arrays and sound field decomposition for dynamic binaural recording
B. Bernschütz · 2016
Cited alongside, same era.
The ray space transform: a new framework for wave field processing
L. Bianchi, F. Antonacci, A. Sarti, and S. Tubaro · 2016
Cited alongside, same era.
Improved MVDR beamforming using single-channel mask prediction networks
H. Erdogan, J. R. Hershey, S. Watanabe, M. Mandel, and J. Le Roux · 2016
Cited alongside, same era.
Multisnapshot sparse Bayesian learning for DOA
P. Gerstoft, C. F. Mecklenbräuker, A. Xenaki, and S. Nannuru · 2016
Cited alongside, same era.
Deep Learning
I. Goodfellow, Y. Bengio, and A. Courville · 2016
Cited alongside, same era.
B. Rafaely · 2019
Later among the works it cites.
Sound events detection and direction of arrival estimation using residual net and recurrent neural networks
R. Ranjan, S. Jayabalan, T. N. T. Nguyen, and W.-S. Lim · 2019
Later among the works it cites.
Improvement of DOA estimation by using quaternion output in sound event localization and detection
Y. Sudo, K. Itoyama, K. Nishida, and K. Nakadai · 2019
Later among the works it cites.
Building and evaluation of a real room impulse response dataset
I. Szöke, M. Skácel, L. Mošner, J. Paliesek, and J. Černocký · 2019
Later among the works it cites.
Regression and classification for direction-of-arrival estimation with convolutional recurrent neural networks
Z. Tang, J. D. Kanu, K. Hogan, and D. Manocha · 2019
Later among the works it cites.
Robust speaker localization guided by deep learning-based time-frequency masking
Z. Wang, X. Zhang, and D. Wang · 2019
Later among the works it cites.
Wham!: Extending speech separation to noisy environments
G. Wichern, J. Antognini, M. Flynn, L. R. Zhu, E. McQuinn, D. Crow, E. Manilow, and J. L. Roux · 2019
Later among the works it cites.
Online multi-object tracking based on feature representation and Bayesian filtering within a deep learning architecture
J. Xiang, G. Zhang, and J. Hou · 2019
Later among the works it cites.
Multi-beam and multi-task learning for joint sound event detection and localization
W. Xue, T. Ying, Z. Chao, and D. Guohong · 2019
Later among the works it cites.
Data augmentation and priori knowledge-based regularization for sound event localization and detection
J. Zhang, W. Ding, and L. He · 2019
Later among the works it cites.
Ambisonics: A Practical 3D Audio Theory for Recording, Studio Production, Sound Reinforcement, and Virtual Reality
F. Zotter and M. Frank · 2019
Later among the works it cites.
Semi-supervised source localization with deep generative modeling
M. J. Bianco, S. Gannot, and P. Gerstoft · 2020
Later among the works it cites.
Event-independent network for polyphonic sound event localization and detection
Y. Cao, T. Iqbal, Q. Kong, Y. Zhong, W. Wang, and M. D. Plumbley · 2020
Later among the works it cites.
Convolutional neural network-based DoA estimation using stereo microphones for drone
J. Choi and J.-H. Chang · 2020
Later among the works it cites.
Deep learning in video multi-object tracking: A survey
G. Ciaparrone, F. Luque Sánchez, S. Tabik, L. Troiano, R. Tagliaferri, and F. Herrera · 2020
Later among the works it cites.
Exploiting spatial invariance for scalable unsupervised object tracking
E. Crawford and J. Pineau · 2020
Later among the works it cites.
Time-domain velocity vector for retracing the multipath propagation
J. Daniel and S. Kitić · 2020
Later among the works it cites.
DeepMUSIC: multiple signal classification via deep learning
A. M. Elbir · 2020
Later among the works it cites.
The LOCATA challenge: Acoustic source localization and tracking
C. Evers, H. W. Löllmann, H. Mellmann, A. Schmidt, H. Barfuss, P. A. Naylor, and W. Kellermann · 2020
Later among the works it cites.
Multi-source DoA estimation through pattern recognition of the modal coherence of a reverberant soundfield
A. Fahim, P. N. Samarasinghe, and T. D. Abhayapala · 2020
Later among the works it cites.
High-resolution speaker counting in reverberant rooms using CRNN with Ambisonics features
P.-A. Grumiaux, S. Kitic, L. Girin, and A. Guerin · 2020
Later among the works it cites.
SELD-TCN: sound event localization & detection via temporal convolutional networks
K. Guirguis, C. Schorn, A. Guntoro, S. Abdulatif, and B. Yang · 2020
Later among the works it cites.
Conformer: convolution-augmented Transformer for speech recognition
A. Gulati, J. Qin, C.-C. Chiu, N. Parmar, Y. Zhang, J. Yu, W. Han, S. Wang, Z. Zhang, Y. Wu, and R. Pang · 2020
Later among the works it cites.
Spectral flux-based convolutional neural network architecture for speech source localization and its real-time implementation
Y. Hao, A. Küçük, A. Ganguly, and I. M. S. Panahi · 2020
Later among the works it cites.
Squeeze-and-excitation networks
J. Hu, L. Shen, S. Albanie, G. Sun, and E. Wu · 2020
Later among the works it cites.
A time-domain unsupervised learning based sound source localization method
Y. Huang, X. Wu, and T. Qu · 2020
Later among the works it cites.
Sound event localization and detection using convolutional recurrent neural networks and gated linear units
T. Komatsu, M. Togami, and T. Takahashi · 2020
Later among the works it cites.
Data-driven multi-microphone speaker localization on manifolds
B. Laufer-Goldshtein, R. Talmon, and S. Gannot · 2020
Later among the works it cites.
Learning multiple sound source 2D localization
G. Le Moing, P. Vinayavekhin, T. Inoue, J. Vongkulbhisal, A. Munawar, R. Tachibana, and D. J. Agravante · 2020
Later among the works it cites.
UnOVOST: Unsupervised offline video object segmentation and tracking
J. Luiten, I. E. Zulfikar, and B. Leibe · 2020
Later among the works it cites.
End-to-end microphone permutation and number invariant multi-channel speech separation
Y. Luo, Z. Chen, N. Mesgarani, and T. Yoshioka · 2020
Later among the works it cites.
Signal-aware broadband DoA estimation using attention mechanisms
W. Mack, U. Bharadwaj, S. Chakrabarty, and E. A. P. Habets · 2020
Later among the works it cites.
Self-supervised neural audio-visual sound source localization via probabilistic spatial modeling
Y. Masuyama, Y. Bando, K. Yatabe, Y. Sasaki, M. Onishi, and Y. Oikawa · 2020
Later among the works it cites.
Sound event localization and detection using squeeze-excitation residual CNNs
J. Naranjo-Alcazar, S. Perez-Castanos, J. Ferrandis, P. Zuccarello, and M. Cobos · 2020
Later among the works it cites.
Sound event localization and detection with various loss functions
S. Park, S. Suh, and Y. Jeong · 2020
Later among the works it cites.
A single stage fully convolutional neural network for sound source localization and detection
S. J. Patel, M. Zawodniok, and J. Benesty · 2020
Later among the works it cites.
Audio event detection and localization with multitask regression network
H. Phan, L. Pham, P. Koch, N. Q. K. Duong, I. McLoughlin, and A. Mertins · 2020
Later among the works it cites.
Three-dimensional source localization using sparse Bayesian learning on a spherical microphone array
G. Ping, E. Fernandez-Grande, P. Gerstoft, and Z. Chu · 2020
Later among the works it cites.
Sound event localization and detection based on CRNN using rectangular filters and channel rotation data augmentation
F. Ronchini, D. Arteaga, and A. Pérez-López · 2020
Later among the works it cites.
Sound event detection and localization using CRNN models
A. Sampathkumar and D. Kowerko · 2020
Later among the works it cites.
Exploiting attention-based sequence-to-sequence architectures for sound event localization
C. Schymura, T. Ochiai, M. Delcroix, K. Kinoshita, T. Nakatani, S. Araki, and D. Kolossa · 2020
Later among the works it cites.
Sound event localization and detection using activity-coupled cartesian DoA vector and RD3net
K. Shimada, N. Takahashi, S. Takahashi, and Y. Mitsufuji · 2020
Later among the works it cites.
A sequential system for sound event detection and localization using CRNN
R. Singla, S. Tiwari, and R. Sharma · 2020
Later among the works it cites.
Localization and detection for moving sound sources using consecutive ensembles of 2D-CRNN
J.-m. Song · 2020
Later among the works it cites.
Raw waveform based end-to-end deep convolutional network for spatial localization of multiple acoustic sources
H. Sundar, W. Wang, M. Sun, and C. Wang · 2020
Later among the works it cites.
Multiple CRNN for SELD
C. Tian · 2020
Later among the works it cites.
A deep learning framework for robust DoA estimation using spherical harmonic decomposition
V. Varanasi, H. Gupta, and R. M. Hegde · 2020
Later among the works it cites.
Exploiting periodicity features for joint detection and DoA estimation of speech sources using convolutional neural networks
R. Varzandeh, K. Adiloğlu, S. Doclo, and V. Hohmann · 2020
Later among the works it cites.
Towards domain independence in CNN-based acoustic localization using deep cross correlations
J. M. Vera-Diaz, D. Pizarro, and J. Macias-Guarasa · 2020
Later among the works it cites.
The USTC-IFLYTEK system for sound event localization and detection of DCASE 2020 challenge
Q. Wang, H. Wu, Z. Jing, F. Ma, Y. Fang, Y. Wang, T. Chen, J. Pan, J. Du, and C.-H. Lee · 2020
Later among the works it cites.
Sound event localization and detection based on multiple DoA beamforming and multi-task learning
W. Xue, Y. Tong, C. Zhang, G. Ding, X. He, and B. Zhou · 2020
Later among the works it cites.
Sound event localization based on sound intensity vector refined by DNN-based denoising and source separation
M. Yasuda, Y. Koizumi, S. Saito, H. Uematsu, and K. Imoto · 2020
Later among the works it cites.
A comprehensive survey on transfer learning
F. Zhuang, Z. Qi, K. Duan, D. Xi, Y. Zhu, H. Zhu, H. Xiong, and Q. He · 2020
Later among the works it cites.
Differentiable tracking-based training of deep learning sound source localizers
S. Adavanne, A. Politis, and T. Virtanen · 2021
Closest in time.
A survey of deep neural network in acoustic direction finding
M. Ahmad, M. Muaz, and M. Adeel · 2021
Closest in time.
DCASE 2021 Task 3: SELD system based on Resnet and random segment augmentation
J. Bai, Z. Pu, and J. Chen · 2021
Closest in time.
Semi-supervised source localization in reverberant environments with deep generative modeling
M. J. Bianco, S. Gannot, E. Fernandez-Grande, and P. Gerstoft · 2021
Closest in time.
Exploiting temporal context in CNN based multisource DoA estimation
A. Bohlender, A. Spriet, W. Tirry, and N. Madhu · 2021
Closest in time.
Acoustic reflectors localization from stereo recordings using neural networks
G. Bologni, R. Heusdens, and J. Martinez · 2021
Closest in time.
An improved event-independent network for polyphonic sound event localization and detection
Y. Cao, T. Iqbal, Q. Kong, F. An, W. Wang, and M. D. Plumbley · 2021
Closest in time.
A neural network based microphone array approach to grid-less noise source localization
P. Castellini, N. Giulietti, N. Falcionelli, A. F. Dragoni, and P. Chiariotti · 2021
Closest in time.
Multi-scale network for sound event localization and detection
P. Emmanuel, N. Parrish, and M. Horton · 2021
Closest in time.
DTU three-channel room impulse response dataset for direction of arrival estimation 2020, 2021
E. Fernandez-Grande, M. J. Bianco, S. Gannot, and P. Gerstoft · 2021
Closest in time.
Synthetic data for DNN-based DoA estimation of indoor speech
F. B. Gelderblom, Y. Liu, J. Kvam, and T. A. Myrvoll · 2021
Closest in time.
Dynamical variational autoencoders: A comprehensive review
L. Girin, S. Leglaive, X. Bie, J. Diard, T. Hueber, and X. Alameda-Pineda · 2021
Closest in time.
Deconvoluting acoustic beamforming maps with a deep neural network
W. Gonçalves Pinto, M. Bauerheim, and H. Parisot-Dupuis · 2021
Closest in time.
L3DAS21 Challenge: Machine Learning for 3D Audio Signal Processing, 2021
E. Guizzo, R. F. Gramaccioni, S. Jamili, C. Marinoni, E. Massaro, C. Medaglia, G. Nachira, L. Nucciarelli, L. Paglialunga, M. Pennese, et al · 2021
Closest in time.
Dynamically localizing multiple speakers based on the time-frequency domain
H. Hammer, S. E. Chazan, J. Goldberger, and S. Gannot · 2021
Closest in time.
A polynomial eigenvalue decomposition MUSIC approach for broadband sound source localization
A. O. Hogg, V. W. Neo, S. Weiss, C. Evers, and P. A. Naylor · 2021
Closest in time.
SSELDNET: A fully end-to-end sample-level framework for sound event localization and detection
D. Huang and R. Perez · 2021
Closest in time.
Efficient training data generation for phase-based DoA estimation
F. Hübner, W. Mack, and E. A. P. Habets · 2021
Closest in time.
MeshRIR: A dataset of room impulse responses on meshed grid points for evaluating sound field analysis and synthesis methods
S. Koyama, T. Nishida, K. Kimura, T. Abe, N. Ueno, and J. Brunnström · 2021
Closest in time.
Data diversity for improving DNN-based localization of concurrent sound events
D. Krause, A. Politis, and K. Kowalczyk · 2021
Closest in time.
Deep sound field reconstruction in real rooms: Introducing the ISOBEL sound field dataset, 2021
M. S. Kristoffersen, M. B. Møller, P. Martínez-Nuevo, and J. Østergaard · 2021
Closest in time.
Data-efficient framework for real-world multiple sound source 2D localization
G. Le Moing, P. Vinayavekhin, D. J. Agravante, T. Inoue, J. Vongkulbhisal, A. Munawar, and R. Tachibana · 2021
Closest in time.
Sound event localization and detection using cross-modal attention and parameter sharing for DCASE2021 challenge
S.-H. Lee, J.-W. Hwang, S.-B. Seo, and H.-M. Park · 2021
Closest in time.
Deep learning assisted sound source localization using two orthogonal first-order differential microphone arrays
N. Liu, H. Chen, K. Songgong, and Y. Li · 2021
Closest in time.
Multiple object tracking: A literature review
W. Luo, J. Xing, A. Milan, X. Zhang, W. Liu, and T.-K. Kim · 2021
Closest in time.
Trackformer: Multi-object tracking with transformers, 2021
T. Meinhardt, A. Kirillov, L. Leal-Taixe, and C. Feichtenhofer · 2021
Closest in time.
Sound event localisation and detection using squeeze-excitation residual CNNs
J. Naranjo-Alcazar, S. Perez-Castanos, M. Cobos, F. J. Ferri, and P. Zuccarello · 2021
Closest in time.
DCASE 2021 Task 3: spectrotemporally-aligned features for polyphonic sound event localization and detection
T. N. T. Nguyen, K. Watcharasupat, N. K. Nguyen, D. L. Jones, and W. S. Gan · 2021
Closest in time.
Deep ranking-based DoA tracking algorithm
R. Opochinsky, G. Chechik, and S. Gannot · 2021
Closest in time.
A dataset of dynamic reverberant sound scenes with directional interferers for sound event localization and detection
A. Politis, S. Adavanne, D. Krause, A. Deleforge, P. Srivastava, and T. Virtanen · 2021
Closest in time.
BeamLearning: an end-to-end deep learning approach for the angular localization of sound sources using raw multichannel acoustic pressure data
H. Pujol, E. Bavu, and A. Garcia · 2021
Closest in time.
A combination of various neural networks for sound event localization and detection
D. Rho, S. Lee, J. Park, T. Kim, J. Chang, and J. Ko · 2021
Closest in time.
Room Impulse Response Dataset - ACT, DTU Elektro (011, IEC; plane, sphere)
S. A. V. Riezu and E. F. Grande · 2021
Closest in time.
Probabilistic tracklet scoring and inpainting for multiple object tracking
F. Saleh, S. Aliakbarian, H. Rezatofighi, M. Salzmann, and S. Gould · 2021
Closest in time.
Does end-to-end trained deep model always perform better than non-end-to-end counterpart?
I. Sato, G. Liu, K. Ishikawa, T. Suzuki, and M. Tanaka · 2021
Closest in time.
PILOT: introducing Transformers for probabilistic sound event localization
C. Schymura, B. Bönninghoff, T. Ochiai, M. Delcroix, K. Kinoshita, T. Nakatani, S. Araki, and D. Kolossa · 2021
Closest in time.
Ensemble of accdoa- and einv2-based systems with d3nets and impulse response simulation for sound event localization and detection
K. Shimada, N. Takahashi, Y. Koyama, S. Takahashi, E. Tsunoo, M. Takahashi, and Y. Mitsufuji · 2021
Closest in time.
Point cloud audio processing
K. Subramani and P. Smaragdis · 2021
Closest in time.
Assessment of self-attention on learned features for sound event localization and detection
P. A. Sudarsanam, A. Politis, and K. Drossos · 2021
Closest in time.
On improved training of CNN for acoustic source localisation
E. Vargas, J. R. Hopgood, K. Brown, and K. Subr · 2021
Closest in time.
Acoustic source localization with deep generalized cross correlations
J. M. Vera-Diaz, D. Pizarro, and J. Macias-Guarasa · 2021
Closest in time.
Q. Wang, J. Du, H.-X. Wu, J. Pan, F. Ma, and C.-H. Lee · 2021
Closest in time.
Sound event localization and detection based on adaptive hybrid convolution and multi-scale feature extractor
S. Xinghao, Y. Hu, X. Zhu, and L. He · 2021
Closest in time.
The Hitachi DCASE 2021 Task 3 system: handling directive interference with self attention layers
N. Yalta, Y. Sumiyoshi, and Y. Kawaguchi · 2021
Closest in time.
On the representation of wavefronts localized in space-time and wavenumber-frequency domains
E. Zea and M. Laudato · 2021
Closest in time.
A survey on multi-task learning
Y. Zhang and Q. Yang · 2021
Closest in time.
Data augmentation and class-based ensembled CNN-Conformer networks for sound event localization and detection
Y. Zhang, S. Wang, Z. Li, K. Guo, S. Chen, and Y. Pang · 2021
Closest in time.
Signal generator, 2022
E. A. P. Habets · 2022
Closest in time.
Unsupervised multiple-object tracking with a dynamical variational autoencoder, 2022
X. Lin, L. Girin, and X. Alameda-Pineda · 2022
Closest in time.
S. Sadok, S. Leglaive, L. Girin, X. Alameda-Pineda, and R. Séguier · 2022
Closest in time.