Fetching the paper…
Reading the bibliography…
Human brain employs perceptual information about the head and eye movements to update the spatial relationship between the individual and the surrounding environment.
H. Wallach, “The role of head movements and vestibular and visual cues in sound localization.,” Journal of Experimental Psychology
1940
Earlier work this paper cites.
J. B. Allen and D. A. Berkley, “Image method for efficiently simulating small-room acoustics,” The Journal of the Acoustical Society of America
1979
Earlier work this paper cites.
R. Schmidt, “Multiple emitter location and signal parameter estimation,” IEEE transactions on antennas and propagation
1986
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, D. S. Pallett, and N. L. Dahlgren, “Darpa timit acoustic phonetic continuous speech corpus cdrom,” 1993
1993
Earlier work this paper cites.
MIT press, 1997
J. Blauert, Spatial hearing: the psychophysics of human sound localization · 1997
Earlier work this paper cites.
B. G. Shinn-Cunningham, N. I. Durlach, and R. M. Held, “Adapting to supernormal auditory localization cues. i. bias and resolution,” The Journal of the Acoustical Society of America
1998
Earlier work this paper cites.
Princeton university press Princeton, 1999
J. B. Kuipers et al · 1999
Earlier work this paper cites.
Brown University Providence, RI, 2000
J. H. DiBiase, A high-accuracy, low-latency technique for talker localization in reverberant environments using microphone arrays · 2000
Earlier work this paper cites.
S. Madgwick, “An efficient orientation filter for inertial and inertial/magnetic sensor arrays,” Report x-io and University of Bristol (UK)
2010
Earlier work this paper cites.
S. Kumar, W. Sedley, K. V. Nourski, H. Kawasaki, H. Oya, R. D. Patterson, M. A. Howard III, K. J. Friston, and T. D. Griffiths, “Predictive coding and pitch processing in the auditory cortex,” Journal of Cognitive Neuroscience
2011
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” arXiv preprint arXiv:1412.6980
2014
Earlier work this paper cites.
R. Roden, N. Moritz, S. Gerlach, S. Weinzierl, and S. Goetze, “On sound source localization of speech signals using deep neural networks,” 2015
2015
Earlier work this paper cites.
X. Xiao, S. Zhao, X. Zhong, D. L. Jones, E. S. Chng, and H. Li, “A learning-based approach to direction of arrival estimation in noisy and reverberant environments,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2015
Cited alongside, same era.
D. Genzel, U. Firzlaff, L. Wiegrebe, and P. R. MacNeilage, “Dependence of auditory spatial updating on vestibular, proprioceptive, and efference copy signals,” Journal of neurophysiology
2016
Cited alongside, same era.
T. Salimans and D. P. Kingma, “Weight normalization: A simple reparameterization to accelerate training of deep neural networks,” in Advances in neural information processing systems
2016
Cited alongside, same era.
S. Chakrabarty and E. A. Habets, “Broadband doa estimation using convolutional neural networks trained with noise signals,” in 2017 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA)
2017
Cited alongside, same era.
T. Qin, P. Li, and S. Shen, “Vins-mono: A robust and versatile monocular visual-inertial state estimator,” IEEE Transactions on Robotics
2018
Later among the works it cites.
K. Wu, V. G. Reju, and A. W. Khong, “Multisource doa estimation in a reverberant environment using a single acoustic vector sensor,” IEEE/ACM Transactions on Audio, Speech, and Language Processing
2018
Later among the works it cites.
2018
Later among the works it cites.
D. Comminiello, M. Lella, S. Scardapane, and A. Uncini, “Quaternion convolutional neural networks for detection and localization of 3d sound events,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Mur-Artal and J. D. Tardós, “Visual-inertial monocular slam with map reuse,” IEEE Robotics and Automation Letters
2017
Cited alongside, same era.
Z.-Q. Wang, X. Zhang, and D. Wang, “Robust speaker localization guided by deep learning-based time-frequency masking,” IEEE/ACM Transactions on Audio, Speech, and Language Processing
2018
Cited alongside, same era.
S. Adavanne, A. Politis, and T. Virtanen, “Direction of arrival estimation for multiple sound sources using convolutional recurrent neural network,” in 2018 26th European Signal Processing Conference (EUSIPCO)
2018
Cited alongside, same era.
2018
Cited alongside, same era.
P. Sermanet, C. Lynch, Y. Chebotar, J. Hsu, E. Jang, S. Schaal, S. Levine, and G. Brain, “Time-contrastive networks: Self-supervised learning from video,” in 2018 IEEE International Conference on Robotics and Automation (ICRA)
2018
Cited alongside, same era.
H. Zhao, C. Gan, A. Rouditchenko, C. Vondrick, J. McDermott, and A. Torralba, “The sound of pixels,” in Proceedings of the European Conference on Computer Vision (ECCV)
2018
Cited alongside, same era.
A. Owens and A. A. Efros, “Audio-visual scene analysis with self-supervised multisensory features,” in Proceedings of the European Conference on Computer Vision (ECCV)
2018
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
C. Gan, H. Zhao, P. Chen, D. Cox, and A. Torralba, “Self-supervised moving vehicle tracking with stereo sound,” in Proceedings of the IEEE International Conference on Computer Vision
2019
Later among the works it cites.
H. Liu, Z. Zhang, Y. Zhu, and S.-C. Zhu, “Self-supervised incremental learning for sound source localization in complex indoor environment,” in 2019 International Conference on Robotics and Automation (ICRA)
2019
Later among the works it cites.
2019
Later among the works it cites.
V. Tourbabin, J. Donley, B. Rafaely, and R. Mehra, “Direction of arrival estimation in highly reverberant environments using soft time-frequency mask,” in 2019 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA)
2019
Later among the works it cites.