Fetching the paper…
Reading the bibliography…
Given a multi-microphone recording of an unknown number of speakers talking concurrently, we simultaneously localize the sources and separate the individual speakers.
Image method for efficiently simulating small-room acoustics
Jont B Allen and David A Berkley · 1979
Earlier work this paper cites.
Coherent signal-subspace processing for the detection and estimation of angles of arrival of multiple wide-band sources
Hong Wang and Mostafa Kaveh · 1985
Earlier work this paper cites.
Multiple emitter location and signal parameter estimation
Ralph Schmidt · 1986
Earlier work this paper cites.
Array Signal Processing: Concepts and Techniques
Don H. Johnson and Dan E. Dudgeon · 1992
Earlier work this paper cites.
Blind signal separation: statistical principles
J-F Cardoso · 1998
Earlier work this paper cites.
A high-accuracy, low-latency technique for talker localization in reverberant environments using microphone arrays
Joseph Hector DiBiase · 2000
Earlier work this paper cites.
Waves: Weighted average of signal subspaces for robust wideband direction finding
Elio D Di Claudio and Raffaele Parisi · 2001
Earlier work this paper cites.
Robust sound source localization using a microphone array on a mobile robot
J-M Valin, François Michaud, Jean Rouat, and Dominic Létourneau · 2003
Earlier work this paper cites.
Sound source localization and separation based on the em algorithm
Futoshi Asano and Hideki Asoh · 2004
Earlier work this paper cites.
Audio source separation with a single sensor
Laurent Benaroya, Frédéric Bimbot, and Rémi Gribonval · 2005
Earlier work this paper cites.
On ideal binary mask as the computational goal of auditory scene analysis
DeLiang Wang · 2005
Earlier work this paper cites.
Tops: New doa estimator for wideband signals
Yeo-Sun Yoon, Lance M Kaplan, and James H McClellan · 2006
Earlier work this paper cites.
An em algorithm for localizing multiple sound sources in reverberant environments
Michael I Mandel, Daniel P Ellis, and Tony Jebara · 2007
Earlier work this paper cites.
Csr-i (wsj0) complete
John Garofalo, David Graff, Doug Paul, and David Pallett · 2007
Earlier work this paper cites.
Auralization: fundamentals of acoustics, modelling, simulation, algorithms and acoustic virtual reality
Michael Vorländer · 2007
Earlier work this paper cites.
Model-based expectation-maximization source separation and localization
Michael I Mandel, Ron J Weiss, and Daniel PW Ellis · 2009
Earlier work this paper cites.
Convolutive bss of short mixtures by ica recursively regularized across frequencies
Francesco Nesta, Piergiorgio Svaizer, and Maurizio Omologo · 2010
Earlier work this paper cites.
Non-negative matrix factorization based compensation of music for automatic speech recognition
Bhiksha Raj, Tuomas Virtanen, Sourish Chaudhuri, and Rita Singh · 2010
Earlier work this paper cites.
Underdetermined convolutive blind source separation via frequency bin-wise clustering and permutation alignment
Hiroshi Sawada, Shoko Araki, and Shoji Makino · 2010
Earlier work this paper cites.
Enforcing harmonicity and smoothness in bayesian non-negative matrix factorization applied to polyphonic music transcription
Nancy Bertin, Roland Badeau, and Emmanuel Vincent · 2010
Earlier work this paper cites.
Under-determined reverberant audio source separation using a full-rank spatial covariance model
Ngoc QK Duong, Emmanuel Vincent, and Rémi Gribonval · 2010
Earlier work this paper cites.
Beamforming techniques for multichannel audio signal separation
Hidri Adel, Meddeb Souad, Abdulqadir Alaqeeli, and Amiri Hamid · 2012
Earlier work this paper cites.
Supervised and unsupervised speech enhancement using nonnegative matrix factorization
Nasser Mohammadiha, Paris Smaragdis, and Arne Leijon · 2013
Earlier work this paper cites.
Real-time multiple sound source localization and counting using a circular microphone array
Despoina Pavlidi, Anthony Griffin, Matthieu Puigt, and Athanasios Mouchtaris · 2013
Earlier work this paper cites.
A wrapped kalman filter for azimuthal speaker tracking
J. Traa and P. Smaragdis · 2013
Cited alongside, same era.
Localization of multiple speakers under high reverberation using a spherical microphone array and the direct-path dominance test
Or Nadiri and Boaz Rafaely · 2014
Cited alongside, same era.
Multichannel source separation and tracking with ransac and directional statistics
J. Traa and P. Smaragdis · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Deep neural networks for single-channel multi-talker speech recognition
Chao Weng, Dong Yu, Michael L Seltzer, and Jasha Droppo · 2015
Cited alongside, same era.
Wave-u-net: A multi-scale neural network for end-to-end audio source separation
Daniel Stoller, Sebastian Ewert, and Simon Dixon · 2018
Later among the works it cites.
Tasnet: time-domain audio separation network for real-time, single-channel speech separation
Yi Luo and Nima Mesgarani · 2018
Later among the works it cites.
Single channel speech separation with constrained utterance level permutation invariant training using grid lstm
Chenglin Xu, Wei Rao, Xiong Xiao, Eng Siong Chng, and Haizhou Li · 2018
Later among the works it cites.
Multi-microphone neural speech separation for far-field multi-talker speech recognition
Takuya Yoshioka, Hakan Erdogan, Zhuo Chen, and Fil Alleva · 2018
Later among the works it cites.
Multi-channel overlapped speech recognition with location guided speech extraction network
Zhuo Chen, Xiong Xiao, Takuya Yoshioka, Hakan Erdogan, Jinyu Li, and Yifan Gong · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yuval Dorfan, Dani Cherkassky, and Sharon Gannot · 2015
Cited alongside, same era.
Acoustic space learning for sound-source separation and localization on binaural manifolds
Antoine Deleforge, Florence Forbes, and Radu Horaud · 2015
Cited alongside, same era.
Directional nmf for joint source localization and separation
Johannes Traa, Paris Smaragdis, Noah D Stein, and David Wingate · 2015
Cited alongside, same era.
Librispeech: an asr corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur · 2015
Cited alongside, same era.
Generalized wiener filtering with fractional power spectrograms
Antoine Liutkus and Roland Badeau · 2015
Cited alongside, same era.
Multichannel audio source separation with deep neural networks
Aditya Arie Nugraha, Antoine Liutkus, and Emmanuel Vincent · 2016
Cited alongside, same era.
Deep clustering: Discriminative embeddings for segmentation and separation
John R Hershey, Zhuo Chen, Jonathan Le Roux, and Shinji Watanabe · 2016
Cited alongside, same era.
The sound of pixels
Hang Zhao, Chuang Gan, Andrew Rouditchenko, Carl Vondrick, Josh McDermott, and Antonio Torralba · 2018
Later among the works it cites.
Deep neural networks for multiple speaker detection and localization
Weipeng He, Petr Motlicek, and Jean-Marc Odobez · 2018
Later among the works it cites.
Sound event localization and detection of overlapping sources using convolutional recurrent neural networks
Sharath Adavanne, Archontis Politis, Joonas Nikunen, and Tuomas Virtanen · 2018
Later among the works it cites.
Latent gaussian activity propagation: using smoothness and structure to separate and localize sounds in large noisy environments
Daniel Johnson, Daniel Gorelik, Ross E Mawhorter, Kyle Suver, Weiqing Gu, Steven Xing, Cody Gabriel, and Peter Sankhagowit · 2018
Later among the works it cites.
Pyroomacoustics: A python package for audio room simulation and array processing algorithms
Robin Scheibler, Eric Bezzam, and Ivan Dokmanić · 2018
Later among the works it cites.
The 2018 signal separation evaluation campaign
Fabian-Robert Stöter, Antoine Liutkus, and Nobutaka Ito · 2018
Later among the works it cites.
Unsupervised deep clustering for source separation: Direct learning from mixtures using spatial information
Efthymios Tzinis, Shrikant Venkataramani, and Paris Smaragdis · 2019
Later among the works it cites.
Conv-tasnet: Surpassing ideal time–frequency magnitude masking for speech separation
Yi Luo and Nima Mesgarani · 2019
Later among the works it cites.
Self-supervised audio-visual co-segmentation
Andrew Rouditchenko, Hang Zhao, Chuang Gan, Josh McDermott, and Antonio Torralba · 2019
Later among the works it cites.
Multiple sound source localization with svd-phat
François Grondin and James Glass · 2019
Later among the works it cites.
Recursive speech separation for unknown number of speakers
Naoya Takahashi, Sudarsanam Parthasaarathy, Nabarun Goswami, and Yuki Mitsufuji · 2019
Later among the works it cites.
Demucs: Deep extractor for music sources with extra unlabeled data remixed
Alexandre Défossez, Nicolas Usunier, Léon Bottou, and Francis Bach · 2019
Later among the works it cites.
Sdr–half-baked or well done?
Jonathan Le Roux, Scott Wisdom, Hakan Erdogan, and John R Hershey · 2019
Later among the works it cites.
Source separation with deep generative priors
Vivek Jayaram and John Thickstun · 2020
Closest in time.
Enhancing end-to-end multi-channel speech separation via spatial feature learning
Rongzhi Gu, Shi-Xiong Zhang, Lianwu Chen, Yong Xu, Meng Yu, Dan Su, Yuexian Zou, and Dong Yu · 2020
Closest in time.
Real-time binaural speech separation with preserved spatial cues
Cong Han, Yi Luo, and Nima Mesgarani · 2020
Closest in time.
End-to-end microphone permutation and number invariant multi-channel speech separation
Yi Luo, Zhuo Chen, Nima Mesgarani, and Takuya Yoshioka · 2020
Closest in time.
Voice separation with an unknown number of multiple speakers
Eliya Nachmani, Yossi Adi, and Lior Wolf · 2020
Closest in time.