Fetching the paper…
Reading the bibliography…
This paper proposes a multichannel source separation technique called the multichannel variational autoencoder (MVAE) method, which uses a conditional VAE (CVAE) to model and estimate the power spectrograms of the sources in a mixture.
“Non-negative matrix factorization for polyphonic music transcription,”
P. Smaragdis, · 2003
Earlier work this paper cites.
“Independent vector analysis: An extension of ICA to multivariate components,”
T. Kim, T. Eltoft, and T.-W. Lee, · 2006
Earlier work this paper cites.
“Solution of permutation problem in frequency domain ICA using multivariate probability density functions,”
A. Hiroe, · 2006
Earlier work this paper cites.
“Selective amplifier of periodic and non-periodic components in concurrent audio signals with spectral control envelopes,”
H. Kameoka, M. Goto, and S. Sagayama, · 2006
Earlier work this paper cites.
“Performance measurement in blind audio source separation,”
E. Vincent, R. Gribonval, and C. Févotte, · 2006
Earlier work this paper cites.
“Nonnegative matrix factorization with the Itakura-Saito divergence,”
C. Févotte, N. Bertin, and J.-L. Durrieu, · 2009
Earlier work this paper cites.
“Multichannel nonnegative matrix factorization in convolutive mixtures for audio source separation,”
A. Ozerov and C. Févotte, · 2010
Earlier work this paper cites.
“Statistical model of speech signals based on composite autoregressive system with application to blind source separation,”
H. Kameoka, T. Yoshioka, M. Hamamura, J. Le Roux, and K. Kashino, · 2010
Earlier work this paper cites.
“Convergence-guaranteed multiplicative algorithms for non-negative matrix factorization with beta-divergence,”
M. Nakano, H. Kameoka, J. Le Roux, N. Ono, and S. Sagayama, · 2010
Earlier work this paper cites.
“Stable and fast update rules for independent vector analysis based on auxiliary function technique,”
N. Ono, · 2011
Earlier work this paper cites.
“Algorithms for nonnegative matrix factorization with the
C. Févotte and J. Idier, · 2011
Cited alongside, same era.
“Blind separation and dereverberation of speech mixtures by joint optimization,”
T. Yoshioka, T. Nakatani, M. Miyoshi, and H. G. Okuno, · 2011
Cited alongside, same era.
“Multichannel extensions of non-negative matrix factorization with complex-valued data,”
H. Sawada, H. Kameoka, S. Araki, and N. Ueda, · 2013
Cited alongside, same era.
“Auto-encoding variational Bayes,”
D. P. Kingma and M. Welling, · 2014
Cited alongside, same era.
“Semi-supervised learning with deep generative models,”
D. P. Kingma and D. J. Rezendey, S. Mohamedy, and M. Welling, · 2014
Cited alongside, same era.
“Generative adversarial nets,”
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio, · 2014
“Sequence-to-sequence voice conversion with similarity metric learned using generative adversarial networks,”
T. Kaneko, H. Kameoka, K. Hiramatsu, and K. Kashino, · 2017
Later among the works it cites.
“Parallel-data-free voice conversion using cycle-consistent adversarial networks,”
T. Kaneko and H. Kameoka, · 2017
Later among the works it cites.
“Determined blind source separation with independent low-rank matrix analysis,”
D. Kitamura, N. Ono, H. Sawada, H. Kameoka, and H. Saruwatari, · 2018
Closest in time.
“Experimental evaluation of multichannel audio source separation based on IDLMA,”
D. Kitamura, H. Sumino, N. Takamune, S. Takamichi, H. Saruwatari, and N. Ono, · 2018
Closest in time.
“Statistical speech enhancement based on probabilistic integration of variational autoencoder and non-negative matrix factorization,”
Y. Bando, M. Mimura, K. Itoyama, K. Yoshii, and T. Kawahara, · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Adam: A method for stochastic optimization,”
D. Kingma and J. Ba, · 2015
Cited alongside, same era.
“Determined blind source separation unifying independent vector analysis and nonnegative matrix factorization,”
D. Kitamura, N. Ono, H. Sawada, H. Kameoka, and H. Saruwatari, · 2016
Cited alongside, same era.
“Multichannel audio source separation with deep neural networks,”
A. A. Nugraha, A. Liutkus, and E. Vincent, · 2016
Cited alongside, same era.
“Language modeling with gated convolutional networks,”
Y. N. Dauphin, A. Fan, M. Auli, and D. Grangier, · 2017
Cited alongside, same era.
“Generative adversarial source separation,”
Y. Subakan and P. Smaragdis, · 2018
Closest in time.
“StarGAN-VC: Non-parallel many-to-many voice conversion with star generative adversarial networks,”
H. Kameoka, T. Kaneko, K. Tanaka, and N. Hojo, · 2018
Closest in time.
“Deep clustering with gated convolutional networks,”
L. Li and H. Kameoka, · 2018
Closest in time.
“The voice conversion challenge 2018: Promoting development of parallel and nonparallel methods,”
J. Lorenzo-Trueba, J. Yamagishi, T. Toda, D. Saito, F. Villavicencio, T. Kinnunen, and Z. Ling, · 2018
Closest in time.
“Joint separation and dereverberation of reverberant mixtures with determined multichannel non-negative matrix factorization,”
H. Kagami, H. Kameoka, and M. Yukawa, · 2018
Closest in time.