Fetching the paper…
Reading the bibliography…
Deep learning techniques for separating audio into different sound sources face several challenges.
Demucs: Deep Extractor for Music Sources with extra unlabeled data remixed
Défossez, A.; Usunier, N.; Bottou, L.; and Bach, F. R. 2019 · 1909
Earlier work this paper cites.
Object Recognition with Gradient-Based Learning
LeCun, Y.; Haffner, P.; Bottou, L.; and Bengio, Y. 1999 · 1999
Earlier work this paper cites.
D3Net: Densely connected multidilated DenseNet for music source separation
Takahashi, N.; and Mitsufuji, Y. 2020 · 2010
Earlier work this paper cites.
MedleyDB: A Multitrack Dataset for Annotation-Intensive MIR Research
Bittner, R. M.; Salamon, J.; Tierney, M.; Mauch, M.; Cannam, C.; and Bello, J. P. 2014 · 2014
Earlier work this paper cites.
Discriminatively trained recurrent neural networks for single-channel speech separation
Weninger, F.; Hershey, J. R.; Roux, J. L.; and Schuller, B. W. 2014 · 2014
Earlier work this paper cites.
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift
Ioffe, S.; and Szegedy, C. 2015 · 2015
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Kingma, D. P.; and Ba, J. 2015 · 2015
Earlier work this paper cites.
The 2016 Signal Separation Evaluation Campaign
Liutkus, A.; Stöter, F.-R.; Rafii, Z.; Kitamura, D.; Rivet, B.; Ito, N.; Ono, N.; and Fontecave, J. 2017 · 2015
Earlier work this paper cites.
U-Net: Convolutional Networks for Biomedical Image Segmentation
Ronneberger, O.; Fischer, P.; and Brox, T. 2015 · 2015
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition
Simonyan, K.; and Zisserman, A. 2015 · 2015
Earlier work this paper cites.
Audio Set: An ontology and human-labeled dataset for audio events
Gemmeke, J. F.; Ellis, D. P. W.; Freedman, D.; Jansen, A.; Lawrence, W.; Moore, R. C.; Plakal, M.; and Ritter, M. 2017 · 2017
Earlier work this paper cites.
The MUSDB18 corpus for music separation
Rafii, Z.; Liutkus, A.; Stöter, F.-R.; Mimilakis, S. I.; and Bittner, R. 2017 · 2017
Earlier work this paper cites.
Improving music source separation based on deep neural networks through data augmentation and network blending
Uhlich, S.; Porcu, M.; Giron, F.; Enenkl, M.; Kemp, T.; Takahashi, N.; and Mitsufuji, Y. 2017 · 2017
Earlier work this paper cites.
Attention is All you Need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, L.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
Deep Learning using Rectified Linear Units (ReLU)
Agarap, A. F. 2018 · 2018
Earlier work this paper cites.
Fonseca, E.; Plakal, M.; Font, F.; Ellis, D. P. W.; Favory, X.; Pons, J.; and Serra, X. 2018 · 2018
Earlier work this paper cites.
Denoising Auto-Encoder with Recurrent Skip Connections and Residual Regression for Music Source Separation
Liu, J.; and Yang, Y. 2018 · 2018
Earlier work this paper cites.
TaSNet: Time-Domain Audio Separation Network for Real-Time, Single-Channel Speech Separation
Luo, Y.; and Mesgarani, N. 2018 · 2018
Cited alongside, same era.
Improving Single-Network Single-Channel Separation of Musical Audio with Convolutional Layers
Roma, G.; Green, O.; and Tremblay, P. A. 2018 · 2018
Cited alongside, same era.
Wave-U-Net: A Multi-Scale Neural Network for End-to-End Audio Source Separation
Stoller, D.; Ewert, S.; and Dixon, S. 2018 · 2018
Cited alongside, same era.
Mmdenselstm: An Efficient Combination of Convolutional and Recurrent Neural Networks for Audio Source Separation
Takahashi, N.; Goswami, N.; and Mitsufuji, Y. 2018 · 2018
Cited alongside, same era.
The Effect of Explicit Structure Encoding of Deep Neural Networks for Symbolic Music Generation
Chen, K.; Zhang, W.; Dubnov, S.; Xia, G.; and Li, W. 2019 · 2019
Cited alongside, same era.
Source Separation with Weakly Labelled Data: an Approach to Computational Auditory Scene Analysis
Kong, Q.; Wang, Y.; Song, X.; Cao, Y.; Wang, W.; and Plumbley, M. D. 2020b · 2020
Later among the works it cites.
Meta-Learning Extractors for Music Source Separation
Samuel, D.; Ganeshan, A.; and Naradowsky, J. 2020 · 2020
Later among the works it cites.
Sound Event Detection in Synthetic Domestic Environments
Serizel, R.; Turpault, N.; Shah, A. P.; and Salamon, J. 2020 · 2020
Later among the works it cites.
Sudo RM -RF: Efficient Networks for Universal Audio Source Separation
Tzinis, E.; Wang, Z.; and Smaragdis, P. 2020 · 2020
Later among the works it cites.
Controllable Monophonic Music Generation Via Latent Variable Disentanglement
Chen, K. 2021 · 2021
Closest in time.
Learning Audio Embeddings with User Listening Data for Content-Based Music Recommendation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A Deep Residual Network for Large-Scale Acoustic Scene Analysis
Ford, L.; Tang, H.; Grondin, F.; and Glass, J. R. 2019 · 2019
Cited alongside, same era.
Universal Sound Separation
Kavalerov, I.; Wisdom, S.; Erdogan, H.; Patton, B.; Wilson, K. W.; Roux, J. L.; and Hershey, J. R. 2019 · 2019
Cited alongside, same era.
Audio Query-based Music Source Separation
Lee, J. H.; Choi, H.; and Lee, K. 2019 · 2019
Cited alongside, same era.
End-to-End Music Source Separation: Is it Possible in the Waveform Domain?
Lluís, F.; Pons, J.; and Serra, X. 2019 · 2019
Cited alongside, same era.
Open-Unmix - A Reference Implementation for Music Source Separation
Stöter, F.; Uhlich, S.; Liutkus, A.; and Mitsufuji, Y. 2019 · 2019
Cited alongside, same era.
EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks
Tan, M.; and Le, Q. V. 2019 · 2019
Cited alongside, same era.
A Survey of Zero-Shot Learning: Settings, Methods, and Applications
Wang, W.; Zheng, V. W.; Yu, H.; and Miao, C. 2019 · 2019
Cited alongside, same era.
Chen, K.; Liang, B.; Ma, X.; and Gu, M. 2021 · 2021
Closest in time.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; Uszkoreit, J.; and Houlsby, N. 2021 · 2021
Closest in time.
TS-CAM: Token Semantic Coupled Attention Map for Weakly Supervised Object Localization
Gao, W.; Wan, F.; Pan, X.; Peng, Z.; Tian, Q.; Han, Z.; Zhou, B.; and Ye, Q. 2021 · 2021
Closest in time.
AST: Audio Spectrogram Transformer
Gong, Y.; Chung, Y.-A.; and Glass, J. 2021a · 2021
Closest in time.
PSLA: Improving Audio Tagging with Pretraining, Sampling, Labeling, and Aggregation
Gong, Y.; Chung, Y.-A.; and Glass, J. 2021b · 2021
Closest in time.
The Benefit of Temporally-Strong Labels in Audio Event Classification
Hershey, S.; Ellis, D. P. W.; Fonseca, E.; Jansen, A.; Liu, C.; Moore, R. C.; and Plakal, M. 2021 · 2021
Closest in time.
Speech enhancement with weakly labelled data from AudioSet
Kong, Q.; Liu, H.; Du, X.; Chen, L.; Xia, R.; and Wang, Y. 2021 · 2021
Closest in time.
A unified model for zero-shot music source separation, transcription and synthesis
Lin, L.; Xia, G.; Kong, Q.; and Jiang, J. 2021 · 2021
Closest in time.
Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
Liu, Z.; Lin, Y.; Cao, Y.; Hu, H.; Wei, Y.; Zhang, Z.; Lin, S.; and Guo, B. 2021 · 2021
Closest in time.
Training data-efficient image transformers & distillation through attention
Touvron, H.; Cord, M.; Douze, M.; Massa, F.; Sablayrolles, A.; and Jégou, H. 2021 · 2021
Closest in time.
What’s all the Fuss about Free Universal Sound Separation Data?
Wisdom, S.; Erdogan, H.; Ellis, D. P. W.; Serizel, R.; Turpault, N.; Fonseca, E.; Salamon, J.; Seetharaman, P.; and Hershey, J. R. 2021 · 2021
Closest in time.