Fetching the paper…
Reading the bibliography…
Time-frequency (TF) representations provide powerful and intuitive features for the analysis of time series such as audio.
On a Hilbert space of analytic functions and an associated integral transform part i
Bargmann, V · 1961
Earlier work this paper cites.
Functions of one complex variable
Conway, J. B · 1973
Earlier work this paper cites.
Implementation of the digital phase vocoder using the fast fourier transform
Portnoff, M · 1976
Earlier work this paper cites.
Short term spectral analysis, synthesis, and modification by discrete fourier transform
Allen, J · 1977
Earlier work this paper cites.
Signal estimation from modified short-time fourier transform
Griffin, D. and Lim, J · 1984
Earlier work this paper cites.
The phase vocoder: A tutorial
Dolson, M · 1986
Earlier work this paper cites.
Discrete gabor expansions
Wexler, J. and Raz, S · 1990
Earlier work this paper cites.
Calculation of a constant Q spectral transform
Brown, J. C · 1991
Earlier work this paper cites.
A Practical Guide to Data Analysis for Physical Science Students
Lyons, L · 1991
Earlier work this paper cites.
An efficient algorithm for the calculation of a constant Q transform
Brown, J. C. and Puckette, M. S · 1992
Earlier work this paper cites.
Improving the readability of time-frequency and time-scale representations by the reassignment method
Auger, F. and Flandrin, P · 1995
Earlier work this paper cites.
From continuous to discrete Weyl-Heisenberg frames through sampling
Janssen, A · 1997
Earlier work this paper cites.
Numerical algorithms for discrete Gabor expansions
Strohmer, T · 1998
Earlier work this paper cites.
Foundations of Time-Frequency Analysis
Gröchenig, K · 2001
Earlier work this paper cites.
Optimal OFDM system design for time-frequency dispersive channels
Strohmer, T. and Beaver, S · 2003
Earlier work this paper cites.
On signal reconstruction without phase
Balan, R., Casazza, P., and Edidin, D · 2006
Earlier work this paper cites.
Fast signal reconstruction from magnitude STFT spectrogram based on spectrogram consistency
Le Roux, J., Kameoka, H., Ono, N., and Sagayama, S · 2010
Earlier work this paper cites.
Time-frequency processing
Arfib, D., Keiler, F., Zölzer, U., Verfaille, V., and Bonada, J · 2011
Earlier work this paper cites.
On phase-magnitude relationships in the short-time fourier transform
Auger, F., Chassande-Mottin, É., and Flandrin, P · 2012
Earlier work this paper cites.
A framework for invertible, real-time constant-Q transforms
Holighaus, N., Dörfler, M., Velasco, G. A., and Grill, T · 2013
Cited alongside, same era.
A fast Griffin-Lim algorithm
Perraudin, N., Balazs, P., and Søndergaard, P. L · 2013
Cited alongside, same era.
End-to-end learning for music audio
Dieleman, S. and Schrauwen, B · 2014
Cited alongside, same era.
Generative adversarial nets
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y · 2014
Cited alongside, same era.
The Large Time-Frequency Analysis Toolbox 2.0
Průša, Z., Søndergaard, P. L., Holighaus, N., Wiesmeyr, C., and Balazs, P · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. and Ba, J · 2015
Cited alongside, same era.
SampleRNN: An unconditional end-to-end neural audio generation model
Mehri, S., Kumar, K., Gulrajani, I., Kumar, R., Jain, S., Sotelo, J., Courville, A., and Bengio, Y · 2017
Later among the works it cites.
SEGAN: Speech enhancement generative adversarial network
Pascual, S., Bonafonte, A., and Serrà, J · 2017
Later among the works it cites.
End-to-end learning for music audio tagging at scale
Pons, J., Nieto, O., Prockup, M., Schmidt, E. M., Ehmann, A. F., and Serra, X · 2017
Later among the works it cites.
A noniterative method for reconstruction of phase from STFT magnitude
Průša, Z., Balazs, P., and Søndergaard, P · 2017
Later among the works it cites.
Char2wav: End-to-end speech synthesis
Sotelo, J., Mehri, S., Kumar, K., Santos, J. F., Kastner, K., Courville, A., and Bengio, Y · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Phase retrieval for the cauchy wavelet transform
Mallat, S. and Waldspurger, I · 2015
Cited alongside, same era.
STFT and DGT phase conventions and phase derivatives interpretation
Pruša, Z · 2015
Cited alongside, same era.
The pole behavior of the phase derivative of the short-time fourier transform
Balazs, P., Bayer, D., Jaillet, F., and Søndergaard, P · 2016
Cited alongside, same era.
An Introduction to Frames and Riesz Bases
Christensen, O · 2016
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
Radford, A., Metz, L., and Chintala, S · 2016
Cited alongside, same era.
Improved techniques for training GANs
Salimans, T., Goodfellow, I., Zaremba, W., Cheung, V., Radford, A., and Chen, X · 2016
Cited alongside, same era.
Fan, Z.-C., Lai, Y.-L., and Jang, J.-S. R · 2018
Later among the works it cites.
Progressive growing of GANs for improved quality, stability, and variation
Karras, T., Aila, T., Laine, S., and Lehtinen, J · 2018
Later among the works it cites.
A context encoder for audio inpainting
Marafioti, A., Perraudin, N., Holighaus, N., and Majdak, P · 2018
Later among the works it cites.
Improving DNN-based music source separation using phase features
Muth, J., Uhlich, S., Perraudin, N., Kemp, T., Cardinaux, F., and Mitsufuji, Y · 2018
Later among the works it cites.
Audlet filter banks: A versatile analysis/synthesis framework using auditory frequency scales
Necciari, T., Holighaus, N., Balazs, P., Průša, Z., Majdak, P., and Derrien, O · 2018
Later among the works it cites.
Natural TTS synthesis by conditioning WaveNet on mel spectrogram predictions
Shen, J., Pang, R., Weiss, R., Schuster, M., Jaitly, N., Yang, Z., Chen, Z., Zhang, Y., Wang, Y., Skerry-Ryan, R., Saurous, R., Agiomyrgiannakis, Y., and Wu, Y · 2018
Later among the works it cites.
Parallel WaveNet: Fast high-fidelity speech synthesis
Van Den Oord, A., Li, Y., Babuschkin, I., Simonyan, K., Vinyals, O., Kavukcuoglu, K., van den Driessche, G., Lockhart, E., Cobo, L., Stimberg, F., Casagrande, N., Grewe, D., Noury, S., Dieleman, S., Elsen, E., Kalchbrenner, N., Zen, H., Graves, A., King, H., Walters, T., Belov, D., and Hassabis, D · 2018
Later among the works it cites.
Speech commands: A dataset for limited-vocabulary speech recognition
Warden, P · 2018
Later among the works it cites.
Stability estimates for phase retrieval from discrete gabor measurements
Alaifari, R. and Wellershoff, M · 2019
Closest in time.
Large scale GAN training for high fidelity natural image synthesis
Brock, A., Donahue, J., and Simonyan, K · 2019
Closest in time.
Adversarial audio synthesis
Donahue, C., McAuley, J., and Puckette, M · 2019
Closest in time.
GANSynth: Adversarial neural audio synthesis
Engel, J., Agrawal, K. K., Chen, S., Gulrajani, I., Donahue, C., and Roberts, A · 2019
Closest in time.
Characterization of analytic wavelet transforms and a new phaseless reconstruction algorithm
Holighaus, N., Koliander, G., Průša, Z., and Abreu, L. D · 2019
Closest in time.
TimbreTron: A WaveNet (CycleGAN (CQT (Audio))) pipeline for musical timbre transfer
Huang, S., Li, Q., Anil, C., Bao, X., Oore, S., and Grosse, R. B · 2019
Closest in time.