Fetching the paper…
Reading the bibliography…
We propose the multi-head convolutional neural network (MCNN) architecture for waveform synthesis from spectrograms.
D. Griffin and J. Lim, “Signal estimation from modified short-time fourier transform,” IEEE Transactions on Acoustics, Speech, and Signal Processing , vol. 32, no. 2, pp. 236–243, Apr 1984
1984
Earlier work this paper cites.
D. L. Sun and J. O. Smith, III, “Estimating a Signal from a Magnitude Spectrogram via Convex Optimization,” arXiv: 1209.2076 , 2012
2012
Earlier work this paper cites.
H. Jia, Y. Zhang, G. Long, J. Xu, S. Yan, and Y. Li, “Gpuroofline: A model for guiding performance optimizations on gpus,” in Euro-Par 2012 Parallel Processing , C. Kaklamanis, T. Papatheodorou, and P. G. Spirakis, Eds. Berlin, Heidelberg: Springer Berlin Heidelberg, 2012
2012
Earlier work this paper cites.
N. Perraudin, P. Balazs, and P. L. Søndergaard, “A fast griffin-lim algorithm,” in 2013 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics , Oct 2013, pp. 1–4
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
G. T. Beauregard, M. Harish, and L. Wyse, “Single pass spectrogram inversion,” in 2015 IEEE International Conference on Digital Signal Processing (DSP) , July 2015, pp. 427–431
2015
Earlier work this paper cites.
D. Clevert, T. Unterthiner, and S. Hochreiter, “Fast and accurate deep network learning by exponential linear units,” arXiv: 1511.07289 , 2015
2015
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an ASR corpus based on public domain audio books,” in Acoustics, Speech and Signal Processing (ICASSP), 2015 IEEE International Conference on . IEEE, 2015, pp. 5206–5210
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
V. Dumoulin and F. Visin, “A guide to convolution arithmetic for deep learning,” arXiv: 1603.07285 , 2016
2016
Cited alongside, same era.
A. Odena, V. Dumoulin, and C. Olah, “Deconvolution and checkerboard artifacts,” Distill , 2016. [Online]. Available: http://distill.pub/2016/deconv-checkerboard
2016
Cited alongside, same era.
Z. Prusa, P. Balazs, and P. L. Sondergaard, “A noniterative method for reconstruction of phase from stft magnitude,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 25, no. 5, pp. 1154–1164, May 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
N. P. Jouppi, C. Young, N. Patil et al. , “In-datacenter performance analysis of a tensor processing unit,” SIGARCH Comput. Archit. News , vol. 45, no. 2, pp. 1–12, Jun. 2017
2017
Later among the works it cites.
E. Grinstein, N. Q. K. Duong, A. Ozerov, and P. Pérez, “Audio style transfer,” arXiv: 1710.11385 , 2017
2017
Later among the works it cites.
C. Donahue, B. Li, and R. Prabhavalkar, “Exploring speech enhancement with generative adversarial networks for robust speech recognition,” arXiv: 1711.05747 , 2017
2017
Later among the works it cites.
J. Engel, C. Resnick, A. Roberts et al. , “Neural audio synthesis of musical notes with wavenet autoencoders,” arXiv: 1704.01279 , 2017
2017
Later among the works it cites.
S. Ö. Arik, M. Chrzanowski, A. Coates et al. , “Deep voice: Real-time neural text-to-speech,” arXiv: 1702.07825 , 2017
2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
S. Ö. Arik, G. F. Diamos, A. Gibiansky et al. , “Deep voice 2: Multi-speaker neural text-to-speech,” arXiv: 1705.08947 , 2017
2017
Cited alongside, same era.
W. Ping, K. Peng, A. Gibiansky et al. , “Deep voice 3: 2000-speaker neural text-to-speech,” arXiv: 1710.07654 , 2017
2017
Cited alongside, same era.
Y. Wang, R. J. Skerry-Ryan, D. Stanton et al. , “Tacotron: A fully end-to-end text-to-speech synthesis model,” arXiv: 1703.10135 , 2017
2017
Cited alongside, same era.
Later among the works it cites.
S. O. Arik, J. Chen, K. Peng, W. Ping, and Y. Zhou, “Neural Voice Cloning with a Few Samples,” arXiv: 1802.06006 , 2018
2018
Closest in time.
“Fft benchmark methodology,” http://www.fftw.org/speed/method.html
2018
Closest in time.
“Tensorflow profiler and advisor,” https://github.com/tensorflow/tensorflow/blob/master/tensorflow/core/profiler/README.md
2018
Closest in time.