Fetching the paper…
Reading the bibliography…
Time-frequency (TF) representations in audio synthesis have been increasingly modeled with real-valued networks.
Polynomial theory of complex systems
Alexey Grigorevich Ivakhnenko · 1971
Earlier work this paper cites.
Signal estimation from modified short-time fourier transform
Daniel Griffin and Jae Lim · 1984
Earlier work this paper cites.
Learning, invariance, and generalization in high-order neural networks
C Lee Giles and Tom Maxwell · 1987
Earlier work this paper cites.
Real and Complex Analysis
W. Rudin, W.A. RUDIN, and Tata McGraw-Hill Publishing Company · 1987
Earlier work this paper cites.
Modification of backpropagation networks for complex-valued signal processing in frequency domain
M.S. Kim and C.C. Guest · 1990
Earlier work this paper cites.
The pi-sigma network: An efficient higher-order neural network for pattern classification and function approximation
Yoan Shin and Joydeep Ghosh · 1991
Earlier work this paper cites.
Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs
Antony W Rix, John G Beerends, Michael P Hollier, and Andries P Hekstra · 2001
Earlier work this paper cites.
A sigma-pi-sigma neural network (spsnn)
Chien-Kuo Li · 2003
Earlier work this paper cites.
Polynomial neural networks architecture: analysis and design
Sung-Kwun Oh, Witold Pedrycz, and Byoung-Jun Park · 2003
Earlier work this paper cites.
Canonical correlation analysis: An overview with application to learning methods
David R. Hardoon, Sandor R. Szedmak, and John R. Shawe-taylor · 2004
Earlier work this paper cites.
Evaluation of objective quality measures for speech enhancement
Yi Hu and Philipos C Loizou · 2007
Earlier work this paper cites.
Training pi-sigma network by online gradient algorithm with penalty for small weight update
Yan Xiong, Wei Wu, Xidai Kang, and Chao Zhang · 2007
Earlier work this paper cites.
Tensor decompositions and applications
Tamara G Kolda and Brett W Bader · 2009
Earlier work this paper cites.
Mnist handwritten digit database
Yann LeCun, Corinna Cortes, and CJ Burges · 2010
Earlier work this paper cites.
Comparison of complex-and real-valued feedforward neural networks in their generalization ability
Akira Hirose and Shotaro Yoshida · 2011
Earlier work this paper cites.
Crowdmos: An approach for crowdsourcing mean opinion score studies
Flavio P. Ribeiro, Dinei A. F. Florêncio, Cha Zhang, and Michael L. Seltzer · 2011
Earlier work this paper cites.
The diverse environments multi-channel acoustic noise database (demand): A database of multichannel environmental noise recordings
Joachim Thiemann, Nobutaka Ito, and Emmanuel Vincent · 2013
Earlier work this paper cites.
The voice bank corpus: Design, collection and data analysis of a large regional accent speech database
Christophe Veaux, Junichi Yamagishi, and Simon King · 2013
Earlier work this paper cites.
End-to-end learning for music audio
Dieleman, Sander and Schrauwen, Benjamin · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Neuronal synchrony in complex-valued deep networks
David P. Reichert and Thomas Serre · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Deep roto-translation scattering for object classification
Edouard Oyallon and Stéphane Mallat · 2015
Earlier work this paper cites.
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2015
Earlier work this paper cites.
Unitary evolution recurrent neural networks
Martín Arjovsky, Amar Shah, and Yoshua Bengio · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Wavenet: A generative model for raw audio
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Real-time spectrogram inversion using phase gradient heap integration
Zdenek Pruša and Peter L Søndergaard · 2016
Cited alongside, same era.
Improved techniques for training gans
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, Xi Chen, and Xi Chen · 2016
Cited alongside, same era.
Complex embeddings for simple link prediction
Non-local neural networks
Xiaolong Wang, Ross Girshick, Abhinav Gupta, and Kaiming He · 2018
Later among the works it cites.
Speech commands: A dataset for limited-vocabulary speech recognition
Pete Warden · 2018
Later among the works it cites.
Activation maximization generative adversarial nets
Zhiming Zhou, Han Cai, Shu Rong, Yuxuan Song, Kan Ren, Weinan Zhang, Jun Wang, and Yong Yu · 2018
Later among the works it cites.
Large scale GAN training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, and Karen Simonyan · 2019
Later among the works it cites.
Polygan: High-order polynomial generators
Grigorios Chrysos, Stylianos Moschoglou, Yannis Panagakis, and Stefanos Zafeiriou · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Théo Trouillon, Johannes Welbl, Sebastian Riedel, Éric Gaussier, and Guillaume Bouchard · 2016
Cited alongside, same era.
A mathematical motivation for complex-valued convolutional networks
Mark Tygert, Joan Bruna, Soumith Chintala, Yann LeCun, Serkan Piantino, and Arthur Szlam · 2016
Cited alongside, same era.
WaveNet: A Generative Model for Raw Audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
Neural audio synthesis of musical notes with wavenet autoencoders, 2017
Jesse Engel, Cinjon Resnick, Adam Roberts, Sander Dieleman, Douglas Eck, Karen Simonyan, and Mohammad Norouzi · 2017
Cited alongside, same era.
Improved training of wasserstein gans
Ishaan Gulrajani, Faruk Ahmed, Martín Arjovsky, Vincent Dumoulin, and Aaron C. Courville · 2017
Cited alongside, same era.
Deligan: Generative adversarial networks for diverse and limited data
Swaminathan Gurumurthy, Ravi Kiran Sarvadevabhatla, and R Venkatesh Babu · 2017
Cited alongside, same era.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Cited alongside, same era.
Chris Donahue, Julian J. McAuley, and Miller S. Puckette · 2019
Later among the works it cites.
GANSynth: Adversarial neural audio synthesis
Jesse H. Engel, Kumar Krishna Agrawal, Shuo Chen, Ishaan Gulrajani, Chris Donahue, and Adam Roberts · 2019
Later among the works it cites.
A style-based generator architecture for generative adversarial networks
Tero Karras, Samuli Laine, and Timo Aila · 2019
Later among the works it cites.
Adversarial generation of time-frequency features with application in audio synthesis
Andrés Marafioti, Nathanaël Perraudin, Nicki Holighaus, and Piotr Majdak · 2019
Later among the works it cites.
Encrypted speech recognition using deep polynomial networks
Shi-Xiong Zhang, Yifan Gong, and Dong Yu · 2019
Later among the works it cites.
P-nets: Deep polynomial neural networks
Grigorios G. Chrysos, Stylianos Moschoglou, Giorgos Bouritsas, Yannis Panagakis, Jiankang Deng, and Stefanos Zafeiriou · 2020
Later among the works it cites.
Improved speech synthesis using generative adversarial networks
Dineshraj Gunasekaran, Gautham Venkatraj, Eoin Brophy, and Tomas E. Ward · 2020
Later among the works it cites.
Kazi Nazmul Haque, Rajib Rana, John HL Hansen, and Björn Schuller · 2020
Later among the works it cites.
Multiplicative interactions and where to find them
Siddhant M. Jayakumar, Wojciech M. Czarnecki, Jacob Menick, Jonathan Schwarz, Jack Rae, Simon Osindero, Yee Whye Teh, Tim Harley, and Razvan Pascanu · 2020
Later among the works it cites.
Conditional Spoken Digit Generation with StyleGAN
Kasperi Palkama, Lauri Juvela, and Alexander Ilin · 2020
Later among the works it cites.
Convolutional tensor-train lstm for spatio-temporal learning
Jiahao Su, Wonmin Byeon, Jean Kossaifi, Furong Huang, Jan Kautz, and Anima Anandkumar · 2020
Later among the works it cites.
Complex transformer: A framework for modeling complex-valued sequence
Muqiao Yang, Martin Q. Ma, Dongyu Li, Yao-Hung Hubert Tsai, and Ruslan Salakhutdinov · 2020
Later among the works it cites.
A survey of complex-valued neural networks
Joshua Bassey, Lijun Qian, and Xianfang Li · 2021
Later among the works it cites.
Expressivity and trainability of quadratic networks
Feng-Lei Fan, Mengzhou Li, Fei Wang, Rongjie Lai, and Ge Wang · 2021
Later among the works it cites.
Diffwave: A versatile diffusion model for audio synthesis
Zhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao, and Bryan Catanzaro · 2021
Later among the works it cites.
Flowtron: an autoregressive flow-based generative network for text-to-speech synthesis
Rafael Valle, Kevin J. Shih, Ryan Prenger, and Bryan Catanzaro · 2021
Later among the works it cites.
It’s raw! audio generation with state-space models
Karan Goel, Albert Gu, Chris Donahue, and Christopher Ré · 2022
Closest in time.
Efficiently modeling long sequences with structured state spaces
Albert Gu, Karan Goel, and Christopher Re · 2022
Closest in time.