Fetching the paper…
Reading the bibliography…
We introduce a new audio processing technique that increases the sampling rate of signals such as speech or music using deep convolutional neural networks.
Distance measures for speech processing
Augustine Gray and John Markel · 1976
Earlier work this paper cites.
Statistical recovery of wideband speech from narrowband speech
Yan Ming Cheng, Douglas O’Shaughnessy, and Paul Mermelstein · 1994
Earlier work this paper cites.
Linear predictive coding
Jeremy Bradbury · 2000
Earlier work this paper cites.
Narrowband to wideband conversion of speech using gmm based transformation
Kun-Youl Park and Hyung Soon Kim · 2000
Earlier work this paper cites.
Applying lstm to time series predictable through time-window approaches
Felix A Gers, Douglas Eck, and Jürgen Schmidhuber · 2001
Earlier work this paper cites.
Bandwidth extension of audio signals by spectral band replication
Per Ekstrand · 2002
Earlier work this paper cites.
Graphical models and automatic speech recognition
Jeffrey A Bilmes · 2004
Earlier work this paper cites.
Bandwidth expansion of narrowband speech using non-negative matrix factorization
Dhananjay Bansal, Bhiksha Raj, and Paris Smaragdis · 2005
Earlier work this paper cites.
The cocktail party problem
Simon Haykin and Zhe Chen · 2005
Earlier work this paper cites.
Audio bandwidth extension: application of psychoacoustics, signal processing and loudspeaker design
Erik Larsen and Ronald M Aarts · 2005
Earlier work this paper cites.
Automated classification of bird and amphibian calls using machine learning: A comparison of methods
Miguel A Acevedo, Carlos J Corrada-Bravo, Héctor Corrada-Bravo, Luis J Villanueva-Rivera, and T Mitchell Aide · 2009
Cited alongside, same era.
Speech bandwidth extension using gaussian mixture model-based estimation of the highband mel spectrum
Hannu Pulakka, Ulpu Remes, Kalle Palomäki, Mikko Kurimo, and Paavo Alku · 2011
Cited alongside, same era.
Multivariate autoregressive mixture models for music auto-tagging
Emanuele Coviello, Yonatan Vaizman, Antoni B Chan, and Gert RG Lanckriet · 2012
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Geoffrey Hinton, Li Deng, Dong Yu, George E Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara N Sainath, et al · 2012
Cited alongside, same era.
Recurrent neural networks for noise reduction in robust asr
Andrew Maas, Quoc V. Le, Tyler M. ONeil, Oriol Vinyals, Patrick Nguyen, and Andrew Y. Ng · 2012
Cited alongside, same era.
Dnn-based speech bandwidth expansion and its application to adding high-frequency missing features for automatic speech recognition of narrowband speech
Kehuang Li, Zhen Huang, Yong Xu, and Chin-Hui Lee · 2015
Later among the works it cites.
Content-aware collaborative music recommendation using pre-trained neural networks
Dawen Liang, Minshu Zhan, and Daniel PW Ellis · 2015
Later among the works it cites.
Soundnet: Learning sound representations from unlabeled video
Yusuf Aytar, Carl Vondrick, and Antonio Torralba · 2016
Later among the works it cites.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A Efros · 2016
Later among the works it cites.
Photo-realistic single image super-resolution using a generative adversarial network
Christian Ledig, Lucas Theis, Ferenc Huszar, Jose Caballero, Andrew P. Aitken, Alykhan Tejani, Johannes Totz, Zehan Wang, and Wenzhe Shi · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
English multi-speaker corpus for cstr voice cloning toolkit, 2012
Junichi Yamagishi · 2012
Cited alongside, same era.
Beta process sparse nonnegative matrix factorization for music
Dawen Liang, Matthew D. Hoffman, and Daniel P. W. Ellis · 2013
Cited alongside, same era.
Non-negative matrix completion for bandwidth extension: A convex optimization approach
Dennis L. Sun and Rahul Mazumder · 2013
Cited alongside, same era.
Improving content-based and hybrid music recommendation using deep learning
Xinxi Wang and Ye Wang · 2014
Cited alongside, same era.
Image super-resolution using deep convolutional networks
Chao Dong, Chen Change Loy, Kaiming He, and Xiaoou Tang · 2015
Cited alongside, same era.
Samplernn: An unconditional end-to-end neural audio generation model, 2016
Soroush Mehri, Kundan Kumar, Ishaan Gulrajani, Rithesh Kumar, Shubham Jain, Jose Sotelo, Aaron Courville, and Yoshua Bengio · 2016
Later among the works it cites.
Deconvolution and checkerboard artifacts
Augustus Odena, Vincent Dumoulin, and Chris Olah · 2016
Later among the works it cites.
Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network
Wenzhe Shi, Jose Caballero, Ferenc Huszar, Johannes Totz, Andrew P. Aitken, Rob Bishop, Daniel Rueckert, and Zehan Wang · 2016
Later among the works it cites.
Wavenet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew W. Senior, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Colorful image colorization
Richard Zhang, Phillip Isola, and Alexei A Efros · 2016
Later among the works it cites.