Fetching the paper…
Reading the bibliography…
Speech codecs learn compact representations of speech signals to facilitate data transmission.
D. A. Huffman, “A method for the construction of minimum-redundancy codes,”
1952
Earlier work this paper cites.
B. S. Atal and M. R. Schroeder, “Adaptive predictive coding of speech signals,”
1970
Earlier work this paper cites.
M. Schroeder and B. Atal, “Code-excited linear prediction (CELP): High-quality speech at very low bit rates,” in
1985
Earlier work this paper cites.
J. Makhoul, S. Roucos, and H. Gish, “Vector quantization in speech coding,”
1985
Earlier work this paper cites.
L. R. Welch and E. R. Berlekamp, “Error correction for algebraic block codes,” Dec. 30 1986, uS Patent 4,633,470
1986
Earlier work this paper cites.
L.-J. Liu, Z.-H. Ling, Y. Jiang, M. Zhou, and L.-R. Dai, “Wavenet vocoder with limited training data for voice conversion,” in
1987
Earlier work this paper cites.
I. H. Witten, R. M. Neal, and J. G. Cleary, “Arithmetic coding for data compression,”
1987
Earlier work this paper cites.
D. O’Shaughnessy, “Linear predictive coding,”
1988
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, D. S. Pallett, N. L. Dahlgren, and V. Zue, “TIMIT acoustic-phonetic continuous speech corpus,”
1993
Earlier work this paper cites.
K. Brandenburg and G. Stoll, “ISO/MPEG-1 audio: A generic standard for coding of high-quality digital audio,”
1994
Earlier work this paper cites.
K. R. Rao and J. J. Hwang,
1996
Earlier work this paper cites.
P. Noll, “MPEG digital audio coding,”
1997
Earlier work this paper cites.
R. Salami, C. Laflamme, B. Bessette, and J.-P. Adoul, “ITU-TG. 729 annex a: reduced complexity 8 kb/s CS-ACELP codecfor digital simultaneous voice and data,” IEEE Communications Magazine, vol. 35, no. 9, pp. 56–63, 1997
1997
Cited alongside, same era.
G. Schuller, B. Yu, and D. Huang, “Lossless coding of audio signals using cascaded prediction,” in
2001
Cited alongside, same era.
A. Rix, J. Beerends, M. Hollier, and A. Hekstra, “Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,” in
2001
Cited alongside, same era.
B. Bessette, R. Salami, R. Lefebvre, M. Jelinek, J. Rotola-Pukkila, J. Vainio, H. Mikkola, and K. Jarvinen, “The adaptive multirate wideband speech codec (amr-wb),”
2002
Cited alongside, same era.
G. Recommendation, “722.2:“wideband coding of speech at around 16 kbit/s using adaptive multi-rate wideband (amr-wb)”,” 2003
2016
Later among the works it cites.
2016
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Later among the works it cites.
W. Shi, J. Caballero, F. Huszár, J. Totz, A. P. Aitken, R. Bishop, D. Rueckert, and Z. Wang, “Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network,” in
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2003
Cited alongside, same era.
R. B. ITU-R, “1534-1,“method for the subjective assessment of intermediate quality levels of coding systems (mushra)”,”
2003
Cited alongside, same era.
L. Deng, M. Seltzer, D. Yu, A. Acero, A. Mohamed, and G. Hinton, “Binary coding of speech spectrograms using a deep auto-encoder,” in
2010
Cited alongside, same era.
M. Neuendorf, M. Multrus, N. Rettelbach, G. Fuchs, J. Robilliard, J. Lecomte, S. Wilde, S. Bayer, S. Disch, C. Helmrich
2012
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Cited alongside, same era.
M. Kim and P. Smaragdis, “Bitwise neural networks,” in
2015
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Later among the works it cites.
S. Kankanahalli, “End-to-end optimized speech coding with deep neural networks,”
2017
Later among the works it cites.
E. Agustsson, F. Mentzer, M. Tschannen, L. Cavigelli, R. Timofte, L. Benini, and L. V. Gool, “Soft-to-hard vector quantization for end-to-end learning compressible representations,” in
2017
Later among the works it cites.
N. Johnston, D. Vincent, D. Minnen, M. Covell, S. Singh, T. Chinen, S. Jin Hwang, J. Shor, and G. Toderici, “Improved lossy image compression with priming and spatially adaptive bit rates for recurrent networks,” in
2018
Later among the works it cites.
K. Tan, J. Chen, and D. Wang, “Gated residual networks with dilated convolutions for supervised speech separation,” in
2018
Later among the works it cites.
Y. L. Cristina Garbacea, Aaron van den Oord, “Low bit-rate speech coding with vq-vae and a wavenet decoder,” in
2019
Closest in time.
Y. L. Cristina Garbacea, Aaron van den Oord, “High-quality speech coding with samplernn,” in
2019
Closest in time.