Fetching the paper…
Reading the bibliography…
The voice mode of the Opus audio coder can compress wideband speech at bit rates ranging from 6 kb/s to 40 kb/s.
B. S. Atal and S. L. Hanauer, “Speech analysis and synthesis by linear prediction of the speech wave,”
1971
Earlier work this paper cites.
F. Itakura, “Line spectrum representation of linear predictor coefficients of speech signals,”
1975
Earlier work this paper cites.
T. Tremain, “Linear predictive coding systems,” in
1976
Earlier work this paper cites.
J. Makhoul and M. Berouti, “Adaptive noise spectral shaping and entropy coding in predictive coding of speech,”
1979
Earlier work this paper cites.
P. Hedelin, “A tone oriented voice excited vocoder,” in
1981
Earlier work this paper cites.
M. Schroeder and B. Atal, “Code-excited linear prediction (CELP): High-quality speech at very low bit rates,” in
1985
Earlier work this paper cites.
R. J. McAulay and T. F. Quatieri, “Speech analysis/synthesis based on a sinusoidal representation,”
1986
Earlier work this paper cites.
NTT Advanced Technology,
1994
Earlier work this paper cites.
A. McCree, K. Truong, E. George, T. Barnwell, and V. Vishu, “A 2.4 kbit/s MELP coder candidate for the new U.S. federal standard,” in
1996
Earlier work this paper cites.
3GPP, “TS 26.090: Adaptive multi-rate (AMR) speech codec; Transcoding functions,” 2001
2001
Earlier work this paper cites.
B. Bessette, R. Salami, R. Lefebvre, M. Jelinek, J. Rotola-Pukkila, J. Vainio, H. Mikkola, and K. Jarvinen, “The adaptive multirate wideband speech codec (AMR-WB),”
2002
Earlier work this paper cites.
J.-M. Valin, “The Speex codec manual,”
2007
Cited alongside, same era.
J.-M. Valin, K. Vos, and T. B. Terriberry, “Definition of the Opus Audio Codec,” IETF RFC 6716, 2012,
2012
Cited alongside, same era.
K. Vos, K. V. Sørensen, S. S. Jensen, and J.-M. Valin, “Voice coding with Opus,” in
2013
Cited alongside, same era.
J.-M. Valin, G. Maxwell, T. B. Terriberry, and K. Vos, “High-quality, low-delay music coding in the Opus codec,” in
2013
Cited alongside, same era.
2014
Cited alongside, same era.
C. Holmberg, S. Håkansson, and G. Eriksson, “Web real-time communication use cases and requirements,” IETF RFC 7478, Mar. 2015,
J.-M. Valin, K. Vos, and T. B. Terriberry, “Updates to the Opus Audio Codec,” IETF RFC 8251, 2017,
2017
Later among the works it cites.
W. B. Kleijn, F. S. C. Lim, A. Luebs, J. Skoglund, F. Stimberg, Q. Wang, and T. C. Walters, “Wavenet based low rate speech coding,” in
2018
Later among the works it cites.
2018
Later among the works it cites.
Z. Jin, A. Finkelstein, G. J. Mysore, and J. Lu, “FFTNet: A real-time speaker-dependent neural vocoder,” in
2018
Later among the works it cites.
J.-M. Valin, “A hybrid DSP/deep learning approach to real-time full-band speech enhancement,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
A. van den Oord, N. Kalchbrenner, and K. Kavukcuoglu, “Pixel recurrent neural networks,”
2016
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Later among the works it cites.
C. Garbacea, A. v. den Oord, Y. Li, F. S. C. Lim, A. Luebs, O. Vinyals, and T. C. Walters, “Low bit-rate speech coding with VQ-VAE and a WaveNet decoder,” in
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
W3C, “WebRTC 1.0: Real-time communication between browsers,” 2019,
2019
Closest in time.
J.-M. Valin and J. Skoglund, “A real-time wideband neural vocoder at 1.6 kb/s using LPCNet,” in
2019
Closest in time.