Fetching the paper…
Reading the bibliography…
The objective speech quality assessment is usually conducted by comparing received speech signal with its clean reference, while human beings are capable of evaluating the speech quality without any reference, such as in the mean opinion score (MOS) tests.
J. B. Allen and D. A. Berkley, “Image method for efficiently simulation room-small acoustic,”
1979
Earlier work this paper cites.
ITU-T, “Recommendation p.800: Methods for subjective determination of transmission quality,”
1998
Earlier work this paper cites.
Y. Rubner, C. Tomasi, and L. J. Guibas, “The earth mover’s distance as a metric for image retrieval,” in
2000
Earlier work this paper cites.
ITU-T, “Recommendation p.862: Perceptual evaluation of speech quality (pesq), an objective method for endto-end speech quality assessment of narrowband telephone networks and speech codecs,” 2001
2001
Earlier work this paper cites.
E. Levina and P. Bickel, “The earth mover’s distance is the mallows distance: some insights from statistics,” in
2001
Earlier work this paper cites.
J. L. Myers and A. D. Well, “Research design and statistical analysis (2nd ed.),”
2003
Earlier work this paper cites.
——, “Recommendation p.563: Single-ended method for objective speech quality assessment in narrowband telephony applications,”
2004
Earlier work this paper cites.
W. H. Falk and W.-Y. Chan, “Single-ended speech quality measurement using machine learning methods,”
2006
Earlier work this paper cites.
E. Vincent, R. Gribonval, and C. Fevotte, “Performance measurement in blind audio source separation,”
2006
Earlier work this paper cites.
V. Grancharov, D. Y. Zhao, J. Lindblom, and W. B. Kleijn, “Low-complexity, nonintrusive speech quality assessment,”
2006
Earlier work this paper cites.
L. Ding, Z. Lin, A. Radwan, M. S. El-Hennawey, and R. A. Goubran, “Non-intrusive single-ended speech quality assessment in voip,”
2007
Earlier work this paper cites.
——, “Recommendation p.863: Perceptual objective listening quality assessment: An advanced objective perceptual method for end-to-end listening speech quality evaluation of fixed, mobile, and ip-based networks and speech codecs covering narrowband, wideband, and super-wideband signals,” 2011
2011
Earlier work this paper cites.
M. Narwaria, W. Lin, I. V. McLoughlin, S. Emmanuel, and L.-T. Chia, “Nonintrusive quality assessment of noise suppressed speech with mel-filtered energies and support vector regression,”
2012
Earlier work this paper cites.
V. I. Bogachev and A. V. Kolesnikov, “The mongekantorovich problem: achievements, connections, and perspectives,”
2012
Cited alongside, same era.
R. K. Dubey and A. Kumar, “Non-intrusive speech quality assessment using several combinations of auditory features,”
2013
Cited alongside, same era.
D. Sharma, L. Meredith, J. Lainez, D. Barreda, and P. A. Naylor, “A non-intrusive pesq measure,” in
2014
Cited alongside, same era.
A. R. Avila, B. Cauchi, S. Goetze, S. Doclo, and T. Falk, “Performance comparison of intrusive and non-intrusive instrumental quality measures for enhanced speech,” in
2016
Cited alongside, same era.
M. H. Soni and H. A. Patil, “Novel deep autoencoder features for non-intrusive speech quality assessment,” in
2016
Cited alongside, same era.
2018
Later among the works it cites.
C.-C. Lo, S.-W. Fu, W.-C. Huang, X. Wang, J. Yamagishi, Y. Tsao, and H.-M. Wang, “Mosnet: Deep learning-based objective assessment for voice conversion,” in
2019
Later among the works it cites.
H. Gamper, C. Reddy, R. Cutler, I. Tashev, and J. Gehrke, “Intrusive and non-intrusive perceptual speech quality assessment using a convolutional neural network,” in
2019
Later among the works it cites.
A. Avila, H. Gamper, C. Reddy, R. Cutler, I. Tashev, and J. Gehrke, “Non-intrusive speech quality assessment using neural networks,” in
2019
Later among the works it cites.
X. Dong and D. S. Williamson, “A classification-aided framework for non-intrusive speech quality assessment,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
X. Geng, “Label distribution learning,”
2016
Cited alongside, same era.
2016
Cited alongside, same era.
M. Hakami and W. B. Kleijn, “Machine learning based nonintrusive quality estimation with an augmented feature set,” in
2017
Cited alongside, same era.
C. Spille, S. Ewert, B. Kollmeier, and B. Meyer, “Predicting speech intelligibility with deep neural networks,”
2018
Cited alongside, same era.
S.-W. Fu, Y. Tsao, H.-T. Hwang, and H.-M. Wang, “Quality-net: An end-to-end non-intrusive speech quality assessment model based on blstm,” in
2018
Cited alongside, same era.
A. H. Andersen, J. M. Haan, Z. Tan, and J. Jensen, “Non-intrusive speech intelligibility prediction using convolutional neural networks,”
2018
Cited alongside, same era.
B. Gao, H. Zhou, J. Wu, and X. Geng, “Age estimation using expectation of label distribution learning,” in
2018
Cited alongside, same era.
2019
Later among the works it cites.
F. Bahmaninezhad, J. Wu, R. Gu, S.-X. Zhang, Y. Xu, M. Yu, and D. Yu, “A comprehensive study of speech separation: spectrogram vs waveform separation,”
2019
Later among the works it cites.
Y. Luo and N. Mesgarani, “Conv-tasnet: Surpassing ideal time–frequency magnitude masking for speech separation,”
2019
Later among the works it cites.
J. L. Roux, S. Wisdom, H. Erdogan, and J. R. Hershey, “Sdr - half-baked or well done?” in
2019
Later among the works it cites.
C. K. Reddy, V. Gopal, and R. Cutler, “Dnsmos: A non-intrusive perceptual objective speech quality metric to evaluate noise suppressors,”
2020
Later among the works it cites.
——, “An attention enhanced multi-task model for objective speech assessment in real-world environments,” in
2020
Later among the works it cites.
R. Gu, S.-X. Zhang, Y. Xu, L. Chen, Y. Zou, , and D. Yu, “Multi-modal multi-channel target speech separation,”
2020
Later among the works it cites.