On intrusive speech quality measures and a global SNR based metric

Chao Pan,Jingdong Chen,Jacob Benesty

doi:10.1016/j.specom.2024.103044

Abstract

Measuring the quality of noisy speech signals has been an increasingly important problem in the field of speech processing as more and more speech-communication and human-machine-interface systems are deployed in practical applications. In this paper, we study four widely used classical performance measures: signal-to-distortion ratio (SDR), short-time objective intelligibility (STOI), signal-to-noise ratio (SNR), and perceptual evaluation of speech quality (PESQ). Through analyzing these performance measures under the same framework and identifying the relationship between their core parameters, we convert these measures into the corresponding equivalent SNRs. This conversion enables not only some new insights into different quality measures but also a way to combine these measures into a new metric. In the derivation of the equivalent SNRs, we introduce the widely used masking technique into the computation of correlation coefficients, which is subsequently used to analyze STOI. Furthermore, we propose an attention method to compute the core parameters of PESQ, and also an empirical formula to project the equivalent SNRs into PESQ scores. Experiments are carried out and the results justifies the properties of the derived quality measures.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

On intrusive speech quality measures and a global SNR based metric

Abstract

Talk to us

Similar Papers

More From: Speech Communication

Lead the way for us

Similar Papers

New research on monaural speech segregation based on quality assessment
Xiaoping Xie ... Fei Ding
Computer Speech & Language | VOL. 85
Xiaoping Xie, et. al.Xiaoping Xie ... Fei Ding
05 Dec 2023
Computer Speech & Language | VOL. 85

A Methodology for Improving PESQ accuracy for Chinese Speech
Fong Chong ... Ian Mcloughlin
-
Fong Chong, et. al.Fong Chong ... Ian Mcloughlin
01 Nov 2005
01 Nov 2005

Towards real-world objective speech quality and intelligibility assessment using speech-enhancement residuals and convolutional long short-term memory networks.
Xuan Dong ... Donald S Williamson
The Journal of the Acoustical Society of America | VOL. 148
Xuan Dong, et. al.Xuan Dong ... Donald S Williamson
01 Nov 2020
The Journal of the Acoustical Society of America | VOL. 148

Analysis of statistical estimators and neural network approaches for speech enhancement
Ravi Kumar Kandagatla ... Rajeswari K
SciEnggJ | VOL. 17
Ravi Kumar Kandagatla, et. al.Ravi Kumar Kandagatla ... Rajeswari K
13 Feb 2024
SciEnggJ | VOL. 17

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

On intrusive speech quality measures and a global SNR based metric

Abstract

Talk to us

Similar Papers

More From: Speech Communication