Reasons Why Current Speech-enhancement Algorithms Do Not Improve Speech Intelligibility and Suggested Solutions
Overview
Authors
Affiliations
Existing speech enhancement algorithms can improve speech quality but not speech intelligibility, and the reasons for that are unclear. In the present paper, we present a theoretical framework that can be used to analyze potential factors that can influence the intelligibility of processed speech. More specifically, this framework focuses on the fine-grain analysis of the distortions introduced by speech enhancement algorithms. It is hypothesized that if these distortions are properly controlled, then large gains in intelligibility can be achieved. To test this hypothesis, intelligibility tests are conducted with human listeners in which we present processed speech with controlled speech distortions. The aim of these tests is to assess the perceptual effect of the various distortions that can be introduced by speech enhancement algorithms on speech intelligibility. Results with three different enhancement algorithms indicated that certain distortions are more detrimental to speech intelligibility degradation than others. When these distortions were properly controlled, however, large gains in intelligibility were obtained by human listeners, even by spectral-subtractive algorithms which are known to degrade speech quality and intelligibility.
Dong R, Liu P, Tian X, Wang Y, Chen Y, Zhang J Front Neurosci. 2024; 18:1407775.
PMID: 39108313 PMC: 11301946. DOI: 10.3389/fnins.2024.1407775.
Zheng C, Zhang H, Liu W, Luo X, Li A, Li X Trends Hear. 2023; 27:23312165231209913.
PMID: 37956661 PMC: 10658184. DOI: 10.1177/23312165231209913.
Reinten I, de Ronde-Brons I, Houben R, Dreschler W Trends Hear. 2023; 27:23312165231192304.
PMID: 37525630 PMC: 10395179. DOI: 10.1177/23312165231192304.
Enhancement of speech-in-noise comprehension through vibrotactile stimulation at the syllabic rate.
Guilleminot P, Reichenbach T Proc Natl Acad Sci U S A. 2022; 119(13):e2117000119.
PMID: 35312362 PMC: 9060510. DOI: 10.1073/pnas.2117000119.
Kang Y, Zheng N, Meng Q Front Med (Lausanne). 2021; 8:740123.
PMID: 34820392 PMC: 8606413. DOI: 10.3389/fmed.2021.740123.