Next Article in Journal
Energy-Efficient PPG-Based Respiratory Rate Estimation Using Spiking Neural Networks
Next Article in Special Issue
A Data Matrix Code Recognition Method Based on L-Shaped Dashed Edge Localization Using Central Prior
Previous Article in Journal
Digital Twin Sensors in Cultural Heritage Ontology Applications
Previous Article in Special Issue
HAtt-Flow: Hierarchical Attention-Flow Mechanism for Group-Activity Scene Graph Generation in Videos
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

Postfilter for Dual Channel Speech Enhancement Using Coherence and Statistical Model-Based Noise Estimation

School of Electrical Engineering and Computer Science, Gwangju Institute of Science and Technology, Gwangju 61005, Republic of Korea
*
Author to whom correspondence should be addressed.
Sensors 2024, 24(12), 3979; https://doi.org/10.3390/s24123979
Submission received: 11 May 2024 / Revised: 18 June 2024 / Accepted: 18 June 2024 / Published: 19 June 2024
(This article belongs to the Special Issue Audio, Image, and Multimodal Sensing Techniques)

Abstract

A multichannel speech enhancement system usually consists of spatial filters such as adaptive beamformers followed by postfilters, which suppress remaining noise. Accurate estimation of the power spectral density (PSD) of the residual noise is crucial for successful noise reduction in the postfilters. In this paper, we propose a postfilter utilizing proposed a posteriori speech presence probability (SPP) and noise PSD estimators, which are based on both the coherence and the statistical models. We model the coherence-based a posteriori SPP as a simple function of the magnitude of coherence between two microphone signals and combine it with a single-channel SPP based on statistical models. The coherence-based estimator for the PSD of the noise remaining in the beamformer output in the presence of speech is derived using the pseudo-coherence considering the effect of the beamformers, which is used to construct the coherence-based noise PSD estimator. Then, the final noise PSD estimator is obtained by combining the coherence-based and statistical model-based noise PSD estimators with the proposed SPP. The spectral gain function is also modified, incorporating the proposed SPP. Experimental results demonstrate that the proposed method led to more accurate noise PSD estimation and perceptual evaluation of speech quality scores in various diffuse noise environments, and did not degrade the speech quality under the presence of directional interference, although the proposed method utilizes the coherence information.
Keywords: noise PSD estimation; coherence; dual channel speech enhancement; postfilter; speech presence probability estimation noise PSD estimation; coherence; dual channel speech enhancement; postfilter; speech presence probability estimation

Share and Cite

MDPI and ACS Style

Cheong, S.; Kim, M.; Shin, J.W. Postfilter for Dual Channel Speech Enhancement Using Coherence and Statistical Model-Based Noise Estimation. Sensors 2024, 24, 3979. https://doi.org/10.3390/s24123979

AMA Style

Cheong S, Kim M, Shin JW. Postfilter for Dual Channel Speech Enhancement Using Coherence and Statistical Model-Based Noise Estimation. Sensors. 2024; 24(12):3979. https://doi.org/10.3390/s24123979

Chicago/Turabian Style

Cheong, Sein, Minseung Kim, and Jong Won Shin. 2024. "Postfilter for Dual Channel Speech Enhancement Using Coherence and Statistical Model-Based Noise Estimation" Sensors 24, no. 12: 3979. https://doi.org/10.3390/s24123979

APA Style

Cheong, S., Kim, M., & Shin, J. W. (2024). Postfilter for Dual Channel Speech Enhancement Using Coherence and Statistical Model-Based Noise Estimation. Sensors, 24(12), 3979. https://doi.org/10.3390/s24123979

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop