Abstract:This paper describes an online algorithm for enhancing monaural noisy speech. Firstly, a novel phase-corrected low-delay gammatone filterbank is derived for signal subband decomposition and resynthesis; the subband signals are then analyzed frame by frame. Secondly, a novel feature named periodicity degree (PD) is proposed to be used for detecting and estimating the fundamental period (P0) in each frame and for estimating the signal-to-noise ratio (SNR) in each frame-subband signal unit. The PD is calculated in each unit as the multiplication of the normalized autocorrelation and the comb filter ratio, and shown to be robust in various low-SNR conditions. Thirdly, the noise energy level in each signal unit is estimated recursively based on the estimated SNR for units with high PD and based on the noisy signal energy level for units with low PD. Then the a priori SNR is estimated using a decision-directed approach with the estimated noise level. Finally, a revised Wiener gain is calculated, smoothed, and applied to each unit; the processed units are summed across subbands and frames to form the enhanced signal. The P0 detection accuracy of the algorithm was evaluated on two corpora and showed comparable performance on one corpus and better performance on the other corpus when compared to a recently published pitch detection algorithm. The speech enhancement effect of the algorithm was evaluated on one corpus with two objective criteria and showed better performance in one highly non-stationary noise and comparable performance in two other noises when compared to a state-of-the-art statistical-model based algorithm.

Model-Based Speech Enhancement in the Modulation Domain.

Phase-Aware Single-Channel Speech Enhancement With Modulation-Domain Kalman Filtering

On Single-Channel Speech Enhancement and On Non-Linear Modulation-Domain Kalman Filtering

Adaptive two-channel speech enhancement algorithm based on the modulation spectrum

Noise Estimation Using Mean Square Cross Prediction Error for Speech Enhancement

A Hybrid Approach for Speech Enhancement Using MoG Model and Neural Network Phoneme Classifier

Clinical evaluation of "veraviewpocs" digital panoramic X-ray system.

Online Monaural Speech Enhancement Based on Periodicity Analysis and A Priori SNR Estimation

Speech Enhancement by Denoising and Dereverberation Using a Generalized Sidelobe Canceller-Based Multichannel Wiener Filter

Speech Enhancement Using Non-Negative Spectrogram Models With Mel-Generalized Cepstral Regularization

Speech Enhancement for Non-Stationary Noise Environments

Speech Enhancement Algorithm Based on Spectral Subtraction

Speech Enhancement Based on Analysis–Synthesis Framework with Improved Parameter Domain Enhancement

Kalman filter-based microphone array signal processing using the equivalent source model

Speech Enhancement Based On Analysis Synthesis Framework With Improved Pitch Estimation And Spectral Envelope Enhancement

A Speech Enhancement Algorithm Based on Computational Auditory Scene Analysis

Error Modeling Via Asymmetric Laplace Distribution for Deep Neural Network Based Single-Channel Speech Enhancement

Single-channel speech enhancement by using psychoacoustical model inspired fusion framework

MP-SENet: A Speech Enhancement Model with Parallel Denoising of Magnitude and Phase Spectra

Attention-based Speech Enhancement Using Human Quality Perception Modelling

Spectral Modeling Using Neural Autoregressive Distribution Estimators for Statistical Parametric Speech Synthesis