Detection of Time Varying Pitch in Tonal Languages: an Approach Based on Ensemble Empirical Mode Decomposition
Hong,Xiao-hua Zhu,Wei-min Su,Run-tong Geng,Xin-long Wang
DOI: https://doi.org/10.1631/jzus.c1100092
2012-01-01
Journal of Zhejiang University SCIENCE C
Abstract:A method based on ensemble empirical mode decomposition (EEMD) is proposed for accurately detecting the time varying pitch of speech in tonal languages. Unlike frame-, event-, or subspace-based pitch detectors, the time varying information of pitch within the short duration, which is of crucial importance in speech processing of tonal languages, can be accurately extracted. The Chinese Linguistic Data Consortium (CLDC) database for Mandarin Chinese was employed as standard speech data for the evaluation of the effectiveness of the method. It is shown that the proposed method provides more accurate and reliable results, particularly in estimating the tones of non-monotonically varying pitches like the third one in Mandarin Chinese. Also, it is shown that the new method has strong resistance to noise disturbance.