Chinese Journal of Acoustics----Institute Of Acoustics Chinese Academy Of Sciences

International Cooperation

Education & Training

Societies & Publications

Chinese Journal of Acoustics

Location:Home>Chinese Journal of Acoustics

Blind speech source separation via nonlinear time-frequency masking (2008 No.3)

Author：

ArticleSource：

Update time：

2024/07/24

Viewed：

Text Size: A A A

XU Shun　 CHEN Shaorong　 LIU Yulin

(DSP Lab., Chongqing Communication College Chongqiong 400035)

Received Jun.25, 2007

Revised Oct.11, 2007

Abstract Aim at the underdetermined convolutive mixture model, a blind speech source separation method based on nonlinear time-frequency masking was proposed, where the approximate W-disjoint orthogonality (W-Do) property among independent speech signals in time-frequency domain is utilized. In this method, the observation mixture signal from multi-microphones is normalized to be independent of frequency in the time-frequency domain at first, then the dynamic clustering algorithm is adopted to obtain the active source information in each time-frequency slot, a nonlinear function via deflection angle from the cluster center is selected for time-frequency masking, finally the blind separation of mixture speech signals can be achieved by inverse STFT (short-time Fourier transformation). This method can not only solve the problem of frequency permutation which may be met in most classic frequency-domain blind separation techniques, but also suppress the spatial direction diffusion of the separation matrix. The simulation results demonstrate that the proposed separation method is better than the typical BLUES method, the signal-noise-ratio gain (SNRG) increases 1.58dB averagely.

PACS numbers: 43.60, 43.70

Copyright ？ 1996 - 2020 Institute of Acoustics, Chinese Academy of Sciences
No. 21 North 4th Ring Road, Haidian District, 100190 Beijing, China
E-mail: ioa@mail.ioa.ac.cn