A novel short-time spectrum estimation algorithm of speech enhancement based on distortion control is presented to overcome the speech distortion and low speech intelligibility in conventional method. An objective measurement parameter of speech distortion is introduced in the algorithm and a constraint condition is generated according to the parameter. Then
a modified log spectral amplitude estimation algorithm is proposed based on the speech absence probabilityand masking properties of the human auditory system. Finally
the enhanced speech is obtained by relating the constraint condition with the estimation algorithm. The new algorithm gets a better performance and the trade-off between noise reduction and speech distortions is achieved. Tests on objective measurements combined with informal subjective listening show that the proposed algorithm has a better performance of speech intelligibility without any perceptional musicality than other spectral subtraction algorithms under the same level of SNR and noise reduction.
关键词
Keywords
references
BOLL S F. Suppression of acoustic noise in speech using spectral subtraction [J]. IEEE Transactions on Acoustics,Speech and Signal Processing, 1979, 27(2): 113-120.
EPHRAIM Y, MALAH D. Speech enhancement using a minimum mean square error short-time spectral amplitude estimator [J]. IEEE Transactions on Acoustics,Speech and Signal Processing, 1984, 2(6): 1109-1121.
EPHRAIM Y, MALAH D. Speech enhancement using a minimum mean square error log spectral amplitude estimator [J]. IEEE Transactions on Acoustics, Speech, and Signal Processing,1985, 3(2): 443-445.
ZHAO Xiaoqun, HUANG Xiaoshan. Improved speech enhancement based on spectral subtraction and auditory masking effect [J]. Journal of Communication, 2008, 29(9): 73-80.
赵力. 语音信号处理[M]第2版. 北京: 机械工业出版社, 2009: 31-32.
MA Jianfen, LOIZOU P C. SNR loss: a new objective measure for predicting the intelligibility of noise-suppressed speech[EB/OL].(2009-12-02)[2010-10-24]. http:∥dx.doi.org/10.1016/j.specom.2010.10.005.
MA Jianfen, HU Yi, LOIZOU P C. Objective measures for predicting speech intelligibility in noisy conditions based on new band-importance functions [J]. Acoust Soc Amer, 2009, 125(5): 3387-3405.
LOIZOU P C, KIM P. Reasons why current speech-enhancement algorithms do not improve speech intelligibility and suggested solutions [J]. IEEE Trans Audio Speech Lang Process, 2011, 19(1): 47-56.
MARTIN R. Noise power spectral density estimation based on optimal smoothing and minimum statistics [J]. IEEE Trans on Acoustics,Speech,and Signal Processing, 2001, 9(5): 504-512.
PU Fanliang, WANG Weiming, DAI Qijun,et a1. Optimizing speech enhancement based on noise marked probability [J]. Journal of Electronics and Information Technology, 2005, 27(5): 753-756.
JOHSTOM J D. Transform coding of audio signals using perceptual noise criteria [J]. IEEE Journal on Selected Areas in Communications, 1988, 6(2): 314-323.
PORTER J E, BOLL F S. Optimal estimators for spectral restoration of speech [C]∥Proceedings of International Conference on Acoustics, Speech, and Signal Processing. Piscataway, NJ, USA: IEEE, 1984: 53-56.