重庆大学通信工程学院,重庆,400030
网络首发:2011-08-10,
纸质出版:2011
移动端阅览
刘晓明, 班超帆, 冯晓荣. 失真控制下的短时谱估计语音增强算法[J]. 西安交通大学学报, 2011,45(8):78-84.
A Short-Time Spectrum Estimation Algorithm of Speech Enhancement under the Distortion Control[J]. 2011, 45(8): 78-84.
针对传统谱估计增强算法易产生语音畸变、导致语音清晰度低的问题
提出了一种失真控制下的短时谱估计语音增强的新算法. 该算法首先引入语音畸变的客观度量参数
并根据这一参数得到抑制语音畸变的约束条件
然后结合人耳听觉掩蔽特性和无语音概率参数
修正最小均方误差对数谱估计函数
最后联立约束条件和估计函数
得到增强后的语音
从而实现了在噪声抑制和语音畸变之间的折中
改善了语音增强的效果. 主观试听和客观测试结果均表明
与其他谱减法相比
在相同的信噪比和去噪度条件下
新算法的语音畸变度最小且几乎察觉不到音乐噪声.
A novel short-time spectrum estimation algorithm of speech enhancement based on distortion control is presented to overcome the speech distortion and low speech intelligibility in conventional method. An objective measurement parameter of speech distortion is introduced in the algorithm and a constraint condition is generated according to the parameter. Then
a modified log spectral amplitude estimation algorithm is proposed based on the speech absence probabilityand masking properties of the human auditory system. Finally
the enhanced speech is obtained by relating the constraint condition with the estimation algorithm. The new algorithm gets a better performance and the trade-off between noise reduction and speech distortions is achieved. Tests on objective measurements combined with informal subjective listening show that the proposed algorithm has a better performance of speech intelligibility without any perceptional musicality than other spectral subtraction algorithms under the same level of SNR and noise reduction.
BOLL S F. Suppression of acoustic noise in speech using spectral subtraction [J]. IEEE Transactions on Acoustics,Speech and Signal Processing, 1979, 27(2): 113-120.
EPHRAIM Y, MALAH D. Speech enhancement using a minimum mean square error short-time spectral amplitude estimator [J]. IEEE Transactions on Acoustics,Speech and Signal Processing, 1984, 2(6): 1109-1121.
EPHRAIM Y, MALAH D. Speech enhancement using a minimum mean square error log spectral amplitude estimator [J]. IEEE Transactions on Acoustics, Speech, and Signal Processing,1985, 3(2): 443-445.
赵晓群,黄小珊.改进的基于人耳掩蔽效应谱减语音增强算法 [J]. 通信学报,2008, 29(9):73-80.
ZHAO Xiaoqun, HUANG Xiaoshan. Improved speech enhancement based on spectral subtraction and auditory masking effect [J]. Journal of Communication, 2008, 29(9): 73-80.
赵力. 语音信号处理[M]第2版. 北京: 机械工业出版社, 2009: 31-32.
MA Jianfen, LOIZOU P C. SNR loss: a new objective measure for predicting the intelligibility of noise-suppressed speech[EB/OL].(2009-12-02)[2010-10-24]. http:∥dx.doi.org/10.1016/j.specom.2010.10.005.
MA Jianfen, HU Yi, LOIZOU P C. Objective measures for predicting speech intelligibility in noisy conditions based on new band-importance functions [J]. Acoust Soc Amer, 2009, 125(5): 3387-3405.
LOIZOU P C, KIM P. Reasons why current speech-enhancement algorithms do not improve speech intelligibility and suggested solutions [J]. IEEE Trans Audio Speech Lang Process, 2011, 19(1): 47-56.
MARTIN R. Noise power spectral density estimation based on optimal smoothing and minimum statistics [J]. IEEE Trans on Acoustics,Speech,and Signal Processing, 2001, 9(5): 504-512.
朴凡亮, 王为民, 戴启军,等. 基于噪声被掩蔽概率的优化语音增强方法[J].电子与信息学报, 2005, 27(5): 753-756.
PU Fanliang, WANG Weiming, DAI Qijun,et a1. Optimizing speech enhancement based on noise marked probability [J]. Journal of Electronics and Information Technology, 2005, 27(5): 753-756.
JOHSTOM J D. Transform coding of audio signals using perceptual noise criteria [J]. IEEE Journal on Selected Areas in Communications, 1988, 6(2): 314-323.
PORTER J E, BOLL F S. Optimal estimators for spectral restoration of speech [C]∥Proceedings of International Conference on Acoustics, Speech, and Signal Processing. Piscataway, NJ, USA: IEEE, 1984: 53-56.
【本刊相关文献链接】
随机矩阵变换的音频置乱算法. 西安交通大学学报,2010,44(4):13-17.
分组网络环境下的实时语音质量客观评价. 西安交通大学学报,2006,40(8):936-939.
基于心理声学模型的高性能语音质量评价算法. 西安交通大学学报,2006,40(4):437-440.
基于模糊多类支持向量机的语音质量客观评价. 西安交通大学学报,2006,40(2):199-202.
语音信号中相位信息的听觉感知研究.西安交通大学学报,2003,37(12):1288-1291.
一种基于神经网络的小波域音频水印算法. 西安交通大学学报,2003,37(4):355-358.
基于局部余弦变换的2.4 kb/s低比特率语音编码. 西安交通大学学报,2003,37(4):388-391.
声带振动功能模式识别的研究. 西安交通大学学报,2002,36(12):1258-1261.
基于统计模型实现语音信号有声/无声检测的研究. 西安交通大学学报,2002,36(8):839-842.
Internet 环境下服务可控的网络语音体系结构研究. 西安交通大学学报,2002,36(4):402-405.
含噪语音信号中噪声参数的一种估计方法. 西安交通大学学报,2001,35(10): 1096-1097.
抽取音频数据特征的快速离散余弦变换方法. 西安交通大学学报,2001,35(8):854-857.
0
浏览量
4
下载量
1
CSCD
关联资源
相关文章
相关作者
相关机构
京公网安备11010802024621