期刊文献+

基于Gammatone滤波器组的听觉特征提取 被引量:28

Auditory Feature Extraction Based on Gammatone Filter Bank
下载PDF
导出
摘要 目前主流说话人特征参数在噪声环境中的鲁棒性较差。为此,提出一种可用于说话人识别的听觉倒谱特征系数。分析人耳听觉模型的工作机理,采用Gammatone滤波器组代替传统的三角滤波器组模拟人耳耳蜗的听觉模型,用指数压缩代替固定的对数压缩,模拟人耳听觉模型处理信号的非线性特性。在基于高斯混合模型分类器的识别算法下进行仿真实验,结果表明,该听觉特征具有比梅尔频率倒谱系数和线性预测倒谱系数更好的抗噪声能力。 Aiming at the problem that speaker's feature coefficients have poor robustness in noise environment, this paper proposes an auditory cepstral coefficient for speaker recognition. It analyzes the working mechanism of the human auditory model, simulates the auditory model of human ear cochlea by Garnmatone filter banks replaces the traditional triangular filter banks. Based on the nonlinear signal processing capability of human auditory model, exponential compression is used instead of the fixed logarithm compression. Simulation experiment is conducted based on Gaussian Mixed Model(GMM) recognition algorithm. Experimental results show that the auditory feature has better noise robusmess than Mel Frequency Cepstral Coefficient(MFCC) and Linear Prediction Cepstral Coefficient(LPCC).
出处 《计算机工程》 CAS CSCD 2012年第21期168-170,174,共4页 Computer Engineering
关键词 说话人识别 特征提取 Gammatone滤波器 听觉模型 倒谱系数 鲁棒性 speaker recognition feature extraction Gammatone filter auditory model cepstral coefficient robustness
  • 相关文献

参考文献8

二级参考文献173

  • 1李波,王成友,杨聪,蔡宣平,张尔扬.基于语音频谱包络抽取的MFCC算法[J].国防科技大学学报,2004,26(4):42-45. 被引量:4
  • 2陆振波,章新华,胡洪波.水中目标辐射噪声的听觉特征提取[J].系统工程与电子技术,2004,26(12):1801-1803. 被引量:19
  • 3彭圆,王晟,王科俊,李雪耀,林良骥,林正青,王建文.感知线性预测在水下目标分类中的应用研究[J].声学学报,2006,31(2):146-150. 被引量:16
  • 4李朝晖,迟惠生.听觉外周计算模型研究进展[J].声学学报,2006,31(5):449-465. 被引量:22
  • 5Tucker S, Brown G J. Classification of transient sonar sounds using perceptually motivated features [J ]. IEEE Journal of Oceanic Engineering, 2005, 30(3) : 588- 600.
  • 6Parks T W, Weisburn B A. Classifichtion of whale and ice sounds with a cochlear model[J]. ICASSP-92, 1992, 2:481 -484.
  • 7Wan Wanggen. Robust speech recognition based on the secondorder difference cochlear model[C]//Proceedings of 2001 International Symposium on Intelligent multimedia, Video and Speech Processing. Hong Kong: IEEE, 2001:543-546.
  • 8Strope B, Alwan A. A model of dynamic auditory perception and its application to robust word recognition[J ]. IEEE Trans on Speech and Audio Processing, 1997, 5(5) : 451 - 464.
  • 9Moore B C J, Glasberg B R, Baer T. A model for the prediction of thresholds, loudness, and partial loudness [ J ]. J Audio Eng Soc, 1997, 45(4): 224-239.
  • 10Patterson R D, Unoki M, Irino T. Extending the domain of center frequencies for the compressive Gammachirp auditory filter[J]. J Acoust Soc Am, 2003, 114(3) : 1529- 1542.

共引文献56

同被引文献193

引证文献28

二级引证文献81

相关作者

内容加载中请稍等...

相关机构

内容加载中请稍等...

相关主题

内容加载中请稍等...

浏览历史

内容加载中请稍等...
;
使用帮助 返回顶部