跳到主要导航 跳到搜索 跳到主要内容

Acoustic features based on auditory model and adaptive fractional Fourier transform for speech recognition

  • Hui Yin*
  • , Xiang Xie
  • , Jingming Kuang
  • *此作品的通讯作者
  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

It is well known that auditory system of human beings has excellent performance with which automatic speech recognition (ASR) systems can't match, and fractional Fourier transform (FrFT) has unique advantages in nonstationary signal processing. In this paper, the Gammatone filterbank is applied to speech signals for front-end temporal filtering, and then acoustic features of the output subband signals are extracted based on fractional Fourier transform. The transform order is critical for FrFT. An order adaptation method based on the instantaneous frequency is proposed, and its performance is compared with the method based on ambiguity function. ASR experiments are conducted on clean and noisy Mandarin digits, and the results show that the proposed features achieve significantly higher recognition rate than the MFCC baseline, and the order adaptation method based on instantaneous frequency has much lower complexity than that based on ambiguity function. Further more, the FrFT-based features achieve the highest recognition rate using the proposed order adaptation method.

源语言英语
页(从-至)97-103
页数7
期刊Shengxue Xuebao/Acta Acustica
37
1
出版状态已出版 - 1月 2012

学术指纹

探究 'Acoustic features based on auditory model and adaptive fractional Fourier transform for speech recognition' 的科研主题。它们共同构成独一无二的学术指纹。

引用此