摘要
A hidden Markov model (HMM) is used as the statistical model for the acoustic parameters in statistical speech synthesis. This paper presents a quantization method for large scale compression of the HMM based on the acoustic space distance in the speech models. The quality loss caused by the quantization is reduced by an optimal iteration procedure that optimizes the vector quantization codebook. Objective and subjective evaluations show that 90% of the scores indicate no significant reduction of the speech quality with a compression ratio of 0.06 using the proposed spectrum model compression method.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 1196-1200 |
| 页数 | 5 |
| 期刊 | Qinghua Daxue Xuebao/Journal of Tsinghua University |
| 卷 | 51 |
| 期 | 9 |
| 出版状态 | 已出版 - 9月 2011 |
学术指纹
探究 'Large scale compression of HMM for statistical speech synthesis' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver