跳到主要导航 跳到搜索 跳到主要内容

Novel full reference perceptual quality metric for audio-visual asynchrony

  • Yao Du Wei*
  • , Xiang Xie
  • , Jing Ming Kuang
  • , Xin Lu Han
  • *此作品的通讯作者
  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

A full reference model was proposed to evaluate the perceptual quality of audiovisual asynchrony. A standard synchronization process was used to determine the time difference between audio and video. The mapping between the time difference and the perceptual quality was derived by co-inertia analysis. The co-inertia analysis extracted the most related component from audio and video features, and then formed a mapping for each audiovisual sequence. Audiovisual contents were divided into three categories: clean speech, non speech and mixed speech. The clean speech category was further split into two subcategories. Audio and video features were chosen separately for each category. Subjective test results showed that the proposed model conforms well with subjective results.

源语言英语
页(从-至)182-190
页数9
期刊Tongxin Xuebao/Journal on Communications
33
2
出版状态已出版 - 2月 2012

指纹

探究 'Novel full reference perceptual quality metric for audio-visual asynchrony' 的科研主题。它们共同构成独一无二的指纹。

引用此