跳到主要导航 跳到搜索 跳到主要内容

DNA binding protein identification by combining pseudo amino acid composition and profile-based protein representation

  • Bin Liu*
  • , Shanyi Wang
  • , Xiaolong Wang
  • *此作品的通讯作者
  • Harbin Institute of Technology
  • Harbin Institute of Technology Shenzhen

科研成果: 期刊稿件文章同行评审

摘要

DNA-binding proteins play an important role in most cellular processes. Therefore, it is necessary to develop an efficient predictor for identifying DNA-binding proteins only based on the sequence information of proteins. The bottleneck for constructing a useful predictor is to find suitable features capturing the characteristics of DNA binding proteins. We applied PseAAC to DNA binding protein identification, and PseAAC was further improved by incorporating the evolutionary information by using profile-based protein representation. Finally, Combined with Support Vector Machines (SVMs), a predictor called iDNAPro-PseAAC was proposed. Experimental results on an updated benchmark dataset showed that iDNAPro-PseAAC outperformed some state-of-the-art approaches, and it can achieve stable performance on an independent dataset. By using an ensemble learning approach to incorporate more negative samples (non-DNA binding proteins) in the training process, the performance of iDNAPro-PseAAC was further improved. The web server of iDNAPro-PseAAC is available at http://bioinformatics.hitsz.edu.cn/iDNAPro-PseAAC/.

源语言英语
文章编号15479
期刊Scientific Reports
5
DOI
出版状态已出版 - 20 10月 2015
已对外发布

指纹

探究 'DNA binding protein identification by combining pseudo amino acid composition and profile-based protein representation' 的科研主题。它们共同构成独一无二的指纹。

引用此