CORC  > 厦门大学  > 信息技术-已发表论文
nDNA-prot: Identification of DNA-binding proteins based on unbalanced classification
Song, Li ; Li, Dapeng ; Zeng, Xiangxiang ; Wu, Yunfeng ; Guo, Li ; Zou, Quan ; Ceng XX(曾湘祥) ; Wu YF(吴云峰) ; Zou Q(邹权)
刊名http://dx.doi.org/10.1186/1471-2105-15-298
2014
关键词Bioinformatics Classification (of information) DNA DNA sequences Feature extraction Statistical tests
英文摘要Background: DNA-binding proteins are vital for the study of cellular processes. In recent genome engineering studies, the identification of proteins with certain functions has become increasingly important and needs to be performed rapidly and efficiently. In previous years, several approaches have been developed to improve the identification of DNA-binding proteins. However, the currently available resources are insufficient to accurately identify these proteins. Because of this, the previous research has been limited by the relatively unbalanced accuracy rate and the low identification success of the current methods.Results: In this paper, we explored the practicality of modelling DNA binding identification and simultaneously employed an ensemble classifier, and a new predictor (nDNA-Prot) was designed. The presented framework is comprised of two stages: a 188-dimension feature extraction method to obtain the protein structure and an ensemble classifier designated as imDC. Experiments using different datasets showed that our method is more successful than the traditional methods in identifying DNA-binding proteins. The identification was conducted using a feature that selected the minimum Redundancy and Maximum Relevance (mRMR). An accuracy rate of 95.80% and an Area Under the Curve (AUC) value of 0.986 were obtained in a cross validation. A test dataset was tested in our method and resulted in an 86% accuracy, versus a 76% using iDNA-Prot and a 68% accuracy using DNA-Prot.Conclusions: Our method can help to accurately identify DNA-binding proteins, and the web server is accessible at http://datamining.xmu.edu.cn/~songli/nDNA. In addition, we also predicted possible DNA-binding protein sequences in all of the sequences from the UniProtKB/Swiss-Prot database. ? 2014 Song et al.; licensee BioMed Central Ltd.
语种英语
出版者BioMed Central Ltd.
内容类型期刊论文
源URL[http://dspace.xmu.edu.cn/handle/2288/92909]  
专题信息技术-已发表论文
推荐引用方式
GB/T 7714
Song, Li,Li, Dapeng,Zeng, Xiangxiang,et al. nDNA-prot: Identification of DNA-binding proteins based on unbalanced classification[J]. http://dx.doi.org/10.1186/1471-2105-15-298,2014.
APA Song, Li.,Li, Dapeng.,Zeng, Xiangxiang.,Wu, Yunfeng.,Guo, Li.,...&邹权.(2014).nDNA-prot: Identification of DNA-binding proteins based on unbalanced classification.http://dx.doi.org/10.1186/1471-2105-15-298.
MLA Song, Li,et al."nDNA-prot: Identification of DNA-binding proteins based on unbalanced classification".http://dx.doi.org/10.1186/1471-2105-15-298 (2014).
个性服务
查看访问统计
相关权益政策
暂无数据
收藏/分享
所有评论 (0)
暂无评论
 

除非特别说明,本系统中所有内容都受版权保护,并保留所有权利。


©版权所有 ©2017 CSpace - Powered by CSpace