Using support vector machines to distinguish enzymes: Approached by   incorporating wavelet transform

Qiu Jian Ding<sup>*</sup>; Luo San Hua; Huang Jian Hua; Liang Ru Ping

doi:10.1016/j.jtbi.2008.10.026

摘要

The enzymatic attributes of newly found protein sequences are Usually determined either by biochemical analysis of eukaryotic and prokaryotic genomes or by microarray chips. These experimental methods are both time-consuming and costly. With the explosion of protein sequences registered in the databanks, it is highly desirable to develop all automated method to identify whether a given new sequence belongs to enzyme or non-enzyme. The discrete wavelet transform (DWT) and support vector machine (SVM) have been used in this study for distinguishing enzyme structures from non-enzymes. The networks have been trained and tested on two datasets of proteins with different wavelet basis functions, decomposition scales and hydrophobicity data types. Maximum accuracy has been obtained using SVM with a wavelet function of Bior2.4, a decomposition scale j = 5, and Kyte-Doolittle hydrophobicity scales. The results obtained by the self-consistency test, jackknife test and independent dataset test are encouraging, which indicates that the proposed method call be employed as a useful assistant technique for distinguishing enzymes from non-enzymes.

出版日期2009-2-21
单位南昌大学; 萍乡学院

全文

访问全文

收藏分享被引(9) 浏览

更新时间：2019-11-26 06:54

Using support vector machines to distinguish enzymes: Approached by incorporating wavelet transform

摘要

全文

产品服务

站内浏览

服务支持

联系方式

科研之友