摘要

We present a novel classification method for splice sites prediction using support vector machine (SVM). The method first represents input sequences by sequence-based features, including the information of the distribution of tri-nucleotides and the conserved features surrounding the splice sites characterized by Markov model. An F-score based feature selection method is then used to select informative features to improve the performance. Finally, SVM is employed to classify the splice sites with the selected features. Experimental results show that this method improves splice site prediction accuracy and performs better than the existing methods such as MM1-SVM, Reduced MM1-SVM and some other methods.

全文