Carbohydrate-binding proteins are proteins that can interact with sugar chains but do not modify them. They are involved in many physiological functions, and we have developed a method for predicting them from their amino acid sequences. Our method is based on support vector machines (SVMs). We first clarified the definition of carbohydrate-binding proteins and then constructed positive and negative datasets with which the SVMs were trained. By applying the leave-one-out test to these datasets, our method delivered 0.92 of the area under the receiver operating characteristic (ROC) curve. We also examined two amino acid grouping methods that enable effective learning of sequence patterns and evaluated the performance of these methods. When we applied our method in combination with the homology-based prediction method to the annotated human genome database, H-invDB, we found that the true positive rate of prediction was improved.

Download full-text PDF

Source
http://www.ncbi.nlm.nih.gov/pmc/articles/PMC2948896PMC
http://dx.doi.org/10.1155/2010/289301DOI Listing

Publication Analysis

Top Keywords

carbohydrate-binding proteins
12
support vector
8
vector machines
8
amino acid
8
method
5
prediction carbohydrate-binding
4
proteins
4
proteins sequences
4
sequences support
4
machines carbohydrate-binding
4

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!