Abstract
Calcium is one of the most abundant minerals in the human body, playing a critical role in many cellular activities by interacting with different calcium ion (Ca2+)-binding proteins. Therefore, the correct identification of Ca2+-binding residues is essential for protein functional research. In this study, a new method was developed to predict Ca2+-binding residues from the primary sequence without using three-dimensional information. Through statistical analysis, four kinds of feature parameters were extracted from amino acid sequences: the increment of diversity values of amino acid composition, the matrix scoring values of position conservation, the autocross covariance of physicochemical properties, and the center motif. These features served as input for a support vector machine to predict Ca2+-binding residues. This method was tested on four well-established datasets using a five-fold cross-validation. The accuracies and Matthews correlation coefficients were 75.9% and 0.53 (dataset 1), 79.2% and 0.58 (dataset 2), 77.4% and 0.55 (dataset 3), and 79.1% and 0.58 (dataset 4). Comparative results show that the developed method outperforms previous methods. Based on this study, a web server was developed for predicting Ca2+-binding residues from any protein sequence, being publically available at http://202.207.29.245/.
Author supplied keywords
Cite
CITATION STYLE
Jiang, Z., Hu, X. Z., Geriletu, G., Xing, H. R., & Cao, X. Y. (2016). Identification of Ca2+-binding residues of a protein from its primary sequence. Genetics and Molecular Research, 15(2). https://doi.org/10.4238/gmr.15027618
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.