Striped sheets and protein contact prediction

Authors:
Robert M. Maccallum
Affiliations:
Stockholm Bioinformatics Center, Stockholm University, 106 91 Stockholm, Sweden
Venue:
Bioinformatics
Year:
2004

Citing 0
Cited 9

Coordination number prediction using learning classifier systems: performance and interpretability

Proceedings of the 8th annual conference on Genetic and evolutionary computation
Automated alphabet reduction method with evolutionary algorithms for protein structure prediction

Proceedings of the 9th annual conference on Genetic and evolutionary computation
Predicting contact map using Radial Basis Function Neural Network with Conformational Energy Function

International Journal of Bioinformatics Research and Applications
Improving the prediction of helix-residue contacts in all-alpha proteins

NN'08 Proceedings of the 9th WSEAS International Conference on Neural Networks
Protein fold recognition based upon the amino acid occurrence

PRIB'07 Proceedings of the 2nd IAPR international conference on Pattern recognition in bioinformatics
Prediction of inter-residue contact clusters from hydrophobic cores

International Journal of Data Mining and Bioinformatics
Bayesian Models and Algorithms for Protein β-Sheet Prediction

IEEE/ACM Transactions on Computational Biology and Bioinformatics (TCBB)
Ranking Beta Sheet Topologies with Applications to Protein Structure Prediction

Journal of Mathematical Modelling and Algorithms
From HP lattice models to real proteins: coordination number prediction using learning classifier systems

EuroGP'06 Proceedings of the 2006 international conference on Applications of Evolutionary Computing

Quantified Score

Hi-index	3.84

Visualization

Abstract

Motivation: Current approaches to contact map prediction in proteins have focused on amino acid conservation and patterns of mutation at sequentially distant positions. This sequence information is poorly understood and very little progress has been made in this area during recent years. Results: In this study, an observation of 'striped' sequence patterns across β-sheets prompted the development of a new type of contact map predictor. Computer program code was evolved with an evolutionary algorithm (genetic programming) to select residues and residue pairs likely to make contacts based solely on local sequence patterns extracted with the help of self-organizing maps. The mean prediction accuracy is 27% on a validation set of 156 domains up to 400 residues in length, where contacts are separated by at least 8 residues and length/10 pairs are predicted. The retrospective accuracy on a set of 15 CASP5 targets is 27% and 14% for length/10 and length/2 predicted pairs, respectively (both using a minimum residue separation of 24). This compares favourably to the equivalent 21% and 13% obtained for the best automated contact prediction methods at CASP5. The results suggest that protein architectures impose regularities in local sequence environments. Other sources of information, such as correlated/compensatory mutations, may further improve accuracy. Availability: A web-based prediction service is available at http://www.sbc.su.se/~maccallr/contactmaps