Scinovex
article Open AccessTop 10% cited

Can correct protein models be identified?

Protein Science · 2003 · Vol. 12(5) · pp. 1073–1086
Björn WallnerArne Elofsson

Abstract

The ability to separate correct models of protein structures from less correct models is of the greatest importance for protein structure prediction methods. Several studies have examined the ability of different types of energy function to detect the native, or native-like, protein structure from a large set of decoys. In contrast to earlier studies, we examine here the ability to detect models that only show limited structural similarity to the native structure. These correct models are defined by the existence of a fragment that shows significant similarity between this model and the native structure. It has been shown that the existence of such fragments is useful for comparing the performance between different fold recognition methods and that this performance correlates well with performance in fold recognition. We have developed ProQ, a neural-network-based method to predict the quality of a protein model that extracts structural features, such as frequency of atom-atom contacts, and predicts the quality of a model, as measured either by LGscore or MaxSub. We show that ProQ performs at least as well as other measures when identifying the native structure and is better at the detection of correct models. This performance is maintained over several different test sets. ProQ can also be combined with the Pcons fold recognition predictor (Pmodeller) to increase its performance, with the main advantage being the elimination of a few high-scoring incorrect models. Pmodeller was successful in CASP5 and results from the latest LiveBench, LiveBench-6, indicating that Pmodeller has a higher specificity than Pcons alone.

Protein Structure and DynamicsMachine Learning in BioinformaticsEnzyme Structure and FunctionProtein structure predictionSimilarity (geometry)Structural similarityFunction (biology)Protein structureComputer scienceBiological systemSet (abstract data type)Contrast (vision)Artificial intelligence

MeSH terms

Cysteine EndopeptidasesHumansModels, MolecularPeptide FragmentsProtein ConformationProteinsNeural Networks, ComputerProtein Structure, SecondaryCaspases
Citations
709
FWCI
2.96
field-weighted impact
References
90
Percentile
91%
vs. same field & year
Citations per year
Cited by
QMEAN: A comprehensive scoring function for model quality assessment
Proteins Structure Function and Bioinformatics · 2007 · 1,028 citations
The HDOCK server for integrated protein–protein docking
Nature Protocols · 2020 · 1,728 citations
References
Neural networks for pattern recognition
Choice Reviews Online · 1994 · 18,690 citations
A new force field for molecular mechanical simulation of nucleic acids and proteins
Journal of the American Chemical Society · 1984 · 4,632 citations
Ab initio protein structure prediction of CASP III targets using ROSETTA
Proteins Structure Function and Bioinformatics · 1999 · 569 citations
The interpretation of protein structures: Estimation of static accessibility
Journal of Molecular Biology · 1971 · 5,866 citations
Protein Structure Comparison by Alignment of Distance Matrices
Journal of Molecular Biology · 1993 · 3,984 citations
Semianalytical treatment of solvation for molecular mechanics and dynamics
Journal of the American Chemical Society · 1990 · 3,632 citations
Comparative Protein Modelling by Satisfaction of Spatial Restraints
Journal of Molecular Biology · 1993 · 13,125 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.