» Articles » PMID: 25411785

On the Importance of the Distance Measures Used to Train and Test Knowledge-based Potentials for Proteins

Overview
Journal PLoS One
Date 2014 Nov 21
PMID 25411785
Citations 1
Authors
Affiliations
Soon will be listed here.
Abstract

Knowledge-based potentials are energy functions derived from the analysis of databases of protein structures and sequences. They can be divided into two classes. Potentials from the first class are based on a direct conversion of the distributions of some geometric properties observed in native protein structures into energy values, while potentials from the second class are trained to mimic quantitatively the geometric differences between incorrectly folded models and native structures. In this paper, we focus on the relationship between energy and geometry when training the second class of knowledge-based potentials. We assume that the difference in energy between a decoy structure and the corresponding native structure is linearly related to the distance between the two structures. We trained two distance-based knowledge-based potentials accordingly, one based on all inter-residue distances (PPD), while the other had the set of all distances filtered to reflect consistency in an ensemble of decoys (PPE). We tested four types of metric to characterize the distance between the decoy and the native structure, two based on extrinsic geometry (RMSD and GTD-TS*), and two based on intrinsic geometry (Q* and MT). The corresponding eight potentials were tested on a large collection of decoy sets. We found that it is usually better to train a potential using an intrinsic distance measure. We also found that PPE outperforms PPD, emphasizing the benefits of capturing consistent information in an ensemble. The relevance of these results for the design of knowledge-based potentials is discussed.

Citing Articles

An In Silico Design of Peptides Targeting the S1/S2 Cleavage Site of the SARS-CoV-2 Spike Protein.

Ho C, Nazarie W, Lee P Viruses. 2023; 15(9).

PMID: 37766336 PMC: 10536081. DOI: 10.3390/v15091930.

References
1.
Cozzetto D, Kryshtafovych A, Tramontano A . Evaluation of CASP8 model quality predictions. Proteins. 2009; 77 Suppl 9:157-66. DOI: 10.1002/prot.22534. View

2.
Rogen P, Koehl P . Extracting knowledge from protein structure geometry. Proteins. 2013; 81(5):841-51. PMC: 3618491. DOI: 10.1002/prot.24242. View

3.
Lindorff-Larsen K, Piana S, Palmo K, Maragakis P, Klepeis J, Dror R . Improved side-chain torsion potentials for the Amber ff99SB protein force field. Proteins. 2010; 78(8):1950-8. PMC: 2970904. DOI: 10.1002/prot.22711. View

4.
Chopra G, Kalisman N, Levitt M . Consistent refinement of submitted models at CASP using a knowledge-based potential. Proteins. 2010; 78(12):2668-78. PMC: 2911515. DOI: 10.1002/prot.22781. View

5.
Guntert P, Mumenthaler C, Wuthrich K . Torsion angle dynamics for NMR structure calculation with the new program DYANA. J Mol Biol. 1997; 273(1):283-98. DOI: 10.1006/jmbi.1997.1284. View