Mining the Protein Data Bank to improve prediction of changes in protein-protein binding.

Flores SC, Alexiou A, Glaros A

PLoS ONE 16 (11) e0257614 [2021-11-02; online 2021-11-02]

Predicting the effect of mutations on protein-protein interactions is important for relating structure to function, as well as for in silico affinity maturation. The effect of mutations on protein-protein binding energy (ΔΔG) can be predicted by a variety of atomic simulation methods involving full or limited flexibility, and explicit or implicit solvent. Methods which consider only limited flexibility are naturally more economical, and many of them are quite accurate, however results are dependent on the atomic coordinate set used. In this work we perform a sequence and structure based search of the Protein Data Bank to find additional coordinate sets and repeat the calculation on each. The method increases precision and Positive Predictive Value, and decreases Root Mean Square Error, compared to using single structures. Given the ongoing growth of near-redundant structures in the Protein Data Bank, our method will only increase in applicability and accuracy.

PubMed 34727109

DOI 10.1371/journal.pone.0257614

Crossref 10.1371/journal.pone.0257614

pmc: PMC8562805
pii: PONE-D-21-15338


Publications 9.5.1