PubMed Health⌕ Search

PubMed · 11131970

Simple knowledge-based descriptors to predict protein-ligand interactions. methodology and validation.

Abstract

A new type of shape descriptor is proposed to describe the spatial orientation for non-covalent interactions. It is built from simple, anisotropic Gaussian contributions that are parameterised by 10 adjustable values. The descriptors have been used to fit propensity distributions derived from scatter data stored in the IsoStar database. This database holds composite pictures of possible interaction geometries between a common central group and various interacting moieties, as extracted from small-molecule crystal structures. These distributions can be related to probabilities for the occurrence of certain interaction geometries among different functional groups. A fitting procedure is described that generates the descriptors in a fully automated way. For this purpose, we apply a similarity index that is tailored to the problem, the Split Hodgkin Index. It accounts for the similarity in regions of either high or low propensity in a separate way. Although dependent on the division into these two subregions, the index is robust and performs better than the regular Hodgkin index. The reliability and coverage of the fitted descriptors was assessed using SuperStar. SuperStar usually operates on the raw IsoStar data to calculate propensity distributions, e.g., for a binding site in a protein. For our purpose we modified the code to have it operate on our descriptors instead. This resulted in a substantial reduction in calculation time (factor of five to eight) compared to the original implementation. A validation procedure was performed on a set of 130 protein-ligand complexes, using four representative interacting probes to map the properties of the various binding sites: ammonium nitrogen, alcohol oxygen, carbonyl oxygen, and methyl carbon. The predicted 'hot spots' for the binding of these probes were compared to the actual arrangement of ligand atoms in experimentally determined protein-ligand complexes. Results indicate that the version of SuperStar that applies to our descriptors is capable to predict the above-mentioned atom types in ligands correctly with success rates of 59% and 74%, respectively, for all ligand atoms (regardless of their solvent accessibility), and a subset of solvent-inaccessible ones. If not only exact atom-type matches are counted, but also those that identify ligand atoms of similar physicochemical properties, the prediction rates rise to 75% and 89%. These rates are close to those obtained by the original SuperStar method (being 67% and 82%, respectively, for the prediction of exact matching atom types, and 81% and 91% in the case of predicting similar atom types).

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Nissink JWM, M L Verdonk, G Klebe. 2000. Simple knowledge-based descriptors to predict protein-ligand interactions. methodology and validation.. https://doi.org/10.1023/a%3A1008109717641

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Density functional theory augmented with an empirical dispersion term. Interaction energies and geometries of 80 noncovalent complexes compared with ab initio quantum mechanics calculations.

Standard density functional theory (DFT) is augmented with a damped empirical dispersion term. The damping function is optimized on a small, well balanced set of 22 van der Waals (vdW) complexes and verified on a validation set of 58 vdW complexes. Both sets contain biologically relevant molecules such as nucleic acid bases. Results are in remarkable agreement with reference high-level wave function data based on the CCSD(T) method. The geometries obtained by full gradient optimization are in very good agreement with the best available theoretical reference. In terms of the standard deviation and average errors, results including the empirical dispersion term are clearly superior to all pure density functionals investigated-B-LYP, B3-LYP, PBE, TPSS, TPSSh, and BH-LYP-and even surpass the MP2/cc-pVTZ method. The combination of empirical dispersion with the TPSS functional performs remarkably well. The most critical part of the empirical dispersion approach is the damping function. The damping parameters should be optimized for each density functional/basis set combination separately. To keep the method simple, we optimized mainly a single factor, s(R), scaling globally the vdW radii. For good results, a basis set of at least triple-zeta quality is required and diffuse functions are recommended, since the basis set superposition error seriously deteriorates the results. On average, the dispersion contribution to the interaction energy missing in the DFT functionals examined here is about 15 and 100% for the hydrogen-bonded and stacked complexes considered, respectively.

Hydrogen Bonding↗

Unicorns in the world of chemical bonding models.

The appearance and the significance of heuristically developed bonding models are compared with the phenomenon of unicorns in mythical saga. It is argued that classical bonding models played an essential role for the development of the chemical science providing the language which is spoken in the territory of chemistry. The advent and the further development of quantum chemistry demands some restrictions and boundary conditions for classical chemical bonding models, which will continue to be integral parts of chemistry.

Hydrogen Bonding↗