PubMed Health⌕ Search

PubMed · 10346926

Molecular hashkeys: a novel method for molecular characterization and its application for predicting important pharmaceutical properties of molecules.

Abstract

We define a novel numerical molecular representation, called the molecular hashkey, that captures sufficient information about a molecule to predict pharmaceutically interesting properties directly from three-dimensional molecular structure. The molecular hashkey represents molecular surface properties as a linear array of pairwise surface-based comparisons of the target molecule against a common 'basis-set' of molecules. Hashkey-measured molecular similarity correlates well with direct methods of measuring molecular surface similarity. Using a simple machine-learning technique with the molecular hashkeys, we show that it is possible to accurately predict the octanol-water partition coefficient, log P. Using more sophisticated learning techniques, we show that an accurate model of intestinal absorption for a set of drugs can be constructed using the same hashkeys used in the aforementioned experiments. Once a set of molecular hashkeys is calculated, its use in the training and testing of property-based models is very fast. Further, the required amount of data for model construction is very small. Neural network-based hashkey models trained on data sets as small as 30 molecules yield statistically significant prediction of molecular properties. The lack of a requirement for large data sets lends itself well to the prediction of pharmaceutically relevant molecular parameters for which data generation is expensive and slow. Molecular hashkeys coupled with machine-learning techniques can yield models that predict key pharmacological aspects of biologically important molecules and should therefore be important in the design of effective therapeutics.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

A M Ghuloum, C R Sage, A N Jain. 1999-05-20. Molecular hashkeys: a novel method for molecular characterization and its application for predicting important pharmaceutical properties of molecules.. https://doi.org/10.1021/jm980527a

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Strategies to design pyrazolyl urea derivatives for p38 kinase inhibition: a molecular modeling study.

The p38 protein kinase is a serine-threonine mitogen activated protein kinase, which plays an important role in inflammation and arthritis. A combined study of 3D-QSAR and molecular docking has been undertaken to explore the structural insights of pyrazolyl urea p38 kinase inhibitors. The 3D-QSAR studies involved comparative molecular field analysis (CoMFA) and comparative molecular similarity indices (CoMSIA). The best CoMFA model was derived from the atom fit alignment with a cross-validated r (2 )(q (2)) value of 0.516 and conventional r (2) of 0.950, while the best CoMSIA model yielded a q (2) of 0.455 and r (2) of 0.979 (39 molecules in training set, 9 molecules in test set). The CoMFA and CoMSIA contour maps generated from these models provided inklings about the influence of interactive molecular fields in the space on the activity. GOLD, Sybyl (FlexX) and AutoDock docking protocols were exercised to explore the protein-inhibitor interactions. The integration of 3D-QSAR and molecular docking has proffered essential structural features of pyrazolyl urea inhibitors and also strategies to design new potent analogues with enhanced activity.

Drug Design↗