PubMed Health⌕ Search

PubMed · 16311814

Multi-space classification for predicting GPCR-ligands.

Abstract

A classification of molecules depends on the descriptor set which is used to represent the compounds, and each descriptor could be regarded as one perception of a molecule. In this study we show that a combination of several classifiers that are grounded on separate descriptor sets can be superior to a single classifier that was built using all available descriptors. The task of predicting ligands of G-protein coupled receptors (GPCR) served as an example application. The perceptron, multilayer neural networks, and radial basis function (RBF) networks were employed for prediction. We developed classifiers with and without descriptor selection. Prediction accuracy was assessed by the area under the receiver operating characteristic (ROC) curve. In the case with descriptor selection both the selection and the rank order of the descriptors depended on the type and topology of the neural networks. We demonstrate that the overall prediction accuracy of the system can be improved by joining neural network classifiers of different type and topology using a "jury network" that is trained to evaluate the predictions from the individual classifiers. Seventy-one percent correct prediction of GPCR ligands was obtained.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Alireza Givehchi, Gisbert Schneider. 2005. Multi-space classification for predicting GPCR-ligands.. https://doi.org/10.1007/s11030-005-6293-4

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Identification of MHC Ligands Through Allele-Guided Isolation Combined With Machine Learning for Improved MHC Assignment Using ARDisplay-I.

The isolation of major histocompatibility complex (MHC) ligands and subsequent analysis by mass spectrometry is considered the gold standard for defining targets for T cell-based immunotherapies. However, as many targets of high tumor specificity are only presented at low abundance on the cell surface of tumor cells, the efficient isolation of these peptides is crucial for their successful detection. Here, we demonstrate how optimizing the MHC ligand isolation strategy, based on both the presenting MHC alleles and the individual peptide level, enhances the identification of specific MHC ligands. This ideally acknowledges not only the hydrophobicity but also the post-translational modifications of the respective MHC ligands. To further improve the identification and characterization of MHC ligands, we developed an MHC class I ligand prediction algorithm (ARDisplay-I) that outperforms current state-of-the-art tools when benchmarked against competitors such as netMHCpan 4.1, MixMHCpred, or MHCflurry. Implementing these strategies can augment the development of T cell receptor-based therapies by improving the identification of novel immunotherapy targets and enriching the resources available in the computational immunology field through a superior MHC presentation prediction algorithm.

Ligands↗

An image-based protein-ligand binding representation learning framework via multi-level flexible dynamics trajectory pre-training.

MOTIVATION: Accurate prediction of protein-ligand binding (PLB) relationships plays a crucial role in drug discovery, which helps identify drugs that modulate the activity of specific targets. Traditional biological assays for measuring PLB relationships are time consuming and costly. In addition, models for predicting PLB relationships have been developed and widely used in drug discovery tasks. However, learning more accurate PLB representations is essential to meet the stringent standards required for drug discovery. RESULTS: We propose an image-based PLB representation learning framework, called ImagePLB, which equips ligand representation learner (LRL) and protein representation learner (PRL) to accept 3D multi-view ligand images and protein graphs as input, respectively, and learns rich interaction information between ligand and protein through a binding representation learner (BRL). Considering the scarcity of protein-ligand pairs, we further propose a multi-level next trajectory prediction (MLNTP) task to pre-train ImagePLB on the 4D flexible dynamics trajectory of 16 972 complexes, including ligand level, protein level, and complex level, to learn information related to trajectories. Besides, by introducing trajectory regularization (TR), we effectively alleviate the problem of high (even almost identical) feature similarity caused by adjacent trajectories. Compared with the current state-of-the-art methods, ImagePLB has achieved competitive improvements on PLB-related prediction tasks, including protein-ligand affinity and efficacy prediction tasks. This study opens the door to the image-based PLB learning paradigm. AVAILABILITY AND IMPLEMENTATION: All data and implementation details of code can be obtained from https://github.com/HongxinXiang/ImagePLB.

Ligands↗

An iterative knowledge-based scoring function to predict protein-ligand interactions: I. Derivation of interaction potentials.

Using a novel iterative method, we have developed a knowledge-based scoring function (ITScore) to predict protein-ligand interactions. The pair potentials for ITScore were derived from a training set of 786 protein-ligand complex structures in the Protein Data Bank. Twenty-six atom types were used based on the atom type category of the SYBYL software. The iterative method circumvents the long-standing reference state problem in the derivation of knowledge-based scoring functions. The basic idea is to improve pair potentials by iteration until they correctly discriminate experimentally determined binding modes from decoy ligand poses for the ligand-protein complexes in the training set. The iterative method is efficient and normally converges within 20 iterative steps. The scoring function based on the derived potentials was tested on a diverse set of 140 protein-ligand complexes for affinity prediction, yielding a high correlation coefficient of 0.74. Because ITScore uses SYBYL-defined atom types, this scoring function is easy to use for molecular files prepared by SYBYL or converted by software such as BABEL.

Ligands↗