PubMed Health⌕ Search

Biomedical subjects

Paul E Blower

Publications and source records attributed to Paul E Blower.

6 recordsLinked to original sources

Comparison of methods for sequential screening of large compound sets.

Sequential screening is an iterative procedure that can greatly increase hit rates over random screening or non-iterative procedures. We studied the effects of three factors on enrichment rates: the method used to rank compounds, the molecular descriptor set and the selection of initial training set. The primary factor influencing recovery rates was the method of selecting the initial training set. Rates for recovering active compounds were substantially lower with the diverse training sets than they were with training sets selected by other methods. Because structure-activity information is incrementally enhanced in intermediate training sets, sequential screening provides significant improvement in the average rate of recovery of active compounds when compared with non-iterative selection procedures.

Chemistry, Pharmaceutical↗

Decision tree methods in pharmaceutical research.

Decision trees are among the most popular of the new statistical learning methods being used in the pharmaceutical industry for predicting quantitative structure-activity relationships. This article reviews applications of decision trees in drug discovery research and extensions to the basic algorithm using hybrid or ensemble methods that improve prediction accuracy.

Chemistry, Pharmaceutical↗

Building predictive models for protein tyrosine phosphatase 1B inhibitors based on discriminating structural features by reassembling medicinal chemistry building blocks.

A new approach to predicting the biological activity of small molecule pharmaceutics is demonstrated. Structural features of medicinal chemistry building blocks are used as 2-D molecular descriptors. These descriptors include predefined structural features and macrostructures obtained from a supervised process in which features in the core library are reassembled to provide larger features that strongly differentiate the desired biological response variable. Chemical features derived in this manner can serve as predictor variables for diverse modeling algorithms, and application using partial least squares techniques is demonstrated here. Models are presented for inhibition by benzofuran and benzothiophene biphenyl analogues of protein tyrosine phosphatase 1B (PTP1B), a target for insulin-resistant disease states. Results are compared to models for PTP1B inhibitors available in the literature based on CoMFA-related techniques and 3-D molecular descriptors.

Models, Molecular↗

Systematic analysis of large screening sets in drug discovery.

Each year large pharmaceutical companies produce massive amounts of primary screening data for lead discovery. To make better use of the vast amount of information in pharmaceutical databases, companies have begun to scrutinize the lead generation stage to ensure that more and better qualified lead series enter the downstream optimization and development stages. This article describes computational techniques for end to end analysis of large drug discovery screening sets. The analysis proceeds in three stages: In stage 1 the initial screening set is filtered to remove compounds that are unsuitable as lead compounds. In stage 2 local structural neighborhoods around active compound classes are identified, including similar but inactive compounds. In stage 3 the structure-activity relationships within local structural neighborhoods are analyzed. These processes are illustrated by analyzing two large, publicly available databases.

Algorithms↗

Finding discriminating structural features by reassembling common building blocks.

We present a new method for constructing discriminating substructures by reassembling common medicinal chemistry building blocks. The algorithm can be parametrized to meet differing objectives: (1) to build features that discriminate for biological activity in a local structural neighborhood, (2) to build scaffolds for R-group analysis, (3) to construct cluster signatures that discriminate for membership in the cluster and provide a graphical representation for its members, and (4) to identify substructures that characterize major classes in a heterogeneous compound set. We illustrated the results of the algorithm on a literature dataset is of 118 compounds with in vitro inhibition data against recombinant human protein tyrosine phosphatase 1B (PTP-1B).

Algorithms↗

Multiscale and Bayesian approaches to data analysis in genomics high-throughput screening.

Tremendous amounts of data are produced by high-throughput screening methods currently employed in drug discovery and product development. A typical cDNA microarray or oligonucleotide-based gene chip experiment easily generates over 10,000 data points for each array or chip. The challenge of inferring meaningful information is formidable given the size and number of these datasets. This paper reviews the current status of statistical tools available for gene expression analysis, with emphasis on Bayesian approaches and multiscale wavelet filtering. Fundamental concepts of Bayesian and multiscale modeling are discussed from the perspective of their potential to address important issues related to the analysis of gene expression data, such as the fact that genomic data often have non-Gaussian distributions and feature localization and multiple scales in both frequency and measurement dimension. Recent publications in these areas are reviewed. Wavelet filtering and the advantages of multiscale methods are demonstrated by application to publicly available gene expression data from the National Cancer Institute (NCI). Multiscale methods, including multiscale principal component analysis (MSPCA), are applied to extract gene subsets and to visualize data in multidimensions for comparisons. Similarity in cell lines and gene selection are effectively visualized and quantitatively compared.

Animals↗