PubMed Health⌕ Search

PubMed · 16023562

Logical analysis of diffuse large B-cell lymphomas.

Abstract

OBJECTIVE: The goal of this study is to re-examine the oligonucleotide microarray dataset of Shipp et al., which contains the intensity levels of 6817 genes of 58 patients with diffuse large B-cell lymphoma (DLBCL) and 19 with follicular lymphoma (FL), by means of the combinatorics, optimisation, and logic-based methodology of logical analysis of data (LAD). The motivations for this new analysis included the previously demonstrated capabilities of LAD and its expected potential (1) to identify different informative genes than those discovered by conventional statistical methods, (2) to identify combinations of gene expression levels capable of characterizing different types of lymphoma, and (3) to assemble collections of such combinations that if considered jointly are capable of accurately distinguishing different types of lymphoma. METHODS AND MATERIALS: The central concept of LAD is a pattern or combinatorial biomarker, a concept that resembles a rule as used in decision tree methods. LAD is able to exhaustively generate the collection of all those patterns which satisfy certain quality constraints, through a systematic combinatorial process guided by clear optimization criteria. Then, based on a set covering approach, LAD aggregates the collection of patterns into classification models. In addition, LAD is able to use the information provided by large collections of patterns in order to extract subsets of variables, which collectively are able to distinguish between different types of disease. RESULTS: For the differential diagnosis of DLBCL versus FL, a model based on eight significant genes is constructed and shown to have a sensitivity of 94.7% and a specificity of 100% on the test set. For the prognosis of good versus poor outcome among the DLBCL patients, a model is constructed on another set consisting also of eight significant genes, and shown to have a sensitivity of 87.5% and a specificity of 90% on the test set. The genes selected by LAD also work well as a basis for other kinds of statistical analysis, indicating their robustness. CONCLUSION: These two models exhibit accuracies that compare favorably to those in the original study. In addition, the current study also provides a ranking by importance of the genes in the selected significant subsets as well as a library of dozens of combinatorial biomarkers (i.e. pairs or triplets of genes) that can serve as a source of mathematically generated, statistically significant research hypotheses in need of biological explanation.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

G Alexe, S Alexe, D E Axelrod, P L Hammer, D Weissmann. 2005. Logical analysis of diffuse large B-cell lymphomas.. https://doi.org/10.1016/j.artmed.2004.11.004

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

New solid support for the synthesis of 3'-oligonucleotide conjugates through glyoxylic oxime bond formation.

A novel solid support 1 was synthesized to incorporate glyoxylic aldehyde functionality at the oligonucleotide 3'-terminus. 6-mer and 11-mer oligonucleotide sequences containing 3'-glyoxylic aldehyde functionality were prepared by using this support. These modified oligonucleotides were coupled to reporters containing an aminooxy group to prepare oligonucleotide 3'-conjugates through glyoxylic oxime bond formation. The hydrolytic stability of a glyoxylic oxime linkage was also investigated. [reaction: see text].

Combinatorial Chemistry Techniques↗

Chemical genetics: an evolving toolbox for target identification and lead optimization.

Chemical genetics combines chemistry with biology as a means of exploring the function of unknown proteins or identifying the proteins responsible for a particular phenotype. Chemical genetics is thus a valuable tool in the identification of novel drug targets. This chapter describes the application of chemical genetics in traditional and systems-based approaches to drug target discovery and the tools/approaches that appear most promising for guiding future pharmaceutical development.

Combinatorial Chemistry Techniques↗

Protein library design and screening: working out the probabilities.

In designing protein libraries for selection, we must coordinate our capacity to create a large diversity of protein variants with the physical limitations of what we can actually screen. This chapter aims to bring the language of probabilities into the protein engineer's laboratory to answer some of our common questions: How can we most efficiently design a library? What fraction of the theoretical library diversity have we actually sampled at the end of the day? What is the probability of missing an individual of the library? Are the mutations present in the variants we have selected statistically meaningful or the product of random variation? The computation of these criteria throughout the process of experimental protein engineering will enable us to better design and evaluate the products of our libraries of protein variants.

Combinatorial Chemistry Techniques↗