PubMed Health⌕ Search

PubMed · 11933190

Determining sequence length or content in zero, one, and two dimensions.

Abstract

High-throughput assays are essential for the practical application of mutation detection in medicine and research. Moreover, such assays should produce informative data of high quality that have a low-error rate and a low cost. Unfortunately, this is not currently the case. Instead, we typically witness legions of people reviewing imperfect data at astronomical expense yielding uncertain results. To address this problem, for the past decade we have been developing methods that exploit the inherent quantitative nature of DNA experiments. By generating high-quality data, careful DNA-signal quantification permits robust analysis for determining true alleles and certainty measures. We will explore several assays and methods. In a one-dimensional readout, short tandem repeat (STR) data display interesting artifacts. Even with high-quality data, PCR artifacts such as stutter and relative amplification can confound correct or automated scoring. However, by appropriate mathematical analysis, these artifacts can be essentially removed from the data. The result is fully automated data scoring, quality assessment, and new types of DNA analysis. These approaches enable the accurate analysis of pooled DNA samples, for both genetic and forensic applications. On a two-dimensional surface (comprised of zero-dimensional spots) one can perform assays of extremely high-throughput at low cost. The question is how to determine DNA sequence length or content from nonelectrophoretic intensity data. Here again, mathematical analysis of highly quantitative data provides a solution. We will discuss new lab assays that can produce data containing such information; mathematical transformation then determines DNA length or content.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Mark W Perlin, Beata Szabady. 2002. Determining sequence length or content in zero, one, and two dimensions.. https://doi.org/10.1002/humu.10087

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Striping artifact removal in VisiumHD data through nuclear counts modeling.

MOTIVATION: 10x Genomics VisiumHD enables spatial transcriptomics at 2 µm × 2 µm resolution but exhibits slide-specific, non-periodic striping artifacts due to lane-width variability. These multiplicative row/column effects distort bin total counts and can bias downstream analyses. The state-of-the-art destriping approach is the normalization procedure used as a preprocessing step in bin2cell; it applies sequential high-quantile row- then column-wise normalization, which is asymmetric and can introduce edge effects/macro-stripes and distortions of large-scale total-count structure. RESULTS: We propose a statistical destriping approach that leverages nuclei segmentation from the co-registered H&E image. Assuming transcript abundance is constant within each nucleus, we model bin counts with a negative binomial distribution whose mean is a product of a nucleus-specific concentration and row- and column-specific stripe-factors reflecting lane-width variation. We fit all parameters in a generalized linear modeling framework with cross-validated regularization on stripe-factors and iterative dispersion estimation, and use the fitted parameters to correct the observed counts into a destriped image. On synthetic data with known ground truth, our method improves stripe-factor estimation accuracy and reduces error in corrected counts relative to bin2cell and bin2cell-derived baselines. Across four public VisiumHD slides, it consistently lowers striping intensity while substantially better preserving biological signal present in the large-scale global count structure and avoiding the artifacts introduced by other methods. AVAILABILITY AND IMPLEMENTATION: All source code and links to publicly available data used for this study are available at https://github.com/paolamalsot/destriping-GLM.

Artifacts↗

PET/CT: a new imaging technology in nuclear medicine.

This review discusses the technical background of combined PET and CT and considers the clinical applications of PET/CT imaging. Questions addressed include: Is PET/CT superior to PET imaging alone? If so, in which patient populations and in what respect? Can PET/CT imaging affect patient management? Can PET/CT be practiced economically? While much work remains to be done, the available data clearly suggest that PET/CT decreases imaging time per patient and, even for the experienced reader, significantly reduces the number of equivocal PET interpretations. PET/CT also has the ability to improve accuracy of PET image interpretation and to affect clinical decision making, thereby improving patient management. The nuclear medicine community should approach this new technology with an open mind and focus on its clinical usefulness. The decision regarding whether PET/CT should be part of the equipment in a given nuclear medicine or radiology practice largely depends on the specific patient population. It is concluded that present skepticism concerning combined PET/CT will subside once critics of this new modality have had the opportunity to clearly see on images its many advantages compared with either PET alone or conventional image fusion approaches.

Artifacts↗

Position specific variation in the rate of evolution in transcription factor binding sites.

BACKGROUND: The binding sites of sequence specific transcription factors are an important and relatively well-understood class of functional non-coding DNAs. Although a wide variety of experimental and computational methods have been developed to characterize transcription factor binding sites, they remain difficult to identify. Comparison of non-coding DNA from related species has shown considerable promise in identifying these functional non-coding sequences, even though relatively little is known about their evolution. RESULTS: Here we analyse the genome sequences of the budding yeasts Saccharomyces cerevisiae, S. bayanus, S. paradoxus and S. mikatae to study the evolution of transcription factor binding sites. As expected, we find that both experimentally characterized and computationally predicted binding sites evolve slower than surrounding sequence, consistent with the hypothesis that they are under purifying selection. We also observe position-specific variation in the rate of evolution within binding sites. We find that the position-specific rate of evolution is positively correlated with degeneracy among binding sites within S. cerevisiae. We test theoretical predictions for the rate of evolution at positions where the base frequencies deviate from background due to purifying selection and find reasonable agreement with the observed rates of evolution. Finally, we show how the evolutionary characteristics of real binding motifs can be used to distinguish them from artefacts of computational motif finding algorithms. CONCLUSION: As has been observed for protein sequences, the rate of evolution in transcription factor binding sites varies with position, suggesting that some regions are under stronger functional constraint than others. This variation likely reflects the varying importance of different positions in the formation of the protein-DNA complex. The characterization of the pattern of evolution in known binding sites will likely contribute to the effective use of comparative sequence data in the identification of transcription factor binding sites and is an important step toward understanding the evolution of functional non-coding DNA.

Artifacts↗