PubMed Health⌕ Search

Biomedical subjects

G Bartha

Publications and source records attributed to G Bartha.

6 recordsLinked to original sources

Basecalling with LifeTrace.

A pivotal step in electrophoresis sequencing is the conversion of the raw, continuous chromatogram data into the actual sequence of discrete nucleotides, a process referred to as basecalling. We describe a novel algorithm for basecalling implemented in the program LifeTrace. Like Phred, currently the most widely used basecalling software program, LifeTrace takes processed trace data as input. It was designed to be tolerant to variable peak spacing by means of an improved peak-detection algorithm that emphasizes local chromatogram information over global properties. LifeTrace is shown to generate high-quality basecalls and reliable quality scores. It proved particularly effective when applied to MegaBACE capillary sequencing machines. In a benchmark test of 8372 dye-primer MegaBACE chromatograms, LifeTrace generated 17% fewer substitution errors, 16% fewer insertion/deletion errors, and 2.4% more aligned bases to the finished sequence than did Phred. For two sets totaling 6624 dye-terminator chromatograms, the performance improvement was 15% fewer substitution errors, 10% fewer insertion/deletion errors, and 2.1% more aligned bases. The processing time required by LifeTrace is comparable to that of Phred. The predicted quality scores were in line with observed quality scores, permitting direct use for quality clipping and in silico single nucleotide polymorphism (SNP) detection. Furthermore, we introduce a new type of quality score associated with every basecall: the gap-quality. It estimates the probability of a deletion error between the current and the following basecall. This additional quality score improves detection of single basepair deletions when used for locating potential basecalling errors during the alignment. We also describe a new protocol for benchmarking that we believe better discerns basecaller performance differences than methods previously published.

Algorithms↗

Serum cholesterol levels in American (Pima) Indian children and adolescents.

Serum cholesterol levels from birth to adulthood in a population of North American (Pima) Indians are described and compared to those of Caucasian populations. Cholesterol levels at birth (mean +/- SEM, 87 +/- 2.6 mg/100 ml) were similar in Pimas and Caucasians, but levels in Pimas from 5 to 16 years (148 +/- 4.6 mg/100 ml) were 20 to 30 mg/100 ml lower than among most white populations. The levels showed little rise with age from 5 to 16, then rose significantly in both sexes from ages 17 to 25. Cholesterol levels in adult Pimas (190 +/- 1.5 mg/100 ml) were up to 50 to 60 mg/100 ml lower than in American whites, and showed little increase after age 25. Two cohorts of children followed prospectively for six years indicated that the prevalence data reflect sequential changes in the population. Cholesterol levels of those subjects were significantly correlated at the first and last examinations. The Pima, in contrast to Caucasian American populations, have relatively low levels of serum cholesterol and low rates of coronary heart disease, but evidence of a causal relationship with the latter remains to be established.

Adolescent↗