PubMed HealthSearch

PubMed · 9017212

Quantification of DNA patchiness using long-range correlation measures.

Abstract

We introduce and develop new techniques to quantify DNA patchiness, and to quantify characteristics of its mosaic structure. These techniques, which involve calculating two functions, alpha(l) and beta(l), measure correlations at length scale l and detect distinct characteristic patch sizes embedded in scale-invariant patch size distributions. Using these new methods, we address a number of issues relating to the mosaic structure of genomic DNA. We find several distinct characteristic patch sizes in certain genomic sequences, and compare, contrast, and quantify the correlation properties of different sequences, including a number of yeast, human, and prokaryotic sequences. We exclude the possibility that the correlation properties and the known mosaic structure of DNA can be explained either by simple Markov processes or by tandem repeats of dinucleotides. We find that the distinct patch sizes in all 16 yeast chromosomes are similar. Furthermore, we test the hypothesis that, for yeast, patchiness is caused by the alternation of coding and noncoding regions, and the hypothesis that in human sequences patchiness is related to repetitive sequences. We find that, by themselves, neither the alternation of coding and noncoding regions, nor repetitive sequences, can fully explain the long-range correlation properties of DNA.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

G M Viswanathan, S V Buldyrev, S Havlin, H E Stanley. 1997. Quantification of DNA patchiness using long-range correlation measures.. https://doi.org/10.1016/s0006-3495(97)78721-6

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Chromosome-level genome assembly of a cosmopolitan marine harmful algal bloom diatom species Chaetoceros socialis (Chaetocerotaceae).

Chaetoceros socialis is a cosmopolitan diatom species that is crucial for maintaining marine ecosystem structure and driving elemental cycles. C. socialis can form harmful algal blooms (HABs) that may cause a negative impact on the marine ecosystems. Whole-genome information for C. socialis is still unavailable, which may hinder more targeted studies on its ecological adaptive responses and evolutionary drivers. To address this gap, we employed cutting-edge genomic technologies including PacBio single-molecule real-time (SMRT) sequencing and high-throughput chromatin conformation capture (Hi-C) to achieve the first chromosome-level genome assembly of C. socialis. The assembled genome is 60.22 Mb in size with a scaffold N50 of 7.81 Mb and has been anchored to eight pseudochromosomes. A total of 13,378 protein-coding genes were predicted, of which 12,069 (90.22%) were functionally annotated. This high-quality genomic resource provides a fundamental data platform for systematically elucidating the ecological adaptation mechanisms of C. socialis.

Chromosomes

Chromosome-level genome assembly of the Vermilion Snapper (Rhomboplites aurorubens).

Vermilion Snapper (Rhomboplites aurorubens, Lutjanidae) inhabits deep waters (20-300 m) from North America to Brazil and supports significant commercial and recreational fisheries. Despite its economic importance, the understanding of its basic biology remains limited. Classified as Vulnerable on the Red List due to overfishing, populations have declined by over 30% in recent generations. We assembled and annotated the first chromosome-scale genome of this species by combining PacBio long reads, Illumina short reads, and Hi-C data. The resulting assembly is 987.5 Mbp, with a scaffold N50 size of 41.3 Mbp, and includes 135 contigs clustered and ordered onto 24 chromosomes with 34,496 predicted genes. The high-quality assembly and annotation contained about 98% complete and single-copy BUSCO genes. It is the most complete, chromosome-level genome assembly of an Atlantic snapper to date. The genome assembly and supporting data are valuable tools for ecological and comparative genomics studies of snappers and other valuable commercial species within the family.

Chromosomes

Influence of nucleosome structure on the three-dimensional folding of idealized minichromosomes.

BACKGROUND: The closed circular, multinucleosome-bound DNA comprising a minichromosome provides one of the best known examples of chromatin organization beyond the wrapping of the double helix around the core of histone proteins. This higher level of chain folding is governed by the topology of the constituent nucleosomes and the spatial disposition of the intervening protein-free DNA linkers. RESULTS: By simplifying the protein-DNA assembly to an alternating sequence of virtual bonds, the organization of a string of nucleosomes on the minichromosome can be treated by analogy to conventional chemical depictions of macromolecular folding in terms of the bond lengths, valence angles, and torsions of the chain. If the nucleosomes are evenly spaced and the linkers are sufficiently short, regular minichromosome structures can be identified from analytical expressions that relate the lengths and angles formed by the virtual bonds spanning the nucleosome-linker repeating units to the pitch and radius of the organized quaternary structures that they produce. CONCLUSIONS: The resulting models with 4-24 bound nucleosomes illustrate how a minichromosome can adopt the low-writhe folding motifs deduced from biochemical studies, and account for published images of the 30 nm chromatin fiber and the simian virus 40 (SV40) nucleohistone core. The marked sensitivity of global folding to the degree of protein-DNA interactions and the assumed nucleosomal shape suggest potential mechanisms for chromosome rearrangements upon histone modification.

Chromosomes