PubMed Health⌕ Search

PubMed · 14707167

Mapping and initial analysis of human subtelomeric sequence assemblies.

Abstract

Physical mapping data were combined with public draft and finished sequences to derive subtelomeric sequence assemblies for each of the 41 genetically distinct human telomere regions. Sequence gaps that remain on the reference telomeres are generally small,well-defined,and for the most part,restricted to regions directly adjacent to the terminal (TTAGGG)n tract. Of the 20.66 Mb of subtelomeric DNA analyzed, 3.01 Mb are subtelomeric repeat sequences (Srpt),and an additional 2.11 Mb are segmental duplications. The subtelomeric sequence assemblies are enriched >25-fold in short,internal (TTAGGG)n-like sequences relative to the rest of the genome; a total of 114 (TTAGGG)n-like islands were found,55 within Srpt regions,35 within one-copy regions,11 at one-copy/Srpt or Srpt/segmental duplication boundaries,and 13 at the telomeric ends of assemblies. Transcripts were annotated in each assembly,noting their mapping coordinates relative to their respective telomere and whether they originate in duplicated DNA or single-copy DNA. A total of 697 transcripts were found in 15.53 Mb of one-copy DNA,76 transcripts in 2.11 Mb of segmentally duplicated DNA,and 168 transcripts in 3.01 Mb of Srpt sequence. This overall transcript density is similar (within approximately 10%) to that found genome-wide. Zinc finger-containing genes and olfactory receptor genes are duplicated within and between multiple telomere regions.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Harold Riethman, Anthony Ambrosini, Carlos Castaneda, Jeffrey Finklestein, Xue-Lan Hu, Uma Mudunuri, Sheila Paul, Jun Wei. 2004. Mapping and initial analysis of human subtelomeric sequence assemblies.. https://doi.org/10.1101/gr.1245004

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Sp1 is essential for p16 expression in human diploid fibroblasts during senescence.

BACKGROUND: p16(INK4a) tumor suppressor protein has been widely proposed to mediate entrance of the cells into the senescent stage. Promoter of p16(INK4a) gene contains at least five putative GC boxes, named GC-I to V, respectively. Our previous data showed that a potential Sp1 binding site, within the promoter region from -466 to -451, acts as a positive transcription regulatory element. These results led us to examine how Sp1 and/or Sp3 act on these GC boxes during aging in cultured human diploid fibroblasts. METHODOLOGY/PRINCIPAL FINDINGS: Mutagenesis studies revealed that GC-I, II and IV, especially GC-II, are essential for p16(INK4a) gene expression in senescent cells. Electrophoretic mobility shift assays (EMSA) and ChIP assays demonstrated that both Sp1 and Sp3 bind to these elements and the binding activity is enhanced in senescent cells. Ectopic overexpression of Sp1, but not Sp3, induced the transcription of p16(INK4a). Both Sp1 RNAi and Mithramycin, a DNA intercalating agent that interferes with Sp1 and Sp3 binding activities, reduced p16(INK4a) gene expression. In addition, the enhanced binding of Sp1 to p16(INK4a) promoter during cellular senescence appeared to be the result of increased Sp1 binding affinity, not an alteration in Sp1 protein level. CONCLUSIONS/SIGNIFICANCE: All these results suggest that GC- II is the key site for Sp1 binding and increase of Sp1 binding activity rather than protein levels contributes to the induction of p16(INK4a) expression during cell aging.

Base Composition↗

Application of CE for determination of DNA base composition.

DNA base composition expressed as mol% of guanine plus cytosine (% GC) or GC content is a key parameter of bacterial taxonomy and genomic analyses. Direct chemical determination methods such as HPLC as well as indirect methods based on physical properties of deoxyribonucleic acid (DNA), melting point (T(m)), and buoyant density (B(d)) have been conventionally applied to determine the GC content. However, these methods require relatively large amounts of sample DNA, time, and labor. We have developed a protocol to determine the GC content by fine separation of nucleosides with CZE. Genomic DNAs with known GC content from 23 bacterial strains were determined by CE at the optimized conditions of 27 degrees C, 20 kV in 50 mM of NaHCO(3) (pH 9.0) and 70 mM SDS added. Nucleosides from <1 microg of DNA hydrolyzed with nuclease-P1 and bacterial alkaline phosphatase were separated in a 75 microm wide and 80 cm long silica capillary. The nucleoside peak areas were determined at 254 nm in less than 12 min. The CE-based determination of GC content requires only small amounts of DNA, and thus should be applicable to environmental genomics (metagenomics), as >90% of environmental micro-organisms are nonculturable and produce only small amounts of genomic DNA.

Base Composition↗