PubMed Health⌕ Search

Biomedical subjects

Jerzy Jurka

Publications and source records attributed to Jerzy Jurka.

At least 19 recordsLinked to original sources

Annotation, submission and screening of repetitive elements in Repbase: RepbaseSubmitter and Censor.

BACKGROUND: Repbase is a reference database of eukaryotic repetitive DNA, which includes prototypic sequences of repeats and basic information described in annotations. Updating and maintenance of the database requires specialized tools, which we have created and made available for use with Repbase, and which may be useful as a template for other curated databases. RESULTS: We describe the software tools RepbaseSubmitter and Censor, which are designed to facilitate updating and screening the content of Repbase. RepbaseSubmitter is a java-based interface for formatting and annotating Repbase entries. It eliminates many common formatting errors, and automates actions such as calculation of sequence lengths and composition, thus facilitating curation of Repbase sequences. In addition, it has several features for predicting protein coding regions in sequences; searching and including Pubmed references in Repbase entries; and searching the NCBI taxonomy database for correct inclusion of species information and taxonomic position. Censor is a tool to rapidly identify repetitive elements by comparison to known repeats. It uses WU-BLAST for speed and sensitivity, and can conduct DNA-DNA, DNA-protein, or translated DNA-translated DNA searches of genomic sequence. Defragmented output includes a map of repeats present in the query sequence, with the options to report masked query sequence(s), repeat sequences found in the query, and alignments. CONCLUSION: Censor and RepbaseSubmitter are available as both web-based services and downloadable versions. They can be found at http://www.girinst.org/repbase/submission.html (RepbaseSubmitter) and http://www.girinst.org/censor/index.php (Censor).

Animals↗

Distinct catalytic and non-catalytic roles of ARGONAUTE4 in RNA-directed DNA methylation.

DNA methylation has important functions in stable, transcriptional gene silencing, immobilization of transposable elements and genome organization. In Arabidopsis, DNA methylation can be induced by double-stranded RNA through the RNA interference (RNAi) pathway, a response known as RNA-directed DNA methylation. This requires a specialized set of RNAi components, including ARGONAUTE4 (AGO4). Here we show that AGO4 binds to small RNAs including small interfering RNAs (siRNAs) originating from transposable and repetitive elements, and cleaves target RNA transcripts. Single mutations in the Asp-Asp-His catalytic motif of AGO4 do not affect siRNA-binding activity but abolish its catalytic potential. siRNA accumulation and non-CpG DNA methylation at some loci require the catalytic activity of AGO4, whereas others are less dependent on this activity. Our results are consistent with a model in which AGO4 can function at target loci through two distinct and separable mechanisms. First, AGO4 can recruit components that signal DNA methylation in a manner independent of its catalytic activity. Second, AGO4 catalytic activity can be crucial for the generation of secondary siRNAs that reinforce its repressive effects.

Amino Acid Sequence↗

Identification of sequence motifs at the breakpoint junctions in three t(1;9)(p36.3;q34) and delineation of mechanisms involved in generating balanced translocations.

Although approximately 1 in 500 individuals carries a reciprocal translocation, little is known about the mechanisms that result in their formation. We analyzed the sequences surrounding the breakpoints in three unbalanced translocations of 1p and 9q, all of which were designated t(1;9)(p36.3;q34), to investigate the presence of sequence motifs that might mediate nonhomologous end joining (NHEJ). The breakpoint regions were unique in all individuals. Two of three translocations demonstrated insertions and duplications at the junctions, suggesting NHEJ in the formation of the rearrangements. No homology was identified in the breakpoint regions, further supporting NHEJ. We found translin motifs at the breakpoint junctions, suggesting the involvement of translin in the joining of the broken chromosome ends. We propose a model for balanced translocation formation in humans similar to transposition in bacteria, in which staggered nicks are repaired resulting in duplications and insertions at the translocation breakpoints.

Base Sequence↗

Positive selection on the nonhomologous end-joining factor Cernunnos-XLF in the human lineage.

BACKGROUND: Cernunnos-XLF is a nonhomologous end-joining factor that is mutated in patients with a rare immunodeficiency with microcephaly. Several other microcephaly-associated genes such as ASPM and microcephalin experienced recent adaptive evolution apparently linked to brain size expansion in humans. In this study we investigated whether Cernunnos-XLF experienced similar positive selection during human evolution. RESULTS: We obtained or reconstructed full-length coding sequences of chimpanzee, rhesus macaque, canine, and bovine Cernunnos-XLF orthologs from sequence databases and sequence trace archives. Comparison of coding sequences revealed an excess of nonsynonymous substitutions consistent with positive selection on Cernunnos-XLF in the human lineage. The hotspots of adaptive evolution are concentrated around a specific structural domain, whose analogue in the structurally similar XRCC4 protein is involved in binding of another nonhomologous end-joining factor, DNA ligase IV. CONCLUSION: Cernunnos-XLF is a microcephaly-associated locus newly identified to be under adaptive evolution in humans, and possibly played a role in human brain expansion. We speculate that Cernunnos-XLF may have contributed to the increased number of brain cells in humans by efficient double strand break repair, which helps to prevent frequent apoptosis of neuronal progenitors and aids mitotic cell cycle progression. REVIEWERS: This article was reviewed by Chris Ponting and Richard Emes (nominated by Chris Ponting), Kateryna Makova, Gáspár Jékely and Eugene V. Koonin.

Journal Article↗

Self-synthesizing DNA transposons in eukaryotes.

Eukaryotes contain numerous transposable or mobile elements capable of parasite-like proliferation in the host genome. All known transposable elements in eukaryotes belong to two types: retrotransposons and DNA transposons. Here we report a previously uncharacterized class of DNA transposons called Polintons that populate genomes of protists, fungi, and animals, including entamoeba, soybean rust, hydra, sea anemone, nematodes, fruit flies, beetle, sea urchin, sea squirt, fish, lizard, frog, and chicken. Polintons from all these species are characterized by a unique set of proteins necessary for their transposition, including a protein-primed DNA polymerase B, retroviral integrase, cysteine protease, and ATPase. In addition, Polintons are characterized by 6-bp target site duplications, terminal-inverted repeats that are several hundred nucleotides long, and 5'-AG and TC-3' termini. Analogously to known transposable elements, Polintons exist as autonomous and nonautonomous elements. Our data suggest that Polintons have evolved from a linear plasmid that acquired a retroviral integrase at least 1 billion years ago. According to the model of Polinton transposition proposed here, a Polinton DNA molecule excised from the genome serves as a template for extrachromosomal synthesis of its double-stranded DNA copy by the Polinton-encoded DNA polymerase and is inserted back into genome by its integrase.

Adenosine Triphosphatases↗

Sequencing of Aspergillus nidulans and comparative analysis with A. fumigatus and A. oryzae.

The aspergilli comprise a diverse group of filamentous fungi spanning over 200 million years of evolution. Here we report the genome sequence of the model organism Aspergillus nidulans, and a comparative study with Aspergillus fumigatus, a serious human pathogen, and Aspergillus oryzae, used in the production of sake, miso and soy sauce. Our analysis of genome structure provided a quantitative evaluation of forces driving long-term eukaryotic genome evolution. It also led to an experimentally validated model of mating-type locus evolution, suggesting the potential for sexual reproduction in A. fumigatus and A. oryzae. Our analysis of sequence conservation revealed over 5,000 non-coding regions actively conserved across all three species. Within these regions, we identified potential functional elements including a previously uncharacterized TPP riboswitch and motifs suggesting regulation in filamentous fungi by Puf family genes. We further obtained comparative and experimental evidence indicating widespread translational regulation by upstream open reading frames. These results enhance our understanding of these widely studied fungi as well as provide new insight into eukaryotic genome evolution and gene regulation.

Aspergillus fumigatus↗

Origin and diversification of minisatellites derived from human Alu sequences.

We analyze minisatellites derived from Alu fragments corresponding approximately to the first 44 bases of human Alu consensus sequences from different subfamilies. The origin of Alu-derived minisatellites appears to have been mediated by short flanking repeats, as first proposed by Haber and Louis [Haber, J.E., Louis, E.J., 1998. Minisatellite origins in yeast and humans. Genomics 48, 132-135.]. We also present evidence for base substitutions and deletions introduced to minisatellites by gene conversion with partially similar but unrelated flanking regions. Segments flanked by short direct repeats are relatively common in different regions of Alu and other repetitive sequences. Our analysis shows that they can be effectively used in comparative studies of the overall sequence context which may contribute to instability of DNA segments flanked by short direct repeats.

Alu Elements↗

Retroposition of processed pseudogenes: the impact of RNA stability and translational control.

Human processed pseudogenes are copies of cellular RNAs reverse transcribed and inserted into the nuclear genome by the enzymatic machinery of L1 (LINE1) non-LTR retrotransposons. Although it is generally accepted that germline expression is crucial for the heritable retroposition of cellular mRNAs, little is known about the influences of RNA stability, mRNA quality control and compartmentalization of translation on the retroposition of processed pseudogenes. We found that frequently retroposed human mRNAs are derived from stable transcripts with translation-competent functional reading frames that are resistant to nonsense-mediated RNA decay. They are preferentially translated on free cytoplasmic ribosomes and encode soluble proteins. Our results indicate that interactions between mRNAs and L1 proteins seem to occur at free cytoplasmic ribosomes.

Animals↗

The microcephaly ASPM gene is expressed in proliferating tissues and encodes for a mitotic spindle protein.

The most common cause of primary autosomal recessive microcephaly (MCPH) appears to be mutations in the ASPM gene which is involved in the regulation of neurogenesis. The predicted gene product contains two putative N-terminal calponin-homology (CH) domains and a block of putative calmodulin-binding IQ domains common in actin binding cytoskeletal and signaling proteins. Previous studies in mouse suggest that ASPM is preferentially expressed in the developing brain. Our analyses reveal that ASPM is widely expressed in fetal and adult tissues and upregulated in malignant cells. Several alternatively spliced variants encoding putative ASPM isoforms with different numbers of IQ motifs were identified. The major ASPM transcript contains 81 IQ domains, most of which are organized into a higher order repeat (HOR) structure. Another prominent spliced form contains an in-frame deletion of exon 18 and encodes 14 IQ domains not organized into a HOR. This variant is conserved in mouse. Other spliced variants lacking both CH domains and a part of the IQ motifs were also detected, suggesting the existence of isoforms with potentially different functions. To elucidate the biochemical function of human ASPM, we developed peptide specific antibodies to the N- and C-termini of ASPM. In a western analysis of proteins from cultured human and mouse cells, the antibodies detected bands with mobilities corresponding to the predicted ASPM isoforms. Immunostaining of cultured human cells with antibodies revealed that ASPM is localized in the spindle poles during mitosis. This finding suggests that MCPH is the consequence of an impairment in mitotic spindle regulation in cortical progenitors due to mutations in ASPM.

Adult↗

Evolutionary diversity and potential recombinogenic role of integration targets of Non-LTR retrotransposons.

Short interspersed elements (SINEs) make up a significant fraction of total DNA in mammalian genomes, providing a rich substrate for chromosomal rearrangements by SINE-SINE recombinations. Proliferation of mammalian SINEs is mediated primarily by long interspersed element 1 (L1) non-long terminal repeat retrotransposons that preferentially integrate at DNA sequence targets with an average length of approximately 15 bp and containing conserved endonucleolytic nicking signals at both ends. We report that sequence variations in the first of the two nicking signals, represented by a 5'-TT-AAAA consensus sequence, affect the position of the second signal thus leading to target site duplications (TSDs) of different lengths. The length distribution of TSDs appears to be affected also by L1-encoded enzyme variants because targets with the same 5' nicking site can be of different average lengths in different mammalian species. Taking this into account, we reanalyzed the second nicking site and found that it is larger and includes more conserved sites than previously appreciated, with a consensus of 5'-ANTNTN-AA. We also studied potential involvement of the nicking sites in stimulating recombinations between SINEs. We determined that SINEs retaining TSDs with perfect 5'-TT-AAAA nicking sites appear to be lost relatively rapidly from the human and rat genomes and less rapidly from dog. We speculate that the introduction of DNA breaks induced by recurring endonucleolytic attacks at these sites, combined with the ubiquitousness of SINEs, may significantly promote recombination between repetitive elements, leading to the observed losses. At the same time, new L1 subfamilies may be selected for "incompatibility" with preexisting targets. This provides a possible driving force for the continual emergence of new L1 subfamilies which, in turn, may affect selection of L1-dependent SINE subfamilies.

Animals↗

RAG1 core and V(D)J recombination signal sequences were derived from Transib transposons.

The V(D)J recombination reaction in jawed vertebrates is catalyzed by the RAG1 and RAG2 proteins, which are believed to have emerged approximately 500 million years ago from transposon-encoded proteins. Yet no transposase sequence similar to RAG1 or RAG2 has been found. Here we show that the approximately 600-amino acid "core" region of RAG1 required for its catalytic activity is significantly similar to the transposase encoded by DNA transposons that belong to the Transib superfamily. This superfamily was discovered recently based on computational analysis of the fruit fly and African malaria mosquito genomes. Transib transposons also are present in the genomes of sea urchin, yellow fever mosquito, silkworm, dog hookworm, hydra, and soybean rust. We demonstrate that recombination signal sequences (RSSs) were derived from terminal inverted repeats of an ancient Transib transposon. Furthermore, the critical DDE catalytic triad of RAG1 is shared with the Transib transposase as part of conserved motifs. We also studied several divergent proteins encoded by the sea urchin and lancelet genomes that are 25%-30% identical to the RAG1 N-terminal domain and the RAG1 core. Our results provide the first direct evidence linking RAG1 and RSSs to a specific superfamily of DNA transposons and indicate that the V(D)J machinery evolved from transposons. We propose that only the RAG1 core was derived from the Transib transposase, whereas the N-terminal domain was assembled from separate proteins of unknown function that may still be active in sea urchin, lancelet, hydra, and starlet sea anemone. We also suggest that the RAG2 protein was not encoded by ancient Transib transposons but emerged in jawed vertebrates as a counterpart of RAG1 necessary for the V(D)J recombination reaction.

Aedes↗

Analysis of the human Alu Ye lineage.

BACKGROUND: Alu elements are short (approximately 300 bp) interspersed elements that amplify in primate genomes through a process termed retroposition. The expansion of these elements has had a significant impact on the structure and function of primate genomes. Approximately 10 % of the mass of the human genome is comprised of Alu elements, making them the most abundant short interspersed element (SINE) in our genome. The majority of Alu amplification occurred early in primate evolution, and the current rate of Alu retroposition is at least 100 fold slower than the peak of amplification that occurred 30-50 million years ago. Alu elements are therefore a rich source of inter- and intra-species primate genomic variation. RESULTS: A total of 153 Alu elements from the Ye subfamily were extracted from the draft sequence of the human genome. Analysis of these elements resulted in the discovery of two new Alu subfamilies, Ye4 and Ye6, complementing the previously described Ye5 subfamily. DNA sequence analysis of each of the Alu Ye subfamilies yielded average age estimates of approximately 14, approximately 13 and approximately 9.5 million years old for the Alu Ye4, Ye5 and Ye6 subfamilies, respectively. In addition, 120 Alu Ye4, Ye5 and Ye6 loci were screened using polymerase chain reaction (PCR) assays to determine their phylogenetic origin and levels of human genomic diversity. CONCLUSION: The Alu Ye lineage appears to have started amplifying relatively early in primate evolution and continued propagating at a low level as many of its members are found in a variety of hominoid (humans, greater and lesser ape) genomes. Detailed sequence analysis of several Alu pre-integration sites indicated that multiple types of events had occurred, including gene conversions, near-parallel independent insertions of different Alu elements and Alu-mediated genomic deletions. A potential hotspot for Alu insertion in the Fer1L3 gene on chromosome 10 was also identified.

Alu Elements↗

Dynamic structure of the SPANX gene cluster mapped to the prostate cancer susceptibility locus HPCX at Xq27.

Genetic linkage studies indicate that germline variations in a gene or genes on chromosome Xq27-28 are implicated in prostate carcinogenesis. The linkage peak of prostate cancer overlies a region of approximately 750 kb containing five SPANX genes (SPANX-A1, -A2, -B, -C, and -D) encoding sperm proteins associated with the nucleus; their expression was also detected in a variety of cancers. SPANX genes are >95% identical and reside within large segmental duplications (SDs) with a high level of similarity, which confounds mutational analysis of this gene family by routine PCR methods. In this work, we applied transformation-associated recombination cloning (TAR) in yeast to characterize individual SPANX genes from prostate cancer patients showing linkage to Xq27-28 and unaffected controls. Analysis of genomic TAR clones revealed a dynamic nature of the replicated region of linkage. Both frequent gene deletion/duplication and homology-based sequence transfer events were identified within the region and were presumably caused by recombinational interactions between SDs harboring the SPANX genes. These interactions contribute to diversity of the SPANX coding regions in humans. We speculate that the predisposition to prostate cancer in X-linked families is an example of a genomic disease caused by a specific architecture of the SPANX gene cluster.

Base Sequence↗

Traffic of genetic information between segmental duplications flanking the typical 22q11.2 deletion in velo-cardio-facial syndrome/DiGeorge syndrome.

Velo-cardio-facial syndrome/DiGeorge syndrome results from unequal crossing-over events between two 240-kb low-copy repeats termed LCR22 (LCR22-2 and LCR22-4) on Chromosome 22q11.2, comprised of modules, each of which are >99% identical in sequence. To delineate regions in the LCR22s that might contain hotspots for 22q11.2 rearrangements, we scanned the interval for increased rates of recombination with the hypothesis that these regions might be more prone to breakage. We generated an algorithm to detect sites of altered recombination by searching for single nucleotide polymorphic positions in BAC clones from different libraries mapped to LCR22-2 and LCR22-4. This method distinguishes single nucleotide polymorphisms from paralogous sequence variants and complex polymorphic positions. Sites of shared polymorphism are considered potential sites of gene conversion or double cross-over between the two LCR22s. We found an inverse correlation between regions of paralogous sequence variants that are unique to a given position within one LCR22 and clusters of shared polymorphic sites, suggesting that these clusters depict altered recombination and not remnants of ancestral single nucleotide polymorphisms. We postulate that most shared polymorphic sites are products of past transfers of DNA information between the LCR22s, suggesting that frequent traffic of genetic material may induce genomic instability in the two LCR22s. We also found that gaps up to 1.5 kb long can be transferred between LCR22s.

Algorithms↗

The genome of the diatom Thalassiosira pseudonana: ecology, evolution, and metabolism.

Diatoms are unicellular algae with plastids acquired by secondary endosymbiosis. They are responsible for approximately 20% of global carbon fixation. We report the 34 million-base pair draft nuclear genome of the marine diatom Thalassiosira pseudonana and its 129 thousand-base pair plastid and 44 thousand-base pair mitochondrial genomes. Sequence and optical restriction mapping revealed 24 diploid nuclear chromosomes. We identified novel genes for silicic acid transport and formation of silica-based cell walls, high-affinity iron uptake, biosynthetic enzymes for several types of polyunsaturated fatty acids, use of a range of nitrogenous compounds, and a complete urea cycle, all attributes that allow diatoms to prosper in aquatic environments.

Adaptation, Physiological↗

Evolution of the tumor suppressor BRCA1 locus in primates: implications for cancer predisposition.

Germ-line mutations in the BRCA1 gene predispose affected individuals to breast and ovarian cancer syndromes. In an attempt to systematically analyze a broader spectrum of genetic changes ranging from frequent exon deletions and duplications to amino acid replacements and protein truncations, we isolated and characterized full size BRCA1 homologues from a representative group of non-human primates. Our analysis represents the first comprehensive sequence comparison of primate BRCA1 loci and corresponding proteins. The comparison revealed an unusually high proportion of indels in non-coding DNA. The major force driving evolutionary changes in non-coding BRCA1 sequences was Alu-mediated rearrangements, including Alu transpositions and Alu-associated deletions, indicating that structural instability of this locus may be intrinsic in anthropoids. Analysis of the non-synonymous/synonymous ratio in coding portions of the gene revealed the presence of both conserved and rapidly evolving regions in the BRCA1 protein. Previously, a rapidly evolving region with evidence of positive evolutionary selection in human and chimpanzee had been identified only in exon 11. Here, we show that most of the internal BRCA1 sequence is variable between primates and evolved under positive selection. In contrast, the terminal regions of BRCA1, which encode the RING finger and BRCT domains, experienced negative selection, which left them almost identical between the compared primates. Distribution of the reported missense mutations, but not frameshift and nonsense mutations, is positively correlated with BRCA1 protein conservation. Finally, on the basis of protein sequence conservation, we identified missense changes that are likely to compromise BRCA1 function.

Alu Elements↗

Accelerated evolution of the ASPM gene controlling brain size begins prior to human brain expansion.

Primary microcephaly (MCPH) is a neurodevelopmental disorder characterized by global reduction in cerebral cortical volume. The microcephalic brain has a volume comparable to that of early hominids, raising the possibility that some MCPH genes may have been evolutionary targets in the expansion of the cerebral cortex in mammals and especially primates. Mutations in ASPM, which encodes the human homologue of a fly protein essential for spindle function, are the most common known cause of MCPH. Here we have isolated large genomic clones containing the complete ASPM gene, including promoter regions and introns, from chimpanzee, gorilla, orangutan, and rhesus macaque by transformation-associated recombination cloning in yeast. We have sequenced these clones and show that whereas much of the sequence of ASPM is substantially conserved among primates, specific segments are subject to high Ka/Ks ratios (nonsynonymous/synonymous DNA changes) consistent with strong positive selection for evolutionary change. The ASPM gene sequence shows accelerated evolution in the African hominoid clade, and this precedes hominid brain expansion by several million years. Gorilla and human lineages show particularly accelerated evolution in the IQ domain of ASPM. Moreover, ASPM regions under positive selection in primates are also the most highly diverged regions between primates and nonprimate mammals. We report the first direct application of TAR cloning technology to the study of human evolution. Our data suggest that evolutionary selection of specific segments of the ASPM sequence strongly relates to differences in cerebral cortical size.

Animals↗