PubMed Health⌕ Search

Biomedical subjects

Lee Rowen

Publications and source records attributed to Lee Rowen.

At least 19 recordsLinked to original sources

A distinct effector B cell population drives autoantibody production in SARS-CoV-2 infection.

Autoantibodies (autoAbs) are linked to mortality and Long COVID, yet their cellular origins remain unclear. We analyzed the INCOV cohort and identified 12 age- and sex-matched participants with varying autoAb abundance and integrated single-cell RNA-seq and ATAC-seq data from B cells, plasma proteomics, proteome-wide autoAb profiling, clinical data, and in vitro assays. AutoAb abundance inversely correlated with neutralizing IgG and declined as infection resolved, paralleling the contraction of atypical memory B cells (AtMs). In vitro, AtMs preferentially differentiated into autoAb-producing antibody-secreting cells upon TLR7/8 stimulation. CD11c+ AtMs (double-negative 2, DN2s) in autoAb-high individuals exhibited increased TLR7 signaling, oxidative stress, and isotype switching, regulated by transcription factors T-bet and XBP1. Integrated genetic and genomic analyses showed that DN2s had the strongest enrichment for autoimmune trait heritability and inferred regulatory effects of autoimmune risk variants among B cell subsets. These findings identify DN2s as key precursors of autoAb-producing cells during SARS-CoV-2 infection.

B cell↗

Analysis of the DNA sequence and duplication history of human chromosome 15.

Here we present a finished sequence of human chromosome 15, together with a high-quality gene catalogue. As chromosome 15 is one of seven human chromosomes with a high rate of segmental duplication, we have carried out a detailed analysis of the duplication structure of the chromosome. Segmental duplications in chromosome 15 are largely clustered in two regions, on proximal and distal 15q; the proximal region is notable because recombination among the segmental duplications can result in deletions causing Prader-Willi and Angelman syndromes. Sequence analysis shows that the proximal and distal regions of 15q share extensive ancient similarity. Using a simple approach, we have been able to reconstruct many of the events by which the current duplication structure arose. We find that most of the intrachromosomal duplications seem to share a common ancestry. Finally, we demonstrate that some remaining gaps in the genome sequence are probably due to structural polymorphisms between haplotypes; this may explain a significant fraction of the gaps remaining in the human genome.

Animals↗

Unusual gene order and organization of the sea urchin hox cluster.

While the highly consistent gene order and axial colinear patterns of expression seem to be a feature of vertebrate hox gene clusters, this pattern may be less well conserved across the rest of the bilaterians. We report the first deuterostome instance of an intact hox cluster with a unique gene order where the paralog groups are not expressed in a sequential manner. The finished sequence from BAC clones from the genome of the sea urchin, Strongylocentrotus purpuratus, reveals a gene order wherein the anterior genes (Hox1, Hox2 and Hox3) lie nearest the posterior genes in the cluster such that the most 3' gene is Hox5. (The gene order is 5'-Hox1, 2, 3, 11/13c, 11/13b, 11/13a, 9/10, 8, 7, 6, 5-3'.) The finished sequence result is corroborated by restriction mapping evidence and BAC-end scaffold analyses. Comparisons with a putative ancestral deuterostome Hox gene cluster suggest that the rearrangements leading to the sea urchin gene order were many and complex.

Animals↗

The evolution of vertebrate Toll-like receptors.

The complete sequences of Takifugu Toll-like receptor (TLR) loci and gene predictions from many draft genomes enable comprehensive molecular phylogenetic analysis. Strong selective pressure for recognition of and response to pathogen-associated molecular patterns has maintained a largely unchanging TLR recognition in all vertebrates. There are six major families of vertebrate TLRs. This repertoire is distinct from that of invertebrates. TLRs within a family recognize a general class of pathogen-associated molecular patterns. Most vertebrates have exactly one gene ortholog for each TLR family. The family including TLR1 has more species-specific adaptations than other families. A major family including TLR11 is represented in humans only by a pseudogene. Coincidental evolution plays a minor role in TLR evolution. The sequencing phase of this study produced finished genomic sequences for the 12 Takifugu rubripes TLRs. In addition, we have produced >70 gene models, including sequences from the opossum, chicken, frog, dog, sea urchin, and sea squirt.

Animals↗

Interchromosomal segmental duplications explain the unusual structure of PRSS3, the gene for an inhibitor-resistant trypsinogen.

Homo sapiens possess several trypsinogen or trypsinogen-like genes of which three (PRSS1, PRSS2, and PRSS3) produce functional trypsins in the digestive tract. PRSS1 and PRSS2 are located on chromosome 7q35, while PRSS3 is found on chromosome 9p13. Here, we report a variation of the theme of new gene creation by duplication: the PRSS3 gene was formed by segmental duplications originating from chromosomes 7q35 and 11q24. As a result, PRSS3 transcripts display two variants of exon 1. The PRSS3 transcript whose gene organization most resembles PRSS1 and PRSS2 encodes a functional protein originally named mesotrypsinogen. The other variant is a fusion transcript, called trypsinogen IV. We show that the first exon of trypsinogen IV is derived from the noncoding first exon of LOC120224, a chromosome 11 gene. LOC120224 codes for a widely conserved transmembrane protein of unknown function. Comparative analyses suggest that these interchromosomal duplications occurred after the divergence of Old World monkeys and hominids. PRSS3 transcripts consist of a mixed population of mRNAs, some expressed in the pancreas and encoding an apparently functional trypsinogen and others of unknown function expressed in brain and a variety of other tissues. Analysis of the selection pressures acting on the trypsinogen gene family shows that, while the apparently functional genes are under mild to strong purifying selection overall, a few residues appear under positive selection. These residues could be involved in interactions with inhibitors.

Chromosomes, Human↗

T1DBase, a community web-based resource for type 1 diabetes research.

T1DBase (http://T1DBase.org) is a public website and database that supports the type 1 diabetes (T1D) research community. The site is currently focused on the molecular genetics and biology of T1D susceptibility and pathogenesis. It includes the following datasets: annotated genome sequence for human, rat and mouse; information on genetically identified T1D susceptibility regions in human, rat and mouse, and genetic linkage and association studies pertaining to T1D; descriptions of NOD mouse congenic strains; the Beta Cell Gene Expression Bank, which reports expression levels of genes in beta cells under various conditions, and annotations of gene function in beta cells; data on gene expression in a variety of tissues and organs; and biological pathways from KEGG and BioCarta. Tools on the site include the GBrowse genome browser, site-wide context dependent search, Connect-the-Dots for connecting gene and other identifiers from multiple data sources, Cytoscape for visualizing and analyzing biological networks, and the GESTALT workbench for genome annotation. All data are open access and all software is open source.

Animals↗

An enigmatic fourth runt domain gene in the fugu genome: ancestral gene loss versus accelerated evolution.

BACKGROUND: The runt domain transcription factors are key regulators of developmental processes in bilaterians, involved both in cell proliferation and differentiation, and their disruption usually leads to disease. Three runt domain genes have been described in each vertebrate genome (the RUNX gene family), but only one in other chordates. Therefore, the common ancestor of vertebrates has been thought to have had a single runt domain gene. RESULTS: Analysis of the genome draft of the fugu pufferfish (Takifugu rubripes) reveals the existence of a fourth runt domain gene, FrRUNT, in addition to the orthologs of human RUNX1, RUNX2 and RUNX3. The tiny FrRUNT packs six exons and two putative promoters in just 3 kb of genomic sequence. The first exon is located within an intron of FrSUPT3H, the ortholog of human SUPT3H, and the first exon of FrSUPT3H resides within the first intron of FrRUNT. The two gene structures are therefore "interlocked". In the human genome, SUPT3H is instead interlocked with RUNX2. FrRUNT has no detectable ortholog in the genomes of mammals, birds or amphibians. We consider alternative explanations for an apparent contradiction between the phylogenetic data and the comparison of the genomic neighborhoods of human and fugu runt domain genes. We hypothesize that an ancient RUNT locus was lost in the tetrapod lineage, together with FrFSTL6, a member of a novel family of follistatin-like genes. CONCLUSIONS: Our results suggest that the runt domain family may have started expanding in chordates much earlier than previously thought, and exemplify the importance of detailed analysis of whole-genome draft sequence to provide new insights into gene evolution.

Amino Acid Sequence↗

The human GRINL1A gene defines a complex transcription unit, an unusual form of gene organization in eukaryotes.

Sequencing of genomic DNA and cloned transcripts from the 200-kb human GRINL1A gene on chromosome 15 revealed a complex gene structure comprising at least 28 exons. In one gene model, transcription begins at exon 1 and ends at exon 15b. Another gene model begins transcription at exon 20 and terminates at exon 23, 24, or 28. In a third gene model, transcription begins at exon 1 and ends at exon 23, thus conjoining two apparently discrete genes into a third combined gene. Exon 15 can function as a terminating exon or as an alternatively spliced internal exon, or it can be skipped altogether. Exons 11, 14, 15a, 16, 17, 18, 19, 20a, and 20f are found only in transcripts that do not terminate at exon 15b. Combined transcripts that convert two genes into a third provide evidence for an unusual form of gene organization and expression that we call the complex transcription unit (CTU). Organization of exons into a CTU increases the extractable information content of a segment of genomic DNA and constitutes a potentially significant mechanism for augmenting the proteome of a genome.

DNA, Complementary↗

Genetic divergence of the rhesus macaque major histocompatibility complex.

The major histocompatibility complex (MHC) is comprised of the class I, class II, and class III regions, including the MHC class I and class II genes that play a primary role in the immune response and serve as an important model in studies of primate evolution. Although nonhuman primates contribute significantly to comparative human studies, relatively little is known about the genetic diversity and genomics underlying nonhuman primate immunity. To address this issue, we sequenced a complete rhesus macaque MHC spanning over 5.3 Mb, and obtained an additional 2.3 Mb from a second haplotype, including class II and portions of class I and class III. A major expansion of from six class I genes in humans to as many as 22 active MHC class I genes in rhesus and levels of sequence divergence some 10-fold higher than a similar human comparison were found, averaging from 2% to 6% throughout extended portions of class I and class II. These data pose new interpretations of the evolutionary constraints operating between MHC diversity and T-cell selection by contrasting with models predicting an optimal number of antigen presenting genes. For the clinical model, these data and derivative genetic tools can be implemented in ongoing genetic and disease studies that involve the rhesus macaque.

Animals↗

The organization and evolution of the dipteran and hymenopteran Down syndrome cell adhesion molecule (Dscam) genes.

The Drosophila melanogaster Down syndrome cell adhesion molecule (Dscam) gene encodes an axon guidance receptor and can generate 38,016 different isoforms via the alternative splicing of 95 variable exons. Dscam contains 10 immunoglobulin (Ig), six Fibronectin type III, a transmembrane (TM), and cytoplasmic domains. The different Dscam isoforms vary in the amino acid sequence of three of the Ig domains and the TM domain. Here, we have compared the organization of the Dscam gene from three members of the Drosophila subgenus (D. melanogaster, D. pseudoobscura, and D. virilis), the mosquito Anopheles gambiae, and the honeybee Apis mellifera. Each of these organisms contains numerous alternative exons and can potentially synthesize tens of thousands of isoforms. Interestingly, most of the alternative exons in one species are more similar to one another than to the corresponding alternative exons in the other species. These observations provide strong evidence that many of the alternative exons have arisen by reiterative exon duplication and deletion events. In addition, these findings suggest that the expression of a large Dscam repertoire is more important for the development and function of the insect nervous system than the actual sequence of each isoform.

Animals↗

Majority of divergence between closely related DNA samples is due to indels.

It was recently shown that indels are responsible for more than twice as many unmatched nucleotides as are base substitutions between samples of chimpanzee and human DNA. A larger sample has now been examined and the result is similar. The number of indels is approximately 1/12th of the number of base substitutions and the average length of the indels is 36 nt, including indels up to 10 kb. The ratio (R(u)) of unpaired nucleotides attributable to indels to those attributable to substitutions is 3.0 for this 2 million-nt chimp DNA sample compared with human. There is similar evidence of a large value of R(u) for sea urchins from the polymorphism of a sample of Strongylocentrotus purpuratus DNA (R(u) = 3-4). Other work indicates that similarly, per nucleotide affected, large differences are seen for indels in the DNA polymorphism of the plant Arabidopsis thaliana (R(u) = 51). For the insect Drosophila melanogaster a high value of R(u) (4.5) has been determined. For the nematode Caenorhabditis elegans the polymorphism data are incomplete but high values of R(u) are likely. Comparison of two strains of Escherichia coli O157:H7 shows a preponderance of indels. Because these six examples are from very distant systematic groups the implication is that in general, for alignments of closely related DNA, indels are responsible for many more unmatched nucleotides than are base substitutions. Human genetic evidence suggests that indels are a major source of gene defects, indicating that indels are a significant source of evolutionary change.

Animals↗

The DNA sequence and analysis of human chromosome 14.

Chromosome 14 is one of five acrocentric chromosomes in the human genome. These chromosomes are characterized by a heterochromatic short arm that contains essentially ribosomal RNA genes, and a euchromatic long arm in which most, if not all, of the protein-coding genes are located. The finished sequence of human chromosome 14 comprises 87,410,661 base pairs, representing 100% of its euchromatic portion, in a single continuous segment covering the entire long arm with no gaps. Two loci of crucial importance for the immune system, as well as more than 60 disease genes, have been localized so far on chromosome 14. We identified 1,050 genes and gene fragments, and 393 pseudogenes. On the basis of comparisons with other vertebrate genomes, we estimate that more than 96% of the chromosome 14 genes have been annotated. From an analysis of the CpG island occurrences, we estimate that 70% of these annotated genes are complete at their 5' end.

5' Untranslated Regions↗

Analysis of the gene-dense major histocompatibility complex class III region and its comparison to mouse.

In mammals, the Major Histocompatibility Complex class I and II gene clusters are separated by an approximately 700-kb stretch of sequence called the MHC class III region, which has been associated with susceptibility to numerous diseases. To facilitate understanding of this medically important and architecturally interesting portion of the genome, we have sequenced and analyzed both the human and mouse class III regions. The cross-species comparison has facilitated the identification of 60 genes in human and 61 in mouse, including a potential RNA gene for which the introns are more conserved across species than the exons. Delineation of global organization, gene structure, alternative splice forms, protein similarities, and potential cis-regulatory elements leads to several conclusions: (1) The human MHC class III region is the most gene-dense region of the human genome: >14% of the sequence is coding, approximately 72% of the region is transcribed, and there is an average of 8.5 genes per 100 kb. (2) Gene sizes, number of exons, and intergenic distances are for the most part similar in both species, implying that interspersed repeats have had little impact in disrupting the tight organization of this densely packed set of genes. (3) The region contains a heterogeneous mixture of genes, only a few of which have a clearly defined and proven function. Although many of the genes are of ancient origin, some appear to exist only in mammals and fish, implying they might be specific to vertebrates. (4) Conserved noncoding sequences are found primarily in or near the 5'-UTR or the first intron of genes, and seldom in the intergenic regions. Many of these conserved blocks are likely to be cis-regulatory elements.

Alternative Splicing↗

QUOD ERAT FACIENDUM: sequence analysis of the H2-D and H2-Q regions of 129/SvJ mice.

The H2-D and -Q regions of the mouse major histocompatibility complex ( Mhc or H2) have been sequenced from strain 129/SvJ (haplotype bc), revealing a D/Q region different from all other investigated haplotypes, including the closely related b haplotype. The 300-kb class I-rich region consists of the classical class I, H2-D, and 11 non-classical class I genes. The Q region was formed by two series of tandem duplications. Comparison of the segment between the D and Q1 genes with the H2-K region provides evidence that class I genes were translocated from the K region to the D region, and gives a new explanation for the weak locus specificity of the H-Kand H2-Dalleles.

Amino Acid Sequence↗

Whole-genome shotgun assembly and analysis of the genome of Fugu rubripes.

The compact genome of Fugu rubripes has been sequenced to over 95% coverage, and more than 80% of the assembly is in multigene-sized scaffolds. In this 365-megabase vertebrate genome, repetitive DNA accounts for less than one-sixth of the sequence, and gene loci occupy about one-third of the genome. As with the human genome, gene loci are not evenly distributed, but are clustered into sparse and dense regions. Some "giant" genes were observed that had average coding sequence sizes but were spread over genomic lengths significantly larger than those of their human orthologs. Although three-quarters of predicted human proteins have a strong match to Fugu, approximately a quarter of the human proteins had highly diverged from or had no pufferfish homologs, highlighting the extent of protein evolution in the 450 million years since teleosts and mammals diverged. Conserved linkages between Fugu and human genes indicate the preservation of chromosomal segments from the common vertebrate ancestor, but with considerable scrambling of gene order.

Animals↗

Patchy interspecific sequence similarities efficiently identify positive cis-regulatory elements in the sea urchin.

We demonstrate that interspecific sequence conservation can provide a systematic guide to the identification of functional cis-regulatory elements within a large expanse of genomic DNA. The test was carried out on the otx gene of Strongylocentrotus purpuratus. This gene plays a major role in the gene regulatory network that underlies endomesoderm specification in the embryo. The cis-regulatory organization of the otx gene is expected to be complex, because the gene has three different start sites (X. Li, C.-K. Chuang, C.-A. Mao, L. M. Angerer, and W. H. Klein, 1997, Dev. Biol. 187, 253-266), and it is expressed in many different spatial domains of the embryo. BAC recombinants containing the otx gene were isolated from Strongylocentrotus purpuratus and Lytechinus variegatus libraries, and the ordered sequence of these BACs was obtained and annotated. Sixty kilobases of DNA flanking the gene, and included in the BAC sequence from both species, were scanned computationally for short conserved sequence elements. For this purpose, we used a newly constructed software package assembled in our laboratory, "FamilyRelations." This tool allows detection of sequence similarities above a chosen criterion within sliding windows set at 20-50 bp. Seventeen partially conserved regions, most a few hundred base pairs long, were amplified from the S. purpuratus BAC DNA by PCR, inserted in an expression vector driving a CAT reporter, and tested for cis-regulatory activity by injection into fertilized S. purpuratus eggs. The regulatory activity of these constructs was assessed by whole-mount in situ hybridization (WMISH) using a probe against CAT mRNA. Of the 17 constructs, 11 constructs displayed spatially restricted regulatory activity, and 6 were inactive in this test. The domains within which the cis-regulatory constructs were expressed are approximately consistent with results from a WMISH study on otx expression in the embryo, in which we used probes specific for the mRNAs generated from each of the three transcription start sites. Four separate cis-regulatory elements that specifically produce endomesodermal expression were identified, as well as ubiquitously active elements, and ectoderm-specific elements. We confirm predictions from other work with respect to target sites for specific transcription factors within the elements that express in the endoderm.

Animals↗

A provisional regulatory gene network for specification of endomesoderm in the sea urchin embryo.

We present the current form of a provisional DNA sequence-based regulatory gene network that explains in outline how endomesodermal specification in the sea urchin embryo is controlled. The model of the network is in a continuous process of revision and growth as new genes are added and new experimental results become available; see http://www.its.caltech.edu/~mirsky/endomeso.htm (End-mes Gene Network Update) for the latest version. The network contains over 40 genes at present, many newly uncovered in the course of this work, and most encoding DNA-binding transcriptional regulatory factors. The architecture of the network was approached initially by construction of a logic model that integrated the extensive experimental evidence now available on endomesoderm specification. The internal linkages between genes in the network have been determined functionally, by measurement of the effects of regulatory perturbations on the expression of all relevant genes in the network. Five kinds of perturbation have been applied: (1) use of morpholino antisense oligonucleotides targeted to many of the key regulatory genes in the network; (2) transformation of other regulatory factors into dominant repressors by construction of Engrailed repressor domain fusions; (3) ectopic expression of given regulatory factors, from genetic expression constructs and from injected mRNAs; (4) blockade of the beta-catenin/Tcf pathway by introduction of mRNA encoding the intracellular domain of cadherin; and (5) blockade of the Notch signaling pathway by introduction of mRNA encoding the extracellular domain of the Notch receptor. The network model predicts the cis-regulatory inputs that link each gene into the network. Therefore, its architecture is testable by cis-regulatory analysis. Strongylocentrotus purpuratus and Lytechinus variegatus genomic BAC recombinants that include a large number of the genes in the network have been sequenced and annotated. Tests of the cis-regulatory predictions of the model are greatly facilitated by interspecific computational sequence comparison, which affords a rapid identification of likely cis-regulatory elements in advance of experimental analysis. The network specifies genomically encoded regulatory processes between early cleavage and gastrula stages. These control the specification of the micromere lineage and of the initial veg(2) endomesodermal domain; the blastula-stage separation of the central veg(2) mesodermal domain (i.e., the secondary mesenchyme progenitor field) from the peripheral veg(2) endodermal domain; the stabilization of specification state within these domains; and activation of some downstream differentiation genes. Each of the temporal-spatial phases of specification is represented in a subelement of the network model, that treats regulatory events within the relevant embryonic nuclei at particular stages.

Animals↗

A genomic regulatory network for development.

Development of the body plan is controlled by large networks of regulatory genes. A gene regulatory network that controls the specification of endoderm and mesoderm in the sea urchin embryo is summarized here. The network was derived from large-scale perturbation analyses, in combination with computational methodologies, genomic data, cis-regulatory analysis, and molecular embryology. The network contains over 40 genes at present, and each node can be directly verified at the DNA sequence level by cis-regulatory analysis. Its architecture reveals specific and general aspects of development, such as how given cells generate their ordained fates in the embryo and why the process moves inexorably forward in developmental time.

Animals↗