PubMed Health⌕ Search

Biomedical subjects

Jane Rogers

Publications and source records attributed to Jane Rogers.

At least 19 recordsLinked to original sources

DNA sequence of human chromosome 17 and analysis of rearrangement in the human lineage.

Chromosome 17 is unusual among the human chromosomes in many respects. It is the largest human autosome with orthology to only a single mouse chromosome, mapping entirely to the distal half of mouse chromosome 11. Chromosome 17 is rich in protein-coding genes, having the second highest gene density in the genome. It is also enriched in segmental duplications, ranking third in density among the autosomes. Here we report a finished sequence for human chromosome 17, as well as a structural comparison with the finished sequence for mouse chromosome 11, the first finished mouse chromosome. Comparison of the orthologous regions reveals striking differences. In contrast to the typical pattern seen in mammalian evolution, the human sequence has undergone extensive intrachromosomal rearrangement, whereas the mouse sequence has been remarkably stable. Moreover, although the human sequence has a high density of segmental duplication, the mouse sequence has a very low density. Notably, these segmental duplications correspond closely to the sites of structural rearrangement, demonstrating a link between duplication and rearrangement. Examination of the main classes of duplicated segments provides insight into the dynamics underlying expansion of chromosome-specific, low-copy repeats in the human genome.

Animals↗

Human chromosome 11 DNA sequence and analysis including novel gene identification.

Chromosome 11, although average in size, is one of the most gene- and disease-rich chromosomes in the human genome. Initial gene annotation indicates an average gene density of 11.6 genes per megabase, including 1,524 protein-coding genes, some of which were identified using novel methods, and 765 pseudogenes. One-quarter of the protein-coding genes shows overlap with other genes. Of the 856 olfactory receptor genes in the human genome, more than 40% are located in 28 single- and multi-gene clusters along this chromosome. Out of the 171 disorders currently attributed to the chromosome, 86 remain for which the underlying molecular basis is not yet known, including several mendelian traits, cancer and susceptibility loci. The high-quality data presented here--nearly 134.5 million base pairs representing 99.8% coverage of the euchromatic sequence--provide scientists with a solid foundation for understanding the genetic basis of these disorders and other biological phenomena.

Chromosomes, Human, Pair 11↗

Genomic anatomy of the Tyrp1 (brown) deletion complex.

Chromosome deletions in the mouse have proven invaluable in the dissection of gene function. The brown deletion complex comprises >28 independent genome rearrangements, which have been used to identify several functional loci on chromosome 4 required for normal embryonic and postnatal development. We have constructed a 172-bacterial artificial chromosome contig that spans this 22-megabase (Mb) interval and have produced a contiguous, finished, and manually annotated sequence from these clones. The deletion complex is strikingly gene-poor, containing only 52 protein-coding genes (of which only 39 are supported by human homologues) and has several further notable genomic features, including several segments of >1 Mb, apparently devoid of a coding sequence. We have used sequence polymorphisms to finely map the deletion breakpoints and identify strong candidate genes for the known phenotypes that map to this region, including three lethal loci (l4Rn1, l4Rn2, and l4Rn3) and the fitness mutant brown-associated fitness (baf). We have also characterized misexpression of the basonuclin homologue, Bnc2, associated with the inversion-mediated coat color mutant white-based brown (B(w)). This study provides a molecular insight into the basis of several characterized mouse mutants, which will allow further dissection of this region by targeted or chemical mutagenesis.

Animals↗

Spread of an inactive form of caspase-12 in humans is due to recent positive selection.

The human caspase-12 gene is polymorphic for the presence or absence of a stop codon, which results in the occurrence of both active (ancestral) and inactive (derived) forms of the gene in the population. It has been shown elsewhere that carriers of the inactive gene are more resistant to severe sepsis. We have now investigated whether the inactive form has spread because of neutral drift or positive selection. We determined its distribution in a worldwide sample of 52 populations and resequenced the gene in 77 individuals from the HapMap Yoruba, Han Chinese, and European populations. There is strong evidence of positive selection from low diversity, skewed allele-frequency spectra, and the predominance of a single haplotype. We suggest that the inactive form of the gene arose in Africa approximately 100-500 thousand years ago (KYA) and was initially neutral or almost neutral but that positive selection beginning approximately 60-100 KYA drove it to near fixation. We further propose that its selective advantage was sepsis resistance in populations that experienced more infectious diseases as population sizes and densities increased.

Base Sequence↗

Genetic analysis of completely sequenced disease-associated MHC haplotypes identifies shuffling of segments in recent human history.

The major histocompatibility complex (MHC) is recognised as one of the most important genetic regions in relation to common human disease. Advancement in identification of MHC genes that confer susceptibility to disease requires greater knowledge of sequence variation across the complex. Highly duplicated and polymorphic regions of the human genome such as the MHC are, however, somewhat refractory to some whole-genome analysis methods. To address this issue, we are employing a bacterial artificial chromosome (BAC) cloning strategy to sequence entire MHC haplotypes from consanguineous cell lines as part of the MHC Haplotype Project. Here we present 4.25 Mb of the human haplotype QBL (HLA-A26-B18-Cw5-DR3-DQ2) and compare it with the MHC reference haplotype and with a second haplotype, COX (HLA-A1-B8-Cw7-DR3-DQ2), that shares the same HLA-DRB1, -DQA1, and -DQB1 alleles. We have defined the complete gene, splice variant, and sequence variation contents of all three haplotypes, comprising over 259 annotated loci and over 20,000 single nucleotide polymorphisms (SNPs). Certain coding sequences vary significantly between different haplotypes, making them candidates for functional and disease-association studies. Analysis of the two DR3 haplotypes allowed delineation of the shared sequence between two HLA class II-related haplotypes differing in disease associations and the identification of at least one of the sites that mediated the original recombination event. The levels of variation across the MHC were similar to those seen for other HLA-disparate haplotypes, except for a 158-kb segment that contained the HLA-DRB1, -DQA1, and -DQB1 genes and showed very limited polymorphism compatible with identity-by-descent and relatively recent common ancestry (<3,400 generations). These results indicate that the differential disease associations of these two DR3 haplotypes are due to sequence variation outside this central 158-kb segment, and that shuffling of ancestral blocks via recombination is a potential mechanism whereby certain DR-DQ allelic combinations, which presumably have favoured immunological functions, can spread across haplotypes and populations.

Chromosome Mapping↗

A genome-wide, end-sequenced 129Sv BAC library resource for targeting vector construction.

The majority of gene-targeting experiments in mice are performed in 129Sv-derived embryonic stem (ES) cell lines, which are generally considered to be more reliable at colonizing the germ line than ES cells derived from other strains. Gene targeting is reliant on homologous recombination of a targeting vector with the host ES cell genome. The efficiency of recombination is affected by many factors, including the isogenicity (H. te Riele et al., 1992, Proc. Natl. Acad. Sci. USA 89, 5128-5132) and the length of homologous sequence of the targeting vector and the location of the target locus. Here we describe the double-end sequencing and mapping of 84,507 bacterial artificial chromosomes (BACs) generated from AB2.2 ES cell DNA (129S7/SvEvBrd-Hprtb-m2). We have aligned these BACs against the mouse genome and displayed them on the Ensembl genome browser, DAS: 129S7/AB2.2. This library has an average insert size of 110.68 kb and average depth of genome coverage of 3.63- and 1.24-fold across the autosomes and sex chromosomes, respectively. Over 97% of the mouse genome and 99.1% of Ensembl genes are covered by clones from this library. This publicly available BAC resource can be used for the rapid construction of targeting vectors via recombineering. Furthermore, we show that targeting vectors containing DNA recombineered from this BAC library can be used to target genes efficiently in several 129-derived ES cell lines.

Animals↗

Seasonally hibernating phenotype assessed through transcript screening.

Hibernation is a seasonally entrained and profound phenotypic transition to conserve energy in winter. It involves significant biochemical reprogramming, although our understanding of the underpinning molecular events is fragmentary and selective. We have conducted a large-scale gene expression screen of the golden-mantled ground squirrel, Spermophilus lateralis, to identify transcriptional responses associated specifically with the summer-winter transition and the torpid-arousal transition in winter. We used 112 cDNA microarrays comprising 12,288 probes that cover at least 5,109 genes. In liver, the profiles of torpid and active states in the winter were almost identical, although we identified 102 cDNAs that were differentially expressed between winter and summer, 90% of which were downregulated in the winter states. By contrast, in cardiac tissue, 59 and 115 cDNAs were elevated in interbout arousal and torpor, respectively, relative to the summer active condition, but only 7 were common to both winter states, and during arousal none was downregulated. In brain, 78 cDNAs were found to change in winter, 44 of which were upregulated. Thus transcriptional changes associated with hibernation are qualitatively modest and, since these changes are generally less than twofold, also quantitatively modest. Unbiased Gene Ontology profiling of the transcripts suggests a winter switch to beta-oxidation of lipids in liver and heart, a reduction in metabolism of toxic compounds and the urea cycle in liver, and downregulated electron transport in the brain. We identified just one strongly winter-induced transcript common to all tissues, namely an RNA-binding protein, RBM3. This analysis clearly differentiates responses of the principal tissues, identifies a large number of new genes undergoing regulation, and broadens our understanding of affected cellular processes that, in part, account for the winter-adaptive hibernating phenotype.

Animals↗

Nodulation signaling in legumes requires NSP2, a member of the GRAS family of transcriptional regulators.

Rhizobial bacteria enter a symbiotic interaction with legumes, activating diverse responses in roots through the lipochito oligosaccharide signaling molecule Nod factor. Here, we show that NSP2 from Medicago truncatula encodes a GRAS protein essential for Nod-factor signaling. NSP2 functions downstream of Nod-factor-induced calcium spiking and a calcium/calmodulin-dependent protein kinase. We show that NSP2-GFP expressed from a constitutive promoter is localized to the endoplasmic reticulum/nuclear envelope and relocalizes to the nucleus after Nod-factor elicitation. This work provides evidence that a GRAS protein transduces calcium signals in plants and provides a possible regulator of Nod-factor-inducible gene expression.

Amino Acid Motifs↗

Gene finding in the chicken genome.

BACKGROUND: Despite the continuous production of genome sequence for a number of organisms, reliable, comprehensive, and cost effective gene prediction remains problematic. This is particularly true for genomes for which there is not a large collection of known gene sequences, such as the recently published chicken genome. We used the chicken sequence to test comparative and homology-based gene-finding methods followed by experimental validation as an effective genome annotation method. RESULTS: We performed experimental evaluation by RT-PCR of three different computational gene finders, Ensembl, SGP2 and TWINSCAN, applied to the chicken genome. A Venn diagram was computed and each component of it was evaluated. The results showed that de novo comparative methods can identify up to about 700 chicken genes with no previous evidence of expression, and can correctly extend about 40% of homology-based predictions at the 5' end. CONCLUSIONS: De novo comparative gene prediction followed by experimental verification is effective at enhancing the annotation of the newly sequenced genomes provided by standard homology-based methods.

Animals↗

Complex haplotypes, copy number polymorphisms and coding variation in two recently divergent mouse strains.

Inbred mouse strains provide the foundation for mouse genetics. By selecting for phenotypic features of interest, inbreeding drives genomic evolution and eliminates individual variation, while fixing certain sets of alleles that are responsible for the trait characteristics of the strain. Mouse strains 129Sv (129S5) and C57BL/6J, two of the most widely used inbred lines, diverged from common ancestors within the last century, yet very little is known about the genomic differences between them. By comparative genomic hybridization and sequence analysis of 129S5 short insert libraries, we identified substantial structural variation, a complex fine-scale haplotype pattern with a continuous distribution of diversity blocks, and extensive nucleotide variation, including nonsynonymous coding SNPs and stop codons. Collectively, these genomic changes denote the level and direction of allele fixation that has occurred during inbreeding and provide a basis for defining what makes these mouse strains unique.

Animals↗

Natural genetic variants influencing type 1 diabetes in humans and in the NOD mouse.

The understanding of the genetic basis of type 1 diabetes and other autoimmune diseases and the application of that knowledge to their treatment, cure and eventual prevention has been a difficult goal to reach. Cumulative progress in both mouse and human are finally giving way to some successes and significant insights have been made in the last few years. Investigators have identified key immune tolerance-associated phenotypes in convincingly reliable ways that are regulated by specific diabetes-associated chromosomal intervals. The combination of positional genetics and functional studies is a powerful approach to the identification of downstream molecular events that are causal in disease aetiology. In the case of type 1 diabetes, the availability of several animal models, especially the NOD mouse, has complemented the efforts to localize human genes causing diabetes and has shown that some of the same genes and pathways are associated with autoimmunity in both species. There is also growing evidence that the initiation or progression of many autoimmune diseases is likely to be influenced by some of the same genes.

Animals↗

Transcriptome analysis for the chicken based on 19,626 finished cDNA sequences and 485,337 expressed sequence tags.

We present an analysis of the chicken (Gallus gallus) transcriptome based on the full insert sequences for 19,626 cDNAs, combined with 485,337 EST sequences. The cDNA data set has been functionally annotated and describes a minimum of 11,929 chicken coding genes, including the sequence for 2260 full-length cDNAs together with a collection of noncoding (nc) cDNAs that have been stringently filtered to remove untranslated regions of coding mRNAs. The combined collection of cDNAs and ESTs describe 62,546 clustered transcripts and provide transcriptional evidence for a total of 18,989 chicken genes, including 88% of the annotated Ensembl gene set. Analysis of the ncRNAs reveals a set that is highly conserved in chickens and mammals, including sequences for 14 pri-miRNAs encoding 23 different miRNAs. The data sets described here provide a transcriptome toolkit linked to physical clones for bioinformaticians and experimental biologists who wish to use chicken systems as a low-cost, accessible alternative to mammals for the analysis of vertebrate development, immunology, and cell biology.

Animals↗

Coping with cold: An integrative, multitissue analysis of the transcriptome of a poikilothermic vertebrate.

How do organisms respond adaptively to environmental stress? Although some gene-specific responses have been explored, others remain to be identified, and there is a very poor understanding of the system-wide integration of response, particularly in complex, multitissue animals. Here, we adopt a transcript screening approach to explore the mechanisms underpinning a major, whole-body phenotypic transition in a vertebrate animal that naturally experiences extreme environmental stress. Carp were exposed to increasing levels of cold, and responses across seven tissues were assessed by using a microarray composed of 13,440 cDNA probes. A large set of unique cDNAs (approximately 3,400) were affected by cold. These cDNAs included an expression signature common to all tissues of 252 up-regulated genes involved in RNA processing, translation initiation, mitochondrial metabolism, proteasomal function, and modification of higher-order structures of lipid membranes and chromosomes. Also identified were large numbers of transcripts with highly tissue-specific patterns of regulation. By unbiased profiling of gene ontologies, we have identified the distinctive functional features of each tissue's response and integrate them into a comprehensive view of the whole-body transition from one strongly adaptive phenotype to another. This approach revealed an expression signature suggestive of atrophy in cooled skeletal muscle. This environmental genomics approach by using a well studied but nongenomic species has identified a range of candidate genes endowing thermotolerance and reveals a previously unrecognized scale and complexity of responses that impacts at the level of cellular and tissue function.

Adaptation, Physiological↗

Organization and evolution of a gene-rich region of the mouse genome: a 12.7-Mb region deleted in the Del(13)Svea36H mouse.

Del(13)Svea36H (Del36H) is a deletion of approximately 20% of mouse chromosome 13 showing conserved synteny with human chromosome 6p22.1-6p22.3/6p25. The human region is lost in some deletion syndromes and is the site of several disease loci. Heterozygous Del36H mice show numerous phenotypes and may model aspects of human genetic disease. We describe 12.7 Mb of finished, annotated sequence from Del36H. Del36H has a higher gene density than the draft mouse genome, reflecting high local densities of three gene families (vomeronasal receptors, serpins, and prolactins) which are greatly expanded relative to human. Transposable elements are concentrated near these gene families. We therefore suggest that their neighborhoods are gene factories, regions of frequent recombination in which gene duplication is more frequent. The gene families show different proportions of pseudogenes, likely reflecting different strengths of purifying selection and/or gene conversion. They are also associated with relatively low simple sequence concentrations, which vary across the region with a periodicity of approximately 5 Mb. Del36H contains numerous evolutionarily conserved regions (ECRs). Many lie in noncoding regions, are detectable in species as distant as Ciona intestinalis, and therefore are candidate regulatory sequences. This analysis will facilitate functional genomic analysis of Del36H and provides insights into mouse genome evolution.

Animals↗

Mutagenic insertion and chromosome engineering resource (MICER).

Embryonic stem cell technology revolutionized biology by providing a means to assess mammalian gene function in vivo. Although it is now routine to generate mice from embryonic stem cells, one of the principal methods used to create mutations, gene targeting, is a cumbersome process. Here we describe the indexing of 93,960 ready-made insertional targeting vectors from two libraries. 5,925 of these vectors can be used directly to inactivate genes with an average targeting efficiency of 28%. Combinations of vectors from the two libraries can be used to disrupt both alleles of a gene or engineer larger genomic changes such as deletions, duplications, translocations or inversions. These indexed vectors constitute a public resource (Mutagenic Insertion and Chromosome Engineering Resource; MICER) for high-throughput, targeted manipulation of the mouse genome.

Animals↗

Fine mapping, gene content, comparative sequencing, and expression analyses support Ctla4 and Nramp1 as candidates for Idd5.1 and Idd5.2 in the nonobese diabetic mouse.

At least two loci that determine susceptibility to type 1 diabetes in the NOD mouse have been mapped to chromosome 1, Idd5.1 (insulin-dependent diabetes 5.1) and Idd5.2. In this study, using a series of novel NOD.B10 congenic strains, Idd5.1 has been defined to a 2.1-Mb region containing only four genes, Ctla4, Icos, Als2cr19, and Nrp2 (neuropilin-2), thereby excluding a major candidate gene, Cd28. Genomic sequence comparison of the two functional candidate genes, Ctla4 and Icos, from the B6 (resistant at Idd5.1) and the NOD (susceptible at Idd5.1) strains revealed 62 single nucleotide polymorphisms (SNPs), only two of which were in coding regions. One of these coding SNPs, base 77 of Ctla4 exon 2, is a synonymous SNP and has been correlated previously with type 1 diabetes susceptibility and differential expression of a CTLA-4 isoform. Additional expression studies in this work support the hypothesis that this SNP in exon 2 is the genetic variation causing the biological effects of Idd5.1. Analysis of additional congenic strains has also localized Idd5.2 to a small region (1.52 Mb) of chromosome 1, but in contrast to the Idd5.1 interval, Idd5.2 contains at least 45 genes. Notably, the Idd5.2 region still includes the functionally polymorphic Nramp1 gene. Future experiments to test the identity of Idd5.1 and Idd5.2 as Ctla4 and Nramp1, respectively, can now be justified using approaches to specifically alter or mimic the candidate causative SNPs.

Amino Acid Sequence↗

Complete MHC haplotype sequencing for common disease gene mapping.

The future systematic mapping of variants that confer susceptibility to common diseases requires the construction of a fully informative polymorphism map. Ideally, every base pair of the genome would be sequenced in many individuals. Here, we report 4.75 Mb of contiguous sequence for each of two common haplotypes of the major histocompatibility complex (MHC), to which susceptibility to >100 diseases has been mapped. The autoimmune disease-associated-haplotypes HLA-A3-B7-Cw7-DR15 and HLA-A1-B8-Cw7-DR3 were sequenced in their entirety through a bacterial artificial chromosome (BAC) cloning strategy using the consanguineous cell lines PGF and COX, respectively. The two sequences were annotated to encompass all described splice variants of expressed genes. We defined the complete variation content of the two haplotypes, revealing >18,000 variations between them. Average SNP densities ranged from less than one SNP per kilobase to >60. Acquisition of complete and accurate sequence data over polymorphic regions such as the MHC from large-insert cloned DNA provides a definitive resource for the construction of informative genetic maps, and avoids the limitation of chromosome regions that are refractory to PCR amplification.

Autoimmune Diseases↗