Program description: Strategies for biological annotation of mammalian systems: implementing gene ontologies in mouse genome informatics.
Explore the source record for details and available documents.
SEARCH · PubMed Health
Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
Explore the source record for details and available documents.
Human chromosome 2q33 is an immunologically important region based on the linkage of numerous autoimmune diseases to the CTLA4 locus. Here, we sequenced and assembled 2q33 bacterial artificial chromosome (BAC) clones, resulting in 381,403 bp of contiguous sequence containing genes encoding a NADH: ubiquinone oxidoreductase, the costimulatory receptors CD28, CTLA4, and ICOS, and a HERV-H type endogenous retrovirus located 366 bp downstream of ICOS in the reverse orientation. Genomic microarray expression analysis using differentially activated T-cell RNA against a subcloned CTLA4/ICOS BAC library revealed upregulation of CTLA4 and ICOS sequences, plus antisense ICOS transcripts generated by the HERV-H, suggesting a potential mechanism for ICOS regulation. We identified four nonlinked, polymorphic, simple repetitive sequence elements in this region, which may be used to delineate genetic effects of ICOS and CTLA4 in disease populations. Comparative genomic analysis of mouse genomic Icos sequences revealed 60% sequence identity in the 5' UTR and regions between exon 2 and the 3' UTR, suggesting the importance of ICOS gene function.
A method (three-dimensional position-specific scoring matrix, 3D-PSSM) to recognise remote protein sequence homologues is described. The method combines the power of multiple sequence profiles with knowledge of protein structure to provide enhanced recognition and thus functional assignment of newly sequenced genomes. The method uses structural alignments of homologous proteins of similar three-dimensional structure in the structural classification of proteins (SCOP) database to obtain a structural equivalence of residues. These equivalences are used to extend multiply aligned sequences obtained by standard sequence searches. The resulting large superfamily-based multiple alignment is converted into a PSSM. Combined with secondary structure matching and solvation potentials, 3D-PSSM can recognise structural and functional relationships beyond state-of-the-art sequence methods. In a cross-validated benchmark on 136 homologous relationships unambiguously undetectable by position-specific iterated basic local alignment search tool (PSI-Blast), 3D-PSSM can confidently assign 18 %. The method was applied to the remaining unassigned regions of the Mycoplasma genitalium genome and an additional 13 regions were assigned with 95 % confidence. 3D-PSSM is available to the community as a web server: http://www.bmm.icnet.uk/servers/3dpssm
In order to circumvent limitations of sequence based methods in the process of making functional predictions for proteins, we have developed a methodology that uses a sequence-to-structure-to-function paradigm. First, an approximate three-dimensional structure is predicted. Then, a three-dimensional descriptor of the functional site, termed a Fuzzy Functional Form, or FFF, is used to screen the structure for the presence of the functional site of interest (Fetrow et al., 1998; Fetrow and Skolnick, 1998). Previously, a disulfide oxidoreductase FFF was developed and applied to predicted structures obtained from a small structural database. Here, using a substantially larger structural database, we expand the analysis of the disulfide oxidoreductase FFF to the B. subtilis genome. To ascertain the performance of the FFF, its results are compared to those obtained using both the sequence alignment method BLAST and three local sequence motif databases: PRINTS, Prosite, and Blocks. The FFF method is then compared in detail to Blocks and it is shown that the FFF is more flexible and sensitive in finding a specific function in a set of unknown proteins. In addition, the estimated false positive rate of function prediction is significantly lower using the FFF structural motif, rather than the standard sequence motif methods. We also present a second FFF and describe a specific example of the results of its whole-genome application to D. melanogaster using a newer threading algorithm. Our results from all of these studies indicate that the addition of three-dimensional structural information adds significant value in the prediction of biochemical function of genomic sequences.
Explore the source record for details and available documents.
Management and analysis of the huge amounts of data produced by microarray experiments is becoming one of the major bottlenecks in the utilization of this high-throughput technology. We describe the basic design of a microarray gene expression database to help microarray users and their informatics teams to set up their information services. We describe two data models--a simpler one called ArrayExpressB and the complete model ArrayExpressC, and discuss some implementation issues. For latest developments see http: wwwebi.ac.uk/arrayexpress
Explore the source record for details and available documents.
Explore the source record for details and available documents.
Explore the source record for details and available documents.
Few clinicians would doubt the importance of obtaining smoking histories from their patients. Nevertheless a clinicians' practices of documenting tobacco use in medical records varies substantially among individual providers and between different health care systems. To investigate the chart documentation of patient smoking among Indian Health Service clinicians, we reviewed 545 randomly selected patient records from 22 different Indian Health Service affiliated clinics. We focused on differences in charting of tobacco use by type of clinic and by geographic area within the Indian Health Service. Documentation varied by area, ranging from no documentation in the Albuquerque, Navajo, and Phoenix areas to 51% in the Oklahoma area. We conclude that clinicians practices of documentation of tobacco use vary widely and recommend that this practice be more widely encouraged at all affiliated Indian Health Service clinics.
Explore the source record for details and available documents.
Explore the source record for details and available documents.
Explore the source record for details and available documents.
Explore the source record for details and available documents.
Heinrich Simon Frenkel or Frenkel-Heiden(1860-1931) is almost completely forgotten as a founder of neurorehabilitation and little is known about his life. Frenkel's main contribution, "The treatment of tabetic ataxia by meansof systematic exercise: An exposition of the principles and practice of compensatory movement treatment", was reprinted several times in English (1902, 1905, 1917). Frenkel exerted great influence among his contemporaries, including his direct student Otfrid Foerster (1873-1941) who became one of the most important neurologists and neurosurgeons of the 20th century. A floor mosaic, preserved in the historic building of the "Medizinische Poliklinik" in Munich, is an exact copy of the pattern of traces that Frenkel had published in 1900 for proprioceptive gait exercises in tabes dorsalis.
Expression profiling offers a potential high-throughput phenotype screen for mutant mouse embryonic stem (ES) cells. We have assessed the ability of expression arrays to distinguish among heterozygous mutant ES cell lines and to accurately reflect the normal function of the mutated genes. Two ES cell lines hemizygous for overlapping regions of mouse Chromosome (Chr) 5 differed substantially from the wildtype parental line and from each other. Expression differences included frequent downregulation of hemizygous genes and downstream effects on genes mapping to other chromosomes. Some genes were affected similarly in each deletion line, consistent with the overlap of the deletions. To determine whether such downstream effects reveal pathways impacted by a mutation, we examined ES cell lines heterozygous for mutations in either of two well-characterized genes. A heterozygous mutation in the gene encoding the cell cycle regulator, cyclin D kinase 4 ( Cdk4), affected expression of many genes involved in cell growth and proliferation. A heterozygous mutation in the ATP binding cassette transporter family A, member 1 ( Abca1) gene, altered genes associated with lipid homeostasis, the cytoskeleton, and vesicle trafficking. Heterozygous Abca1 mutation had similar effects in liver, indicating that ES cell expression profile reflects changes in fundamental processes relevant to mutant gene function in multiple cell types.
We have compiled a database of mitochondrial DNA (mtDNA) control region, hypervariable regions 1 (HVR1) and 2 (HVR2) sequences of a total of 14,138 individuals compiled from 103 mtDNA publications before 1 January 2000, 13 data sets published in 2000 and 2001 and 2 unpublished data sets of Iraqi Kurds and Indians from Kerala. By contacting the authors and by other means, we have confirmed and corrected sequence errors, eliminated duplications and harmonised the sequence format. These changes affected all but 26 of the 116 publications. Furthermore, we have implemented a geographic information system ("mtradius") which searches for closest matches to a given mtDNA control region sequence and displays them on a geographic map. A potential application is to estimate a chance matching probability when a forensic stain and a suspect have an identical mtDNA sequence: we suggest that the geographic area with the highest frequency of closely related mtDNA sequence types may be used to define a reference population to give the suspect the maximum benefit of doubt in accordance with the ceiling principle.