PubMed Health⌕ Search

Biomedical subjects

Miguel A Andrade

Publications and source records attributed to Miguel A Andrade.

At least 19 recordsLinked to original sources

A mitochondria-K+ channel axis is suppressed in cancer and its normalization promotes apoptosis and inhibits cancer growth.

The unique metabolic profile of cancer (aerobic glycolysis) might confer apoptosis resistance and be therapeutically targeted. Compared to normal cells, several human cancers have high mitochondrial membrane potential (DeltaPsim) and low expression of the K+ channel Kv1.5, both contributing to apoptosis resistance. Dichloroacetate (DCA) inhibits mitochondrial pyruvate dehydrogenase kinase (PDK), shifts metabolism from glycolysis to glucose oxidation, decreases DeltaPsim, increases mitochondrial H2O2, and activates Kv channels in all cancer, but not normal, cells; DCA upregulates Kv1.5 by an NFAT1-dependent mechanism. DCA induces apoptosis, decreases proliferation, and inhibits tumor growth, without apparent toxicity. Molecular inhibition of PDK2 by siRNA mimics DCA. The mitochondria-NFAT-Kv axis and PDK are important therapeutic targets in cancer; the orally available DCA is a promising selective anticancer agent.

Animals↗

Human Lsg1 defines a family of essential GTPases that correlates with the evolution of compartmentalization.

BACKGROUND: Compartmentalization is a key feature of eukaryotic cells, but its evolution remains poorly understood. GTPases are the oldest enzymes that use nucleotides as substrates and they participate in a wide range of cellular processes. Therefore, they are ideal tools for comparative genomic studies aimed at understanding how aspects of biological complexity such as cellular compartmentalization evolved. RESULTS: We describe the identification and characterization of a unique family of circularly permuted GTPases represented by the human orthologue of yeast Lsg1p. We placed the members of this family in the phylogenetic context of the YlqF Related GTPase (YRG) family, which are present in Eukarya, Bacteria and Archea and include the stem cell regulator Nucleostemin. To extend the computational analysis, we showed that hLsg1 is an essential GTPase predominantly located in the endoplasmic reticulum and, in some cells, in Cajal bodies in the nucleus. Comparison of localization and siRNA datasets suggests that all members of the family are essential GTPases that have increased in number as the compartmentalization of the eukaryotic cell and the ribosome biogenesis pathway have evolved. CONCLUSION: We propose a scenario, consistent with our data, for the evolution of this family: cytoplasmic components were first acquired, followed by nuclear components, and finally the mitochondrial and chloroplast elements were derived from different bacterial species, in parallel with the formation of the nucleolus and the specialization of nuclear components.

Cell Nucleus↗

G2D: a tool for mining genes associated with disease.

BACKGROUND: Human inherited diseases can be associated by genetic linkage with one or more genomic regions. The availability of the complete sequence of the human genome allows examining those locations for an associated gene. We previously developed an algorithm to prioritize genes on a chromosomal region according to their possible relation to an inherited disease using a combination of data mining on biomedical databases and gene sequence analysis. RESULTS: We have implemented this method as a web application in our site G2D (Genes to Diseases). It allows users to inspect any region of the human genome to find candidate genes related to a genetic disease of their interest. In addition, the G2D server includes pre-computed analyses of candidate genes for 552 linked monogenic diseases without an associated gene, and the analysis of 18 asthma loci. CONCLUSION: G2D can be publicly accessed at http://www.ogic.ca/projects/g2d_2/.

Algorithms↗

Inconsistencies over time in 5% of NetAffx probe-to-gene annotations.

BACKGROUND: DNA microarray probes are designed to match particular mRNA transcripts, often based on expressed sequences like ESTs, or cDNAs, many times incomplete. As a result, the relations between probes and genes can change as the sequence data are updated. However, it is frequent that the reported results of microarray analyses are given just as lists of genes without any reference to the underlying probes. RESULTS: We show for a particular commercial microarray design that the number of probes associated to some genes change with time. These changes concern approximately 5% of the probe sets across the history of annotation releases over a two year span. CONCLUSION: We recommend to report probe set identifiers when publishing microarray results, and to submit those analyses to microarray public databases to ensure that the interpretation of the data is updated with the latest set of annotations.

Animals↗

Systematic association of genes to phenotypes by genome and literature mining.

One of the major challenges of functional genomics is to unravel the connection between genotype and phenotype. So far no global analysis has attempted to explore those connections in the light of the large phenotypic variability seen in nature. Here, we use an unsupervised, systematic approach for associating genes and phenotypic characteristics that combines literature mining with comparative genome analysis. We first mine the MEDLINE literature database for terms that reflect phenotypic similarities of species. Subsequently we predict the likely genomic determinants: genes specifically present in the respective genomes. In a global analysis involving 92 prokaryotic genomes we retrieve 323 clusters containing a total of 2,700 significant gene-phenotype associations. Some clusters contain mostly known relationships, such as genes involved in motility or plant degradation, often with additional hypothetical proteins associated with those phenotypes. Other clusters comprise unexpected associations; for example, a group of terms related to food and spoilage is linked to genes predicted to be involved in bacterial food poisoning. Among the clusters, we observe an enrichment of pathogenicity-related associations, suggesting that the approach reveals many novel genes likely to play a role in infectious diseases.

Bacteria↗

Ranking the whole MEDLINE database according to a large training set using text indexing.

BACKGROUND: The MEDLINE database contains over 12 million references to scientific literature, with about 3/4 of recent articles including an abstract of the publication. Retrieval of entries using queries with keywords is useful for human users that need to obtain small selections. However, particular analyses of the literature or database developments may need the complete ranking of all the references in the MEDLINE database as to their relevance to a topic of interest. This report describes a method that does this ranking using the differences in word content between MEDLINE entries related to a topic and the whole of MEDLINE, in a computational time appropriate for an article search query engine. RESULTS: We tested the capabilities of our system to retrieve MEDLINE references which are relevant to the subject of stem cells. We took advantage of the existing annotation of references with terms from the MeSH hierarchical vocabulary (Medical Subject Headings, developed at the National Library of Medicine). A training set of 81,416 references was constructed by selecting entries annotated with the MeSH term stem cells or some child in its sub tree. Frequencies of all nouns, verbs, and adjectives in the training set were computed and the ratios of word frequencies in the training set to those in the entire MEDLINE were used to score references. Self-consistency of the algorithm, benchmarked with a test set containing the training set and an equal number of references randomly selected from MEDLINE was better using nouns (79%) than adjectives (73%) or verbs (70%). The evaluation of the system with 6,923 references not used for training, containing 204 articles relevant to stem cells according to a human expert, indicated a recall of 65% for a precision of 65%. CONCLUSION: This strategy appears to be useful for predicting the relevance of MEDLINE references to a given concept. The method is simple and can be used with any user-defined training set. Choice of the part of speech of the words used for classification has important effects on performance. Lists of words, scripts, and additional information are available from the web address http://www.ogic.ca/projects/ks2004/.

Abstracting and Indexing↗

Study of stem cell function using microarray experiments.

DNA Microarrays are used to simultaneously measure the levels of thousands of mRNAs in a sample. We illustrate here that a collection of such measurements in different cell types and states is a sound source of functional predictions, provided the microarray experiments are analogous and the cell samples are appropriately diverse. We have used this approach to study stem cells, whose identity and mechanisms of control are not well understood, generating Affymetrix microarray data from more than 200 samples, including stem cells and their derivatives, from human and mouse. The data can be accessed online (StemBase; http://www.scgp.ca:8080/StemBase/).

Animals↗

The Shwachman-Bodian-Diamond syndrome protein family is involved in RNA metabolism.

A combination of structural, biochemical, and genetic studies in model organisms was used to infer a cellular role for the human protein (SBDS) responsible for Shwachman-Bodian-Diamond syndrome. The crystal structure of the SBDS homologue in Archaeoglobus fulgidus, AF0491, revealed a three domain protein. The N-terminal domain, which harbors the majority of disease-linked mutations, has a novel three-dimensional fold. The central domain has the common winged helix-turn-helix motif, and the C-terminal domain shares structural homology with known RNA-binding domains. Proteomic analysis of the SBDS sequence homologue in Saccharomyces cerevisiae, YLR022C, revealed an association with over 20 proteins involved in ribosome biosynthesis. NMR structural genomics revealed another yeast protein, YHR087W, to be a structural homologue of the AF0491 N-terminal domain. Sequence analysis confirmed them as distant sequence homologues, therefore related by divergent evolution. Synthetic genetic array analysis of YHR087W revealed genetic interactions with proteins involved in RNA and rRNA processing including Mdm20/Nat3, Nsr1, and Npl3. Our observations, taken together with previous reports, support the conclusion that SBDS and its homologues play a role in RNA metabolism.

Acetyltransferases↗

ACRATA: a novel electron transfer domain associated to apoptosis and cancer.

BACKGROUND: Recently, several members of a vertebrate protein family containing a six trans-membrane (6TM) domain and involved in apoptosis and cancer (e.g. STEAP, STAMP1, TSAP6), have been identified in Golgi and cytoplasmic membranes. The exact function of these proteins remains unknown. METHODS: We related this 6TM domain to distant protein families using intermediate sequences and methods of iterative profile sequence similarity search. RESULTS: Here we show for the first time that this 6TM domain is homolog to the 6TM heme binding domain of both the NADPH oxidase (Nox) family and the YedZ family of bacterial oxidoreductases. CONCLUSIONS: This finding gives novel insights about the existence of a previously undetected electron transfer system involved in apoptosis and cancer, and suggests further steps in the experimental characterization of these evolutionarily related families.

Adaptor Proteins, Signal Transducing↗

Gene annotation from scientific literature using mappings between keyword systems.

MOTIVATION: The description of genes in databases by keywords helps the non-specialist to quickly grasp the properties of a gene and increases the efficiency of computational tools that are applied to gene data (e.g. searching a gene database for sequences related to a particular biological process). However, the association of keywords to genes or protein sequences is a difficult process that ultimately implies examination of the literature related to a gene. RESULTS: To support this task, we present a procedure to derive keywords from the set of scientific abstracts related to a gene. Our system is based on the automated extraction of mappings between related terms from different databases using a model of fuzzy associations that can be applied with all generality to any pair of linked databases. We tested the system by annotating genes of the SWISS-PROT database with keywords derived from the abstracts linked to their entries (stored in the MEDLINE database of scientific references). The performance of the annotation procedure was much better for SWISS-PROT keywords (recall of 47%, precision of 68%) than for Gene Ontology terms (recall of 8%, precision of 67%). AVAILABILITY: The algorithm can be publicly accessed and used for the annotation of sequences through a web server at http://www.bork.embl.de/kat

Abstracting and Indexing↗

Global analysis of bacterial transcription factors to predict cellular target processes.

Whole-genome sequences are now available for >100 bacterial species, giving unprecedented power to comparative genomics approaches. We have applied genome-context methods to predict target processes that are regulated by transcription factors (TFs). Of 128 orthologous groups of proteins annotated as TFs, to date, 36 are functionally uncharacterized; in our analysis we predict a probable cellular target process or biochemical pathway for half of these functionally uncharacterized TFs.

Bacteria↗

Update on XplorMed: A web server for exploring scientific literature.

As scientific literature databases like MEDLINE increase in size, so does the time required to search them. Scientists must frequently inspect long lists of references manually, often just reading the titles. XplorMed is a web tool that aids MEDLINE searching by summarizing the subjects contained in the results, thus allowing users to focus on subjects of interest. Here we describe new features added to XplorMed during the last 2 years (http://www.bork.embl-heidelberg.de/xplormed/).

Bibliography of Medicine↗

Information extraction from full text scientific articles: where are the keywords?

BACKGROUND: To date, many of the methods for information extraction of biological information from scientific articles are restricted to the abstract of the article. However, full text articles in electronic version, which offer larger sources of data, are currently available. Several questions arise as to whether the effort of scanning full text articles is worthy, or whether the information that can be extracted from the different sections of an article can be relevant. RESULTS: In this work we addressed those questions showing that the keyword content of the different sections of a standard scientific article (abstract, introduction, methods, results, and discussion) is very heterogeneous. CONCLUSIONS: Although the abstract contains the best ratio of keywords per total of words, other sections of the article may be a better source of biologically relevant data.

Anatomy↗

Evaluation of annotation strategies using an entire genome sequence.

MOTIVATION: Genome-wide functional annotation either by manual or automatic means has raised considerable concerns regarding the accuracy of assignments and the reproducibility of methodologies. In addition, a performance evaluation of automated systems that attempt to tackle sequence analyses rapidly and reproducibly is generally missing. In order to quantify the accuracy and reproducibility of function assignments on a genome-wide scale, we have re-annotated the entire genome sequence of Chlamydia trachomatis (serovar D), in a collaborative manner. RESULTS: We have encoded all annotations in a structured format to allow further comparison and data exchange and have used a scale that records the different levels of potential annotation errors according to their propensity to propagate in the database due to transitive function assignments. We conclude that genome annotation may entail a considerable amount of errors, ranging from simple typographical errors to complex sequence analysis problems. The most surprising result of this comparative study is that automatic systems might perform as well as the teams of experts annotating genome sequences.

Amino Acid Sequence↗

The way we write.

Explore the source record for details and available documents.

Biomedical Research↗

A protocol for the update of references to scientific literature in biological databases.

Entries in biological databases are usually linked to scientific references. To generate those links and to keep them up-to-date, database maintainers have to continuously scan the scientific literature to select references that are relevant for each single database entry. The continuous growth of both the corpus of scientific literature and the size of biological databases makes this task very hard. We present a protocol intended to assist the updating of an existing set of literature (abstract) links from a single database entry with new references. It consists of taking the set of MEDLINE neighbour references of the existing linked abstracts and evaluating their relevance according to the existing set of abstracts. To test the applicability of the algorithm, we did a simple benchmark of the system using the references associated with the entries of a protein domain database. Human experts found the references that the algorithm scored highly were more relevant to the database entry than those scored lowly, suggesting that the algorithm was useful.

Abstracting and Indexing↗

NEAT: a domain duplicated in genes near the components of a putative Fe3+ siderophore transporter from Gram-positive pathogenic bacteria.

BACKGROUND: Iron uptake from the host is essential for bacteria that infect animals. To find potential targets for drugs active against pathogenic bacteria, we have searched all completely sequenced genomes of pathogenic bacteria for genes relevant for iron transport. RESULTS: We identified a protein domain that appears in variable copy number in bacterial genes that are usually in the vicinity of a putative Fe3+ siderophore transporter. Accordingly, we have denoted this domain NEAT for 'near transporter'. Most of the bacterial species containing this domain are pathogenic. Sequence features indicate that the domain is anchored to the extracellular side of the membrane. The domain seems to be under high selective pressure for rapid independent duplications that are typical of sequences involved in signaling and binding. CONCLUSIONS: The NEAT domain might be functionally related to iron transport. The taxonomic specificity of this domain and its predicted extracellular position could make it an interesting target for designing new drugs against some highly pathogenic bacteria.

Amino Acid Sequence↗