PubMed Health⌕ Search

Biomedical subjects

P Hraber

Publications and source records attributed to P Hraber.

3 recordsLinked to original sources

The phytophthora genome initiative database: informatics and analysis for distributed pathogenomic research.

The Phytophthora Genome Initiative (PGI) is a distributed collaboration to study the genome and evolution of a particularly destructive group of plant pathogenic oomycete, with the goal of understanding the mechanisms of infection and resistance. NCGR provides informatics support for the collaboration as well as a centralized data repository. In the pilot phase of the project, several investigators prepared Phytophthora infestans and Phytophthora sojae EST and Phytophthora sojae BAC libraries and sent them to another laboratory for sequencing. Data from sequencing reactions were transferred to NCGR for analysis and curation. An analysis pipeline transforms raw data by performing simple analyses (i.e., vector removal and similarity searching) that are stored and can be retrieved by investigators using a web browser. Here we describe the database and access tools, provide an overview of the data therein and outline future plans. This resource has provided a unique opportunity for the distributed, collaborative study of a genus from which relatively little sequence data are available. Results may lead to insight into how better to control these pathogens. The homepage of PGI can be accessed at http:www.ncgr.org/pgi, with database access through the database access hyperlink.

Databases, Factual↗

Initial assessment of gene diversity for the oomycete pathogen Phytophthora infestans based on expressed sequences.

A total of 1000 expressed sequence tags (ESTs) corresponding to 760 unique sequence sets were identified using random sequencing of clones from a cDNA library constructed from mycelial RNA of Phytophthora infestans. A number of software programs, represented by a relational database and an analysis pipeline, were developed for the automated analysis and storage of the EST sequence data. A set of 419 nonredundant sequences, which correspond to a total of 632 ESTs (63.2%), were identified as showing significant matches to sequences deposited in public databases. A putative cellular identity and role was assigned to all 419 sequences. All major functional categories were represented by at least several ESTs. Four novel cDNAs containing sequences related to elicitins, a family of structurally related proteins that induce the hypersensitive response and condition avirulence of P. infestans on Nicotiana plants, were among the most notable genes identified. Two of these elicitin-like cDNAs were among the most abundant cDNAs examined. The set also contained several ESTs with high sequence similarity to unique plant genes.

Actins↗

The Genome Sequence DataBase (GSDB): improving data quality and data access.

In 1997 the primary focus of the Genome Sequence DataBase (GSDB; www. ncgr.org/gsdb ) located at the National Center for Genome Resources was to improve data quality and accessibility. Efforts to increase the quality of data within the database included two major projects; one to identify and remove all vector contamination from sequences in the database and one to create premier sequence sets (including both alignments and discontiguous sequences). Data accessibility was improved during the course of the last year in several ways. First, a graphical database sequence viewer was made available to researchers. Second, an update process was implemented for the web-based query tool, Maestro. Third, a web-based tool, Excerpt, was developed to retrieve selected regions of any sequence in the database. And lastly, a GSDB flatfile that contains annotation unique to GSDB (e.g., sequence analysis and alignment data) was developed. Additionally, the GSDB web site provides a tool for the detection of matrix attachment regions (MARs), which can be used to identify regions of high coding potential. The ultimate goal of this work is to make GSDB a more useful resource for genomic comparison studies and gene level studies by improving data quality and by providing data access capabilities that are consistent with the needs of both types of studies.

Base Sequence↗