PubMed Health⌕ Search

PubMed · 10089484

Macromolecular structure databases: past progress and future challenges.

Abstract

Databases containing macromolecular structure data provide a crystallographer with important tools for use in solving, refining and understanding the functional significance of their protein structures. Given this importance, this paper briefly summarizes past progress by outlining the features of the significant number of relevant databases developed to date. One recent database, PDB+, containing all current and obsolete structures deposited with the Protein Data Bank (PDB) is discussed in more detail. PDB+ has been used to analyze the self-consistency of the current (1 January 1998) corpus of over 7000 structures. A summary of those findings is presented (a full discussion will appear elsewhere) in the form of global and temporal trends within the data. These trends indicate that challenges exist if crystallographers are to provide the community with complete and consistent structural results in the future. It is argued that better information management practices are required to meet these challenges.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

H Weissig, I N Shindyalov, P E Bourne. 1998-11-01. Macromolecular structure databases: past progress and future challenges.. https://doi.org/10.1107/s0907444998009846

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The Longhorn Array Database (LAD): an open-source, MIAME compliant implementation of the Stanford Microarray Database (SMD).

BACKGROUND: The power of microarray analysis can be realized only if data is systematically archived and linked to biological annotations as well as analysis algorithms. DESCRIPTION: The Longhorn Array Database (LAD) is a MIAME compliant microarray database that operates on PostgreSQL and Linux. It is a fully open source version of the Stanford Microarray Database (SMD), one of the largest microarray databases. LAD is available at http://www.longhornarraydatabase.org CONCLUSIONS: Our development of LAD provides a simple, free, open, reliable and proven solution for storage and analysis of two-color microarray data.

Database Management Systems↗

Digital extractor: analysis of digital differential display output.

Digital Extractor is a program for the high-throughput processing of data sets derived from digital differential display-based comparisons of EST libraries. These comparisons can be utilized to identify discrete subsets of genes whose expression is restricted to distinct tissue types. The program facilitates these investigations by permitting parallel annotation of genes identified as being differentially expressed.

Database Management Systems↗

Achieving evolvable Web-database bioscience applications using the EAV/CR framework: recent advances.

The EAV/CR framework, designed for database support of rapidly evolving scientific domains, utilizes metadata to facilitate schema maintenance and automatic generation of Web-enabled browsing interfaces to the data. EAV/CR is used in SenseLab, a neuroscience database that is part of the national Human Brain Project. This report describes various enhancements to the framework. These include (1) the ability to create "portals" that present different subsets of the schema to users with a particular research focus, (2) a generic XML-based protocol to assist data extraction and population of the database by external agents, (3) a limited form of ad hoc data query, and (4) semantic descriptors for interclass relationships and links to controlled vocabularies such as the UMLS.

Database Management Systems↗