PubMed Health⌕ Search

PubMed · 15153304

Biological database design and implementation.

Abstract

We present our experience of building biological databases. Such databases have most aspects in common with other complex databases in other fields. We do not believe that biological data are that different from complex data in other fields. Our experience has led us to emphasise simplicity and conservative technology choices when building these databases. This is a short paper of advice that we hope is useful to people designing their own biological database.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Ewan Birney, Michele Clamp. 2004. Biological database design and implementation.. https://doi.org/10.1093/bib%2F5.1.31

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

PhD: a web database application for phenotype data management.

A database application has been developed for phenotype data management employing the Entity-Attribute-Value (EAV) model. By applying the EAV model, this application allows users to manage arbitrary phenotypes and customize data entry forms; therefore, it is suitable for different and multi-center projects.

Database Management Systems↗

Distribution of immunodeficiency fact files with XML--from Web to WAP.

BACKGROUND: Although biomedical information is growing rapidly, it is difficult to find and retrieve validated data especially for rare hereditary diseases. There is an increased need for services capable of integrating and validating information as well as proving it in a logically organized structure. A XML-based language enables creation of open source databases for storage, maintenance and delivery for different platforms. METHODS: Here we present a new data model called fact file and an XML-based specification Inherited Disease Markup Language (IDML), that were developed to facilitate disease information integration, storage and exchange. The data model was applied to primary immunodeficiencies, but it can be used for any hereditary disease. Fact files integrate biomedical, genetic and clinical information related to hereditary diseases. RESULTS: IDML and fact files were used to build a comprehensive Web and WAP accessible knowledge base ImmunoDeficiency Resource (IDR) available at http://bioinf.uta.fi/idr/. A fact file is a user oriented user interface, which serves as a starting point to explore information on hereditary diseases. CONCLUSION: The IDML enables the seamless integration and presentation of genetic and disease information resources in the Internet. IDML can be used to build information services for all kinds of inherited diseases. The open source specification and related programs are available at http://bioinf.uta.fi/idml/.

Database Management Systems↗

Querying and computing with BioCyc databases.

We describe multiple methods for accessing and querying the complex and integrated cellular data in the BioCyc family of databases: access through multiple file formats, access through Application Program Interfaces (APIs) for LISP, Perl and Java, and SQL access through the BioWarehouse relational database.

Database Management Systems↗