PubMed HealthSearch

PubMed · 9019093

Database and knowledge base integration--a data mapping method for Arden Syntax knowledge modules.

Abstract

One of the most important categories of decision-support systems in medicine are data driven systems where the inference engine is linked to a database. It is, therefore, important to find methods that facilitate the implementation of database queries referred to in the knowledge modules. A method is described for linking clinical databases to a knowledge base with Arden Syntax modules. The method is based on a query meta-database including templates for SQL queries which is maintained by a database administrator. During knowledge module authoring the medical expert refers only to a code in the query meta-database; no knowledge is needed about the database model or the naming of attributes and relations. The method uses standard tools, such as C+2 and ODBC, which makes it possible to implement the method at many platforms and to link to different clinical databases in a standardized way.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

B Johansson, N Shahsavar, H Ahlfeldt, O Wigertz. 1996. Database and knowledge base integration--a data mapping method for Arden Syntax knowledge modules.. https://pubmed.ncbi.nlm.nih.gov/9019093/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Quality assessment of NMR structures: a statistical survey.

A statistical analysis is reported of experimental data and coordinates of a set of 97 NMR structures deposited in the PDB. The aim is to assess the quality of these structures in relation to the amount of experimental information. Experimental restraints were analysed using the program AQUA. Many nomenclature inconsistencies between deposited restraint and coordinate files were observed. The experimental restraint files were found to contain a high proportion of redundant restraints. Procedures for analysing and correcting the inconsistencies and restraint counts are described. The analysis of NOE restraint violations (using AQUA) and of a wide variety of geometrical quality indicators (using PROCHECK-NMR and WHAT IF) provides a reference for other NMR structure determinations. The extent of NOE violations is anti-correlated with the quality of the Ramachandran map. The precision as measured by the circular variance of backbone dihedral angles, does increase with the amount of experimental data, as expected, but is sometimes overestimated. Bond lengths, bond angles and planarity of groups can deviate considerably from ideal values. Outliers appear to cluster per laboratory, indicating that the results depend on particulars of refinement protocols and/or software. We have identified a problem of atom overlap in a number of refined structures.We recommend adhering to the standard nomenclature as put forward by an IUPAC Task Group, to ensure consistency between restraints and coordinates, and to omit redundant restraints from the deposition. The results obtained from this analysis and the AQUA program are available through the World Wide Web.

Databases, Factual

Dictionary of interfaces in proteins (DIP). Data bank of complementary molecular surface patches.

Molecular surface areas of proteins are responsible for selective binding of ligands and protein-protein recognition, and are considered the basis for specific interactions between different parts of a protein. This basic principle leads us to study the interfaces within proteins as a learning set for intermolecular recognition processes of ligands like substrates, coenzymes, etc., and for prediction of contacts occurring during protein folding and association. For this purpose, we defined interfaces as pairs of matching molecular surface patches between neighboring secondary structural elements. All such interfaces from known protein structures were collected in a comprehensive data bank of interfaces in proteins (DIP). The up-to-date DIP contains interface files for 351 selected Brookhaven Protein Data Bank entries with a total of about 160,000 surface elements formed by 12,475 secondary structures. For special purposes, the inclusion of additional structures or selection of subgroups of proteins can be performed in an easy and straightforward manner. Atomic coordinates of the constituents of molecular surface patches are directly accessible as well as the corresponding contact distances from given atoms to their neighboring secondary structural elements. As a rule, independent of the type of secondary structure, the molecular surface patches of the secondary structural elements can be described as quite flat bodies with a length to width to depth ratio of about 3:2:1 for patches consisting of more than ten atoms. The relative orientation between two docking patches is strongly restricted, due to the narrow distribution of the distances between their centers of mass and of the angles between their normal lines, respectively. The existing retrieval system for the DIP allows selection (out of the set of molecular patches) according to different criteria, such as geometric features, atomic composition, type of secondary structure, contacts, etc. A fast, sequence-independent 3-D superposition procedure was developed for automatic searches for geometrically similar surface areas. Using this procedure, we found a large number of structurally similar interfaces of up to 30 atoms in completely unrelated protein structures.

Databases, Factual

Phases and phase transitions of the phosphatidylcholines.

LIPIDAT (http://www.lipidat.chemistry.ohio-state.edu) is an Internet accessible, computerized relational database providing access to the wealth of information scattered throughout the literature concerning synthetic and biologically derived polar lipid polymorphic and mesomorphic phase behavior and molecular structures. Here, a review of the data subset referring to phosphatidylcholines is presented together with an analysis of these data. This subset represents ca. 60% of all LIPIDAT records. It includes data collected over a 43-year period and consists of 12,208 records obtained from 1573 articles in 106 different journals. An analysis of the data in the subset identifies trends in phosphatidylcholine phase behavior reflecting changes in lipid chain length, unsaturation (number, isomeric type and position of double bonds), asymmetry and branching, type of chain-glycerol linkage (ester, ether, amide), position of chain attachment to the glycerol backbone (1,2- vs. 1,3-) and head group modification. Also included is a summary of the data concerning the effect of pressure, pH, stereochemical purity, and different additives such as salts, saccharides, amino acids and alcohols, on phosphatidylcholine phase behavior. Information on the phase behavior of biologically derived phosphatidylcholines is also presented. This review includes 651 references.

Databases, Factual