PubMed HealthSearch

SEARCH · PubMed Health

Results for “Database”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2Linked to original sources

The effect of a multiple literature database search--a numerical evaluation in the domain of Japanese life science.

In literature database searching, we show that it is necessary to use plural databases for a more improved search. We also compare the results of a single database search with that of multiple database search in the domain of Japanese life sciences. We searched the MEDLINE and EMBASE using the same search terms. There were some differences in the results, owing to differences in the journals and recording methods. We herein show some of the differences in the journals contained in both databases. Furthermore, we show the differences in the number of papers derived from the same journal. Next, as an example of a practical search, we selected some universities in Japan, searched both databases regarding papers published from these universities and then merged the results by hand. According to our results, only 63% of all papers were common to both databases.

Biological Science Disciplines

The free form database program as a research tool.

Two types of database programs for the IBM compatible personal computer (PC) are described: the fixed form database and the free form database. The use of the latter in compiling bibliographic databases and in content analysis of interview transcripts is described. Other uses for the free form database program are also discussed. It is suggested that the free form database has advantages over some other custom-made analysis programs in terms of its simplicity and ease of use. It is also suggested that the free form database can be a useful tool for nurse educators and students.

Databases, Bibliographic

Automated assembly of protein blocks for database searching.

A system is described for finding and assembling the most highly conserved regions of related proteins for database searching. First, an automated version of Smith's algorithm for finding motifs is used for sensitive detection of multiple local alignments. Next, the local alignments are converted to blocks and the best set of non-overlapping blocks is determined. When the automated system was applied successively to all 437 groups of related proteins in the PROSITE catalog, 1764 blocks resulted; these could be used for very sensitive searches of sequence databases. Each block was calibrated by searching the SWISS-PROT database to obtain a measure of the chance distribution of matches, and the calibrated blocks were concatenated into a database that could itself be searched. Examples are provided in which distant relationships are detected either using a set of blocks to search a sequence database or using sequences to search the database of blocks. The practical use of the blocks database is demonstrated by detecting previously unknown relationships between oxidoreductases and by evaluating a proposed relationship between HIV Vif protein and thiol proteases.

Algorithms

Expanding vaginal microbiome pangenomes via a custom MIDAS database reveals Lactobacillus crispatus accessory genes associated with cervical dysplasia.

The vaginal microbiome plays a central role in reproductive health. Vaginal microbiome dysbiosis is associated with many adverse reproductive health outcomes, but most studies have focused on associations at the species level. The potential contribution of intraspecies microbial variation, especially gene content differences across bacterial strains, remains underexplored in reproductive health contexts. The Metagenomic Intra-Species Diversity Analysis (MIDAS) framework enables such analyses, but depends on comprehensive reference databases. We constructed a MIDAS-compatible pangenome database from over 18,000 genomes in the Vaginal Microbiome Genome Collection (VMGC). Compared to the Genome Taxonomy Database (GTDB)-derived reference, the VMGC-derived database expanded the pangenomes of prevalent vaginal species, better capturing vaginal-specific intraspecies diversity. Applying this database to vaginal samples from a cervical dysplasia cohort, we identified 13 Lactobacillus crispatus accessory genes significantly associated with cervical dysplasia, including a HicAB toxin-antitoxin system, three transcriptional regulators, and three phage-derived genes. These findings highlight the utility of body site-specific reference resources and shotgun metagenomic sequencing for uncovering intraspecies microbial variation relevant to reproductive health.IMPORTANCEThe vaginal microbiome plays a critical role in reproductive health, and different bacteria from the same species can carry different genes that influence how the strains interact with the host and other microbes. These strain-level differences are often overlooked when microbiomes are analyzed only at the species level. Existing genomic reference databases are heavily biased toward gut and environmental bacteria, leaving the genetic diversity of vaginal microbes understudied. We built a specialized reference database from over 18,000 vaginal bacterial genomes that better reflects this diversity. We then applied this resource to quantify gene-level variation in vaginal samples from a cervical dysplasia cohort. Focusing on Lactobacillus crispatus, a prevalent and often beneficial vaginal species, we identified 13 genes that were more common in women with cervical dysplasia than in controls. This work demonstrates that body site-specific genomic resources are essential for uncovering strain-level bacterial differences relevant to reproductive health.

Lactobacillus crispatus

Carcinogenicity evaluations and ongoing studies: the IARC databases.

Many thousands of chemicals are produced industrially and many more occur naturally. Information on the toxicology of these chemicals is often minimal or absent. The International Agency for Research on Cancer (IARC) has published evaluations of the carcinogenic risk to humans of over 700 chemicals, groups of chemicals, and complex mixtures as a regular series of monographs. A database has been created containing summaries of all the relevant epidemiological, animal carcinogenicity, and other relevant biological data for each chemical or mixture evaluated. Additional databases have been created for ongoing epidemiological studies of cancer in humans and for long-term carcinogenicity studies in rodents, as well as a database containing information on genotoxic and related effects of chemicals. Some of these databases have been published in print form. IARC now plans to publish them electronically, together with other databases, in the form of a CDROM (compact disk, read-only memory). The objective will be to make the entire IARC database of cancer information as widely available as possible in an integrated format conducive to efficient and combined exploitation of all the component databases.

Animals

[A new database system for radiological reports].

We have designed and developed a new database system to facilitate automatic feedback of the content of radiology reports to radiologists. The prototype of this database system has been implemented in the RGSS-IDJ, a developmental computer system that applies artificial intelligence methods to a reporting system. This prototype system was constructed to test the feasibility of overcoming the limitations of conventional database systems. The new database system is based on our semantic model for radiology reports and is able to treat data with unnormalized relations. Operations specific to our database system include the ability to acquire information about a set of reports that contains any semantic expression included in the lexicon and the ability to obtain the expressions that belong to a set of several semantic expressions in the reports. Thus, our new database system will offer a more powerful tool for analyzing the content of reports than conventional database systems.

Databases, Bibliographic

Up-to-date, and taxonomy-curated mcrA reference databases for methanogen community profiling.

The methyl-coenzyme M reductase subunit alpha gene (mcrA) is an important phylogenetic marker for high throughput ecological profiling of methanogenic archaea, central to industrial biological methane production and greenhouse gas emissions. Yet, dedicated reference databases predate current relevant NCBI sequence accumulation and archaeal taxonomic revision. We present three updated mcrA reference databases: (i) one derived from NCBI-catalogued methanogen genomes (1572 sequences); (ii) a database built by expansion of a previously published reference dataset, leveraging the NCBI nucleotide collection (27,942 sequences); (iii) a curated-taxonomy version of the latter. The updated amplicon databases provide a ∼ 3.5-fold sequence richness expansion, extend genus-level richness from 31 to 83 taxa, more than 4-fold species-level richness, and incorporate novel lineages compared with the previous reference dataset (e.g. Thermoplasmatota-encompassed). All databases were formatted to support analysis with relevant contemporary software pipelines and packages. Overall, the generated databases facilitate a highly improved characterization of methanogen diversity and ecology.

Archaea

Characteristics of the U.S. EPA's Office of Pesticide Programs' toxicity information databases.

The United States Environmental Protection Agency's Office of Pesticide Programs (OPP) requires that data from toxicity testing be submitted to the OPP to support the registration of pesticide chemicals. Once the toxicity data are submitted, they are entered into various toxicity databases. The studies are listed in an archival database to catalog and allow retrieval of the study for review. Reviews of toxicity studies are then placed into a separate database that can be retrieved to support a regulatory position. Toxicity information for health effects other than cancer and gene mutations from chronic exposure is reviewed through a reference dose (RfD) approach, and these decisions and supporting data are entered into an RfD database. Carcinogenicity data are reviewed by a peer review process, and these decisions are entered into a newly developed database to show the regulatory decision with supporting data. The mutagenicity data are reviewed and acceptable data are entered into the Genetic Activity Profile system to catalog and display the submitted information. These databases contain the information used for hazard evaluations as part of the OPP review of pesticide chemicals.

Animals

Non-sequence databases for biological activity and physicochemical properties.

A biological activity database and a physicochemical property database are described. They are intended to complement the protein sequence database of PIR-International. The Biological Activity Database and the Physicochemical Property Database contain information regarding the biological activity and the physicochemical properties of proteins, respectively. In addition they also provide information about wild-type molecules with which information concerning variant molecules may be compared. Data on artificial variant molecules are stored in the Artificial Variant Database which is described separately.

Amino Acid Sequence

The gene-protein database of Escherichia coli: edition 4.

The gene-protein database of Escherichia coli has as its core an index that links each of the protein spots from a two-dimensional polyacrylamide gel to the gene that encodes the protein. Additional information about each protein and its gene is generated from two-dimensional gel analysis or collated from the literature to form the database. Earlier editions of the database have provided periodic updates of information. The current edition does this, but also introduces a new reference gel image produced by an electrophoresis system recently adopted in this laboratory. The new gel system was chosen because it offers an improved opportunity for other investigations to produce close replicas of the reference gel pattern, thereby allowing easier access to the information of the database and encouraging independent contribution to the database. The new gel format also is larger and hence more compatible with computer assisted image analysis, which has become essential for a project of this magnitude. This edition continues the use of the former reference gel images, but adds a reference image of an equilibrium gel of E. coli strain W3110 produced by the new standardized gel system. At this time, 55% of the protein spots annotated on the previous equilibrium reference gel for this organism have been located on the new reference image, and these identifications are included in the tables of the database.

Bacterial Proteins

Mouse liver protein database: a catalog of proteins detected by two-dimensional gel electrophoresis.

Alterations in the abundance or structure of mouse liver proteins are being studied using two-dimensional gel electrophoresis (2-DE) to build a database of protein changes correlating with exposure to ionizing radiation or toxic chemicals. Thus far, studies have included the analysis of proteins from the offspring of exposed parents or from the exposed individuals themselves. In order to characterize and identify proteins found altered by such exposures, sex- and strain-related differences in protein patterns have been analyzed, and the subcellular locations of a large portion of the mapped proteins have been determined. As part of these studies, data are collected and stored using a variety of computer hardware and software tools that allow the accumulation of information on the origin of samples, gel identification, experiment description, and protein similarities and differences. This accumulation of information constitutes the mouse liver protein database. Relational database software is used to tie the different facets of the database together so that the results of a variety of experiments can be compared and interrelated. The database optimizes the information obtained from 2-DE gel sets by allowing use of the data for many purposes, including monitoring of gel resolution to ensure the collection of high quality data and correlation of protein effects induced by different agents. This first edition of the Argonne National Laboratory mouse liver protein database lays the foundation for future work and communication that should elucidate the significance of observed protein effects as possible markers of exposure to toxic agents.

Animals

Practice databases and their uses in clinical research.

A few large clinical information databases have been established within larger medical information systems. Although they are smaller than claims databases, these clinical databases offer several advantages: accurate and timely data, rich clinical detail, and continuous parameters (for example, vital signs and laboratory results). However, the nature of the data vary considerably, which affects the kinds of secondary analyses that can be performed. These databases have been used to investigate clinical epidemiology, risk assessment, post-marketing surveillance of drugs, practice variation, resource use, quality assurance, and decision analysis. In addition, practice databases can be used to identify subjects for prospective studies. Further methodologic developments are necessary to deal with the prevalent problems of missing data and various forms of bias if such databases are to grow and contribute valuable clinical information.

Clinical Medicine

Genome-related datasets within the E. coli Genetic Stock Center database.

The contents of the E. coli Genetic Stock Center database and the availability in electronic form of the subset of information most relevant to sequence databases are described. The database uses the long-standing Stock Center records (developed and curated by Dr B.J.Bachmann) in describing genotypes of mutant derivatives of E.coli K-12 in terms of alleles, structural mutations, mating type, and plasmids as well as the derivation, names and originators of the strain, and references. The database includes descriptions of mutations, mutation properties, genes, gene properties, and gene products, with EC number identifiers for enzymes. Sequence information is not included, but entries refer to sequence database accession numbers for sequenced regions. A gene is described as a subtype of a more general category of chromosome interval called Site. Since sites are used to describe any chromosomal interval, mapping information is associated with sites. Alleles are described as mutations of those sites and they are not primary map objects, but inherit map position information from the corresponding site description. The database design is intended to preserve richness of detail where it is known and uncertainty of measurements or information as it occurs in order to represent the stock center records as accurately as possible.

Bacterial Proteins

Comparison and evaluation of nine bibliographic databases concerning adverse drug reactions.

Few evaluations and statistical comparisons of bibliographic databases have been published. As a drug information center, we were particularly interested in databases providing references on adverse drug reactions (ADRs). Ten drugs were randomly chosen from the 2000 files at our center. Nine databases were selected according to the high frequency of references concerning ADRs: eight online systems (MEDLINE, BIOSIS, TOXLINE, Iowa Drug Information System, PASCAL, EMBASE, PHARMLINE, and International Pharmaceutical Abstracts [IPA]), and one Compact Disk Read Only Memory (CD-ROM) system (Core MEDLINE). The total number of references, the number of references from 1987 to 1989, and the number of relevant references from 1987 to 1989 were analyzed using the Friedman two-way ANOVA by ranks. The overlap between databases for only one drug, carboplatin, and the quality:cost ratio were also studied. Considering the total number of references, TOXLINE and EMBASE were significantly superior to IPA, PHARMLINE, PASCAL, and Core MEDLINE. For the period 1987-1989, EMBASE was significantly superior to PASCAL, IPA, PHARMLINE, and Core MEDLINE with regard to total number of references, and significantly superior to PASCAL, Core MEDLINE, and IPA with regard to relevance. MEDLINE, TOXLINE, and EMBASE had the best quality:cost ratio. EMBASE had the slightest overlap of references, with 53 percent of the unique references on carboplatin. This comparative evaluation showed that the ability of bibliographic databases to provide information on ADRs is dependent on both the size and the quality of each database.

Databases, Bibliographic

Examples of uses of databases for quantitative and qualitative correlation studies between genotoxicity and carcinogenicity.

In this paper we give some examples of using databases of genotoxicity and carcinogenicity for quantitative and qualitative correlation studies between short-term tests and carcinogenicity. The quality of the databases is obviously important, but one of the major deficiencies of present databases is that they are too small. Using relatively small, different databases, different results can be obtained. With small databases it is difficult to disaggregate data for homogeneous chemical classes or other types of subsets. Using the databases of Gold (carcinogenicity) and Würgler (genotoxicity), we have investigated the carcinogenic potency of genotoxic and nongenotoxic carcinogens for different chemical classes.

Animals

Aspects of database construction and interrogation of relevance to the accurate prediction of rodent carcinogenicity and mutagenicity.

Attempts to reconcile qualitative carcinogenicity databases with qualitative mutagenicity database continue to indicate that there is no useful relationship between mutagenicity/genotoxicity and rodent carcinogenicity. It is suggested that recognition of two classes of carcinogen, genotoxic and nongenotoxic, is the first step in finding meaningful correlations between the above parameters. This then leads to purposeful intervention into the databases, including rejecting low quality data, abandoning some assays from the database, and clustering certain end points as repetitive rather that independent of each other. Seeking specific correlations within a focused database may yield knowledge from the current wealth of information. The effort required to build databases, particularly quantitative ones, has so far prevented the equally arduous task of their correct interrogation. Preliminary indications are the mutagenicity is closely correlated with genotoxic carcinogenesis and completely independent of nongenotoxic carcinogenesis.

Animals

Creating a resource database for nursing service administration.

In response to the current information explosion in nursing service administration (NSA), the authors felt a need to collect and organize available resources for use by their faculty and graduate students. An electronic database was developed to facilitate the use of the collected print and software resources. This article describes the creation of the NSA Resource Database from the time the need for it was realized to its completion. There is discussion regarding the criteria used for writing the database, what the database screens look like and why and what the database contains. The article also discusses the use and users of the NSA Resource Database to date.

Databases, Bibliographic

Methods for the analysis and assessment of clinical databases: the clinician's perspective.

Innovative approaches to analysing clinical databases can be considered from a perspective of innovations that improve the analytical approach or from a more global perspective in which clinical databases themselves are evaluated as a technology. The analytic approach for using a database to estimate risk can be considered as a matrix of three methodologic concerns: the predictive method; the assessment of the quality of the predictions; and the assessment of the validity or generalizability of the predictions. Considering databases as a technology places in perspective the merit of clinical databases and defines their potential value to the health care system. An awareness of both the clinical and analytic problem encourages innovation and can lead to creative solutions to the many problems present in the analysis of clinical databases.

Clinical Medicine