PubMed Health⌕ Search

PubMed · 10683130

Creating a computerized database from administrative claims data.

Abstract

The creation of a computerized database from Medicaid administrative claims data for research purposes is described. Researchers should consult with computer experts at their institution before selecting software for data manipulation and conversion. It is essential to have an accurate layout of the file record before attempting to convert raw claims data into data sets or other data formats. The location of data elements within the claim will vary depending on whether the record comes from a provider, an institution, or a pharmacy. Each claim contains a common header, a variable header, and a claim detail section. The difficulty in analyzing data elements within a claim detail lies in locating the starting point of the claim detail section. So that data elements not in character or numeric formats can be converted, the file record layout must describe the exact format of each data element and its COBOL notation. A data element dictionary is necessary for translating data element coding into usable data. Data elements not necessary for any planned analysis must be eliminated. The data are then "cleaned" to remove any denied or reversed claims and claims that contain incomplete or erroneous data. Regardless of the format data are obtained in, an accurate file record layout and a data element dictionary are essential to the conversion of administrative claims data into a computerized database for data analysis and research purposes.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

L T Piecoro, L S Wang, W S Dixon, R J Crovo. 1999-07-01. Creating a computerized database from administrative claims data.. https://doi.org/10.1093/ajhp%2F56.13.1326

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Retention index database for identification of general green leaf volatiles in plants by coupled capillary gas chromatography-mass spectrometry.

A series of ubiquitously occurring saturated and monounsaturated six-carbon aldehydes, alcohols and esters thereof is summarised as 'green leaf volatiles' (GLVs). The present study gives a comprehensive data collection of retention indices of 35 GLVs on commonly used non-polar DB-5, mid-polar DB-1701, and polar DB-Wax stationary phases. Seventeen commercially not available compounds were synthesised. Thus, the present study allows reliable identification of most known GLV in natural plant volatile samples. Applications revealed the presence of several seldom reported GLVs in headspace samples of mechanically damaged plant leaves of Carpinus betulus and Fagus sylvatica.

Database Management Systems↗

Toward general methods of targeted library design: topomer shape similarity searching with diverse structures as queries.

A promising strategy for selecting synthetic targets is similarity-based searching of very large "virtual libraries", which comprise all structures accessible by linking two or three commercially available building blocks with combinatorial syntheses. To assess the general applicability of this strategy, leading structures taken from each of 34 recent medicinal chemistry publications were used as queries to search a virtual library containing 2.6 x 10(13) products from seven reactions, using a topomer shape similarity metric. Eighty-five percent of these searches succeeded, by yielding, with a search radius no greater than 120 topomer shape units, either at least 400 hits or hits from at least six sublibraries. From these 34 sets of search results, 122 representative structures were selected, illustrating potential "lead hops", or otherwise novel structures. Overall shape similarity to the query structure was confirmed for up to 95% of these representative structures, according to FLEXS, an algorithmically distinct program. Experimentally, there were 28 structures among those reported in the 34 query publications that were identified within the virtual library. Among these, the frequency of high activity was 87% for the 16 structures whose similarity to their query was 90 topomer units or less, compared to a frequency of 50% for the other 12 structures.

Database Management Systems↗

NCBI's LocusLink and RefSeq.

The NCBI has introduced two new web resources-LocusLink and RefSeq-that facilitate retrieval of gene-based information and provide reference sequence standards. These resources are designed to provide a non-redundant view of current knowledge about human genes, transcripts and proteins. Additional information about these resources is available on the LocusLink web site at http://www.ncbi.nlm.nih.gov/LocusLink/

Database Management Systems↗