PubMed HealthSearch

PubMed · 8131284

Evaluating test methods by estimating total error.

Abstract

A common procedure for evaluating a test method by comparison with another, well-accepted method has been to use a repeated measurements design, in which several individual subjects' specimens are assayed with both methods. We propose the use of the intrasubject relative mean square error, which is a function of the intrasubject relative bias and the coefficient of variation of the test method, as a measure of total error. We construct for each individual subject a score that is based on how well an individual's estimate of total error compares with a maximum allowable value. If the individual's score is > 100%, then that individual's estimate of total error exceeds the maximum allowable value. We present a distribution-free statistical methodology for evaluating the sample of scores. This involves the construction of an upper tolerance limit to determine whether the test method yields values of the total error that are acceptable for most of the population with some level of confidence. Our definition of total error is very different from that defined in the National Cholesterol Education Program (NCEP) guidelines. The NCEP bound for total error has three main problems: (a) it incorrectly assumes that the standard error of the estimated relative bias is the test coefficient of variation; (b) it incorrectly assumes that the individual estimated relative biases follow gaussian distributions; (c) it is based on requiring the relative bias of the average individual in the population to lie within prescribed limits, whereas we believe it is more important to require the total error for most of the individuals in the population, say 95%, to lie within prescribed limits.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

V M Chinchilli, W G Miller. 1994. Evaluating test methods by estimating total error.. https://pubmed.ncbi.nlm.nih.gov/8131284/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Assessment of blinding in pharmacotherapy and noninvasive neuromodulation randomized controlled trials for neuropathic pain in adults.

In randomized controlled trials (RCTs), study participants and research personnel are often blinded to minimize biases related to knowing treatment allocation. To determine if blinding was effective, participants may be asked which treatment they believe they received ("treatment guess"). This descriptive review characterized blinding assessment (BA) reporting in pharmacotherapy and neuromodulation neuropathic pain RCTs. Of 288 papers, 36 (12.5%) reported a BA. One paper reported the results of 2 studies, so in total 37 studies with a BA were assessed. Of these, 19 were crossover, 17 parallel, and 1 partial crossover in design. All 37 studies assessed participant blinding, and 10 also assessed investigator blinding. Approximately 27% included an "unsure" answer option for treatment guess, and 38% asked the reason for the guess. There were no clear patterns in BA reporting across time nor based on treatment type. Seventeen trials provided sufficient data to calculate Bang Blinding Index (BI) to determine blinding success. Participants remained blinded (BI = 0 &#xb1; 0.2) in 10/17 placebo and 10/17 treatment arms, 6 placebo and 5 treatment arms had a BI > 0.2 suggesting possible unblinding, whereas 1 placebo and 2 treatment arms had a BI < -0.2 suggesting misinformed guessing. Overall, we found that BAs are done in a minority of published neuropathic pain trials and with variable methodology. Given the importance of minimizing risk of bias because of treatment unblinding, future studies should consider including BAs, and further consensus building is necessary to determine if and how BAs should be conducted and interpreted in analgesic clinical trials.

Bias

Alternative approaches for estimating prevalence in epidemiologic surveys with two waves of respondents.

Estimates of prevalence in epidemiologic surveys are prone to bias due to selective response. Therefore, much effort is devoted to reduce the number of nonrespondents. For example, individuals who do not respond in the first round of recruitment in mail surveys are usually contacted a second (or even third or fourth) time yielding consecutive waves of responses. Yet this sequence of waves is often neglected in epidemiologic analyses in that prevalence is simply estimated as the proportion of trait-positive individuals among the total group of respondents. This paper investigates alternative estimates of prevalence that might be used in surveys with two waves of respondents. The estimates are based on different assumptions on the relation of response rates with the trait of interest. As this relation is likely to vary from survey to survey depending on the specific circumstances under which the recruitment of participants is conducted, none of the estimates is universally preferable. The performance of the different estimates is assessed in a variety of hypothetical and empirical examples, and strategies are discussed to make the best use of the different estimates in the analysis of epidemiologic studies.

Bias