PubMed Health⌕ Search

PubMed · 12111875

A new parametric method based on S-distributions for computing receiver operating characteristic curves for continuous diagnostic tests.

Abstract

Receiver operating characteristic (ROC) curves provides a method for evaluating the performance of a diagnostic test. These curves represent the true positive ratio, that is, the true positives among those affected by the disease, as a function of the false positive ratio, that is, the false positives among the healthy, corresponding to each possible value of the diagnostic variable. When the diagnostic variable is continuous, the corresponding ROC curve is also continuous. However, estimation of such curve through the analysis of sample data yields a step-line, unless some assumption is made on the underlying distribution of the considered variable. Since the actual distribution of the diagnostic test is seldom known, it is difficult to select an appropriate distribution for practical use. Data transformation may offer a solution but also may introduce a distortion on the evaluation of the diagnostic test. In this paper we show that the distribution family known as the S-distribution can be used to solve this problem. The S-distribution is defined as a differential equation in which the dependent variable is the cumulative. This special form provides a highly flexible family of distributions that can be used as models for unknown distributions. It has been shown that classical statistical distributions can be represented accurately as S-distributions and that they occur in a definite subspace of the parameter space corresponding to the whole S-distribution family. Consequently, many other distributional forms that do not correspond to known distributions are provided by the S-distribution. This property can be used to model observed data for unknown distributions and is very useful in constructing parametric ROC curves in those cases. After fitting an S-distribution to the observed samples of diseased and healthy populations, ROC curve computation is straightforward. A ROC curve can be considered as the solution of a differential equation in which the dependent variable is the ratio of true positives and the independent variable is the ratio of false positives. This equation can be easily obtained from the S-distributions fitted to observed data. Using these results, we can compute pointwise confidence bands for the ROC curve and the corresponding area under the curve. We shall compare this approach with the empirical and the binormal methods for estimating a ROC curve to show that the S-distribution based method is a useful parametric procedure.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Albert Sorribas, Jaume March, Javier Trujillano. 2002-05-15. A new parametric method based on S-distributions for computing receiver operating characteristic curves for continuous diagnostic tests.. https://doi.org/10.1002/sim.1086

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Kullback-Leibler divergence for evaluating bioequivalence.

In this paper we propose a methodology for evaluating the bioequivalence of two formulations of a drug that encompasses not only average bioequivalence (ABE), but also the more recently introduced measures of population bioequivalence (PBE) and individual bioequivalence (IBE). The latter two measures are concerned with prescribability (PBE) and switchability (IBE). The main idea is to use the Kullback-Leibler divergence (KLD) as a measure of discrepancy between the distributions of the two formulations. Two formulations are declared bioequivalent if the upper bound of a level-alpha confidence interval for the KLD is less than a given goalpost to be set by a regulator. This new methodology overcomes many of the disadvantages of the corresponding measures recommended by the FDA. In particular the KLD: (i) possesses the natural hierarchical property that IBE => PBE => ABE; (ii) satisfies the properties of a true distance metric; (iii) is invariant to monotonic transformations of the data; (iv) generalizes easily to the multivariate case where equivalence on more than one parameter (for example, AUC, C(max) and T(max)) is required; and (v) is applicable over a wide range of distributions of the response variable (for example, those in the exponential family). The performance of the KLD relative to the metric proposed in guidance by the FDA for the evaluation of individual bioequivalence is evaluated using a simulation study. Previously published retrospective analyses using the FDA-proposed metric are contrasted with those based on the KLD. It is concluded that the KLD is a viable alternative to the FDA-proposed metric and that its mathematical and statistical properties make it a readily interpretable measure of the differences between formulations.

Area Under Curve↗