PubMed HealthSearch

PubMed · 8896134

Explained variation for logistic regression.

Abstract

Different measures of the proportion of variation in a dependent variable explained by covariates are reported by different standard programs for logistic regression. We review twelve measures that have been suggested or might be useful to measure explained variation in logistic regression models. The definitions and properties of these measures are discussed and their performance is compared in an empirical study. Two of the measures (squared Pearson correlation between the binary outcome and the predictor, and the proportional reduction of squared Pearson residuals by the use of covariates) give almost identical results, agree very well with the multiple R2 of the general linear model, have an intuitively clear interpretation and perform satisfactorily in our study. For all measures the explained variation for the given sample and also the one expected in future samples can be obtained easily. For small samples an adjustment analogous to Radj2 in the general linear model is suggested. We discuss some aspects of application and recommend the routine use of a suitable measure of explained variation for logistic models.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

M Mittlböck, M Schemper. 1996-10-15. Explained variation for logistic regression.. https://doi.org/10.1002/(sici)1097-0258(19961015)15%3A19%3C1987%3A%3Aaid-sim318%3E3.0.co%3B2-9

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

How many fold types of protein are there in nature?

Many protein structures have now been determined and reveal that protein molecules can adopt the same fold despite having very different sequences. It has been suggested that, owing to different stereochemical constraints, the number of ways that a sequence can fold may be limited. Therefore, it is reasonable to ask how many fold types exist in nature. Several groups have tackled this problem with very different results. In the present study, a novel statistical sampling approach is used to reestimate this number. The results suggest that the number of protein folds in nature is probably several hundreds.

Likelihood Functions

T2 maximum likelihood estimation from multiple spin-echo magnitude images.

An optimal maximum likelihood (ML) method is described for an unbiased estimation of monoexponential T2 from magnitude spin-echo images. The algorithm is based on a Gaussian assumption of noise distribution. The validity of this assumption was checked by a statistical chi 2 test on spin-echo and fast low-angle shot surface coil images. Monte-Carlo simulations of magnitude data showed that the ML estimate standard deviation is lower than that produced by a weighted least-squares fitting on signal logarithm. Correction schemes are proposed to reduce bias deriving from magnitude reconstruction. The variance of the ML estimate converged rapidly toward the theoretical algebraic expression of the Cramér-Rao lower bound.

Likelihood Functions