PubMed Health⌕ Search

PubMed · 16624967

Variable selection for propensity score models.

Abstract

Despite the growing popularity of propensity score (PS) methods in epidemiology, relatively little has been written in the epidemiologic literature about the problem of variable selection for PS models. The authors present the results of two simulation studies designed to help epidemiologists gain insight into the variable selection problem in a PS analysis. The simulation studies illustrate how the choice of variables that are included in a PS model can affect the bias, variance, and mean squared error of an estimated exposure effect. The results suggest that variables that are unrelated to the exposure but related to the outcome should always be included in a PS model. The inclusion of these variables will decrease the variance of an estimated exposure effect without increasing bias. In contrast, including variables that are related to the exposure but not to the outcome will increase the variance of the estimated exposure effect without decreasing bias. In very small studies, the inclusion of variables that are strongly related to the exposure but only weakly related to the outcome can be detrimental to an estimate in a mean squared error sense. The addition of these variables removes only a small amount of bias but can increase the variance of the estimated exposure effect. These simulation studies and other analytical results suggest that standard model-building tools designed to create good predictive models of the exposure will not always lead to optimal PS models, particularly in small studies.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

M Alan Brookhart, Sebastian Schneeweiss, Kenneth J Rothman, Robert J Glynn, Jerry Avorn, Til Stürmer. 2006-04-19. Variable selection for propensity score models.. https://doi.org/10.1093/aje%2Fkwj149

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Psychiatric disorder criteria and their application to research in different racial groups.

BACKGROUND: The advent of standardized classification and assessment of psychiatric disorders, and considerable joint efforts among many countries has led to the reporting of international rates of psychiatric disorders, and inevitably, their comparison between different racial groups. RESULTS: In neurologic diseases with defined genetic etiologies, the same genetic cause has different phenotypes in different racial groups. CONCLUSION: We suggest that genetic differences between races mean that diagnostic criteria refined in one racial group, may not be directly and simply applicable to other racial groups and thus more effort needs to be expended on defining diseases in other groups. Cross-racial confounds (in addition to cultural confounds) make the interpretation of rates in different groups even more hazardous than seems to have been appreciated.

Confounding Factors, Epidemiologic↗