PubMed HealthSearch

PubMed · 3390508

Significance testing for correlated binary outcome data.

Abstract

Multiple logistic regression is a commonly used multivariate technique for analyzing data with a binary outcome. One assumption needed for this method of analysis is the independence of outcome for all sample points in a data set. In ophthalmologic data and other types of correlated binary data, this assumption is often grossly violated and the validity of the technique becomes an issue. A technique has been developed (Rosner, 1984) that utilizes a polychotomous logistic regression model to allow one to look at multiple exposure variables in the context of a correlated binary data structure. This model is an extension of the beta-binomial model, which has been widely used to model correlated binary data when no covariates are present. In this paper, a relationship is developed between the two techniques, whereby it is shown that use of ordinary logistic regression in the presence of correlated binary data can result in true significance levels that are considerably larger than nominal levels in frequently encountered situations. This relationship is explored in detail in the case of a single dichotomous exposure variable. In this case, the appropriate test statistic can be expressed as an adjusted chi-square statistic based on the 2 X 2 contingency table relating exposure to outcome. The test statistic is easily computed as a function of the ordinary chi-square statistic and the correlation between eyes (or more generally between cluster members) for outcome and exposure, respectively. This generalizes some previous results obtained by Koval and Donner (1987, in Festschrift for V. M. Joshi, I. B. MacNeill (ed.), Vol. V, 199-224.(ABSTRACT TRUNCATED AT 250 WORDS)

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

B Rosner, R C Milton. 1988. Significance testing for correlated binary outcome data.. https://pubmed.ncbi.nlm.nih.gov/3390508/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Coloured noise or low-dimensional chaos?

Devising a method capable of distinguishing a low-dimensional chaotic signal that might be embedded in a noisy stochastic process has become a major challenge for those involved in time-series analysis. Here a null hypothesis approach is used in conjunction with a known nonlinear predictive test, to probe for the presence of chaos in epidemiological data. A probabilistic set of rules is used to stimulate a historic record of New York City measles outbreaks, generally understood to be governed by a chaotic attractor. The simulated runs of 'surrogate data' are carefully constructed so as to be free from any underlying low-dimensional chaotic process. They therefore serve as a useful null model against which to test the observed time series. However, despite the assumed differences between the dynamics of measles outbreaks and the null model, a nonlinear predictive scheme is found to be unable to differentiate between their characteristic time series. The methodology confirms that, if there is in fact a chaotic signal in the measles data, it is extremely difficult to detect in time series of such limited length. The results have general relevance to the analysis of physical, ecological and environmental time series.

Biometry