PubMed HealthSearch

PubMed · 6517052

Sample-size calculations in segregation analysis.

Abstract

Segregation analysis, employing nuclear families, is the most frequently used method to evaluate the mode of inheritance of a trait. To our knowledge, there exists no tabular information regarding the sample sizes required of individuals and families needed to perform a significance test of a specific segregation ratio for a predetermined power and significance level. To fill this gap, we have developed sample-size tables based on the asymptotic variance of the maximum likelihood estimate of the segregation ratio and on the normal approximation for two-sided hypothesis testing. Assuming homogeneous sibship size, minimum sample sizes were determined for testing the null hypothesis for the segregation ratio of 1/4 or 1/2 vs. alternative values of .05-.80, for the significance level of .05 and power of .8, for ascertainment probabilities of nearly 0 to 1.0, and sibship sizes 2-7. The results of these calculations indicate a complex interaction of the null and the alternate hypotheses, ascertainment probability, and sibship size in determining the sample size required for simple segregation analysis. The accompanying tables should aid in the appropriate design and cost assessment of future genetic epidemiologic studies.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

F L Wong, J I Rotter. 1984. Sample-size calculations in segregation analysis.. https://pubmed.ncbi.nlm.nih.gov/6517052/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

New estimates of maternal mortality.

A major new study carried out by WHO and the United Nations Children's Fund (UNICEF) indicates that maternal mortality has been substantially underestimated in the past, and that there are nearly 80 000 more pregnancy-related deaths per year than previously reported. According to the study, some 585 000 maternal deaths occur in the world each year, 99% of them in developing countries.

Epidemiologic Methods

Two-stage global search designs for linkage analysis using pairs of affected relatives.

We investigate the two-stage procedure proposed by Elston [(1992) Proceedings of the XVIth International Biometric Conference, Hamilton, New Zealand, December 7-11, 1992, pp 39-51, and (1994) "Genetic Approaches to Mental Disorders." Washington, DC: American Psychiatric Press, pp 3-21] for performing a global search of the genome to locate disease genes by linkage analysis using affected relative pairs. The optimal design depends on the type of pairs studied, the effect of the disease locus, the relative costs of recruiting affected persons and typing markers, how informative the markers are, and the amount of genetic heterogeneity. It is specified by the initial number of markers to use, the number of affected relative pairs to study, the initial significance level alpha* to use at the first stage, and the number of flanking markers to use at the second stage around markers significant at the first stage. Asymptotically, the optimal design does not depend separately on either the desired final significance level or power, but rather on a function of the two. Both as the effect of the disease locus increases and as the relative cost of recruiting a subject increases, the optimal number of initial markers increases and the optimal number of pairs decreases. The expected cost of the study decreases as the effect of the disease locus increases, but increases as the relative cost of recruiting a subject increases. The optimal initial number of markers decreases but the number of pairs increases when there is genetic heterogeneity present; conversely, the optimal initial number of markers increases when markers are less than fully informative. Compared to a one-stage procedure, a two-stage procedure typically halves the cost of a study.

Epidemiologic Methods