PubMed Health⌕ Search

PubMed · 14601759

A new method for choosing sample size for confidence interval-based inferences.

Abstract

Scientists often need to test hypotheses and construct corresponding confidence intervals. In designing a study to test a particular null hypothesis, traditional methods lead to a sample size large enough to provide sufficient statistical power. In contrast, traditional methods based on constructing a confidence interval lead to a sample size likely to control the width of the interval. With either approach, a sample size so large as to waste resources or introduce ethical concerns is undesirable. This work was motivated by the concern that existing sample size methods often make it difficult for scientists to achieve their actual goals. We focus on situations which involve a fixed, unknown scalar parameter representing the true state of nature. The width of the confidence interval is defined as the difference between the (random) upper and lower bounds. An event width is said to occur if the observed confidence interval width is less than a fixed constant chosen a priori. An event validity is said to occur if the parameter of interest is contained between the observed upper and lower confidence interval bounds. An event rejection is said to occur if the confidence interval excludes the null value of the parameter. In our opinion, scientists often implicitly seek to have all three occur: width, validity, and rejection. New results illustrate that neglecting rejection or width (and less so validity) often provides a sample size with a low probability of the simultaneous occurrence of all three events. We recommend considering all three events simultaneously when choosing a criterion for determining a sample size. We provide new theoretical results for any scalar (mean) parameter in a general linear model with Gaussian errors and fixed predictors. Convenient computational forms are included, as well as numerical examples to illustrate our methods.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Michael R Jiroutek, Keith E Muller, Lawrence L Kupper, Paul W Stewart. 2003. A new method for choosing sample size for confidence interval-based inferences.. https://doi.org/10.1111/1541-0420.00068

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Tutorial in biostatistics: competing risks and multi-state models.

Standard survival data measure the time span from some time origin until the occurrence of one type of event. If several types of events occur, a model describing progression to each of these competing risks is needed. Multi-state models generalize competing risks models by also describing transitions to intermediate events. Methods to analyze such models have been developed over the last two decades. Fortunately, most of the analyzes can be performed within the standard statistical packages, but may require some extra effort with respect to data preparation and programming. This tutorial aims to review statistical methods for the analysis of competing risks and multi-state models. Although some conceptual issues are covered, the emphasis is on practical issues like data preparation, estimation of the effect of covariates, and estimation of cumulative incidence functions and state and transition probabilities. Examples of analysis with standard software are shown.

Biometry↗

The role of education in biostatistical consulting.

Medical students, residents, postdoctoral fellows, and faculty commonly consult with biostatistical experts about study design and data analysis when conducting clinical research. The role of biostatistical training during these consultations is examined, and characterizations of the connections between biostatistical consultation and education are reviewed. The presence and kinds of teaching efforts during biostatistical consults at four academic research institutions over various periods of time between 1999 and 2005 (237 consultations in total) were recorded and are described. By site, 67, 70, 78, and 100 per cent of the consulting sessions included biostatistical training, with an overall 78 per cent (95 per cent CI: 73-83 per cent) of consultations including an educational component when all consultations were combined. Training covered a wide range of biostatistical topics. Seventy-five per cent of the consultations with faculty (120/161), 79 per cent with fellows and residents (31/39), and 100 per cent with medical students (10/10) included some degree of instruction in study design or statistical analysis topics. Results show that both the need and the opportunity exist for specialized biostatistical instruction during one-on-one sessions between a consulting biostatistician and physicians, medical students, and research staff. Academic researchers are ideally positioned to absorb this kind of training when they initiate a request for assistance with their own research project.

Biometry↗

Improving the quality of patient care using reliability measures: a classification tree approach.

This paper considers the application and interpretation of new reliability measures for a classification tree-based medical risk assessment tool. Following the construction of a classification tree reliability measures may then be used to provide an estimate of the precision of the classification and the probability in each terminal node of the classification tree. Identification of unreliable nodes (those that have low precision) in this application may indicate patient groups requiring closer monitoring or scenarios in which further information about the patient is required, thereby providing medical practitioners with an avenue for more informed decision making.

Biometry↗