PubMed Health⌕ Search

PubMed · 15220022

Simple solution to a common statistical problem: interpreting multiple tests.

Abstract

BACKGROUND: The misinterpretation of the results of multiple statistical tests is an error commonly made in scientific literature. When testing several outcome variables simultaneously, many researchers declare a statistically significant result for each test having a P value of <0.05, for example. This approach ignores the fact that, based on a probability result called the Bonferroni inequality, the risk of incorrectly declaring as significant > or =1 test result increases with the number of tests conducted. The implication of this practice is that many scientific results are presented as statistically significant when the underlying data do not adequately support such a claim (sometimes referred to as false-positive results). Although the sequentially rejective Bonferroni test is well known among statisticians, it is not used routinely in scientific literature. OBJECTIVE: The intent of this article was to increase the awareness and understanding of the sequentially rejective Bonferroni test, thereby expanding its use. METHODS: This article describes the statistical problem and demonstrates how the use of the sequentially rejective Bonferroni test ensures that incorrect declarations of statistical significance for > or =1 test result are bounded by 0.05, for example. CONCLUSION: The sequentially rejective Bonferroni test is an easily applied, versatile statistical tool that enables researchers to make simultaneous inferences from their data without risking an unacceptably high overall type I error rate.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Toufigh Gordi, Harry Khamis. 2004. Simple solution to a common statistical problem: interpreting multiple tests.. https://doi.org/10.1016/s0149-2918(04)90078-1

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Four-fold table cell frequencies imputation in meta analysis.

Meta analysis is a collection of quantitative methods devoted to combine summary information from related but independent studies. Because research reports usually present only data reductions and summary statistics rather than detailed data, the reviewer must often resort to rather crude methods for constructing summary effect estimate suitable for meta analysis pooling methods. When the studies involve a binary variable, both number of events and sample sizes are required to compute pooled estimate and its confidence interval. Sometimes, only summary statistics and related confidence intervals are provided in the publication. Although it is possible to estimate the standard error of each study's effect measure using the confidence interval from each study, this lack of detailed data compels the reviewers to use the inverse variance method to perform meta analysis, or to exclude the works with incomplete data. This paper shows three methods to reconstruct four-fold tables when summary measures for binary data and related confidence intervals and sample sizes are provided. The methods are discussed through a wider application example to assess the reconstruction precision, and the impact of using reconstructed data on meta analysis results. These methods seem to yield a correct reconstruction if original measures are reported at least with two decimal places. Meta analysis results do not seem seriously affected by the use of reconstructed data. These methods allow the reviewer to use full meta analysis statistical tools, instead of the simple inverse variance method, and can greatly contribute to the completeness of systematic reviews.

Confidence Intervals↗

Polymorphisms in the promoter region of the alpha1A-adrenoceptor gene are associated with schizophrenia/schizoaffective disorder in a Spanish isolate population.

BACKGROUND: Animal models have implicated the alpha(1)-adrenergic subtypes in cognitive functions relevant to schizophrenia, but no consensus exists with regard to the status of noradrenergic receptor populations in psychiatric patients. We focused on one alpha(1)-adrenergic subtype, the alpha(1A)-adrenergic receptor, and proposed that genetic variants within the regulatory region of this gene (ADRA1A) alter the expression of this receptor, influencing susceptibility toward schizophrenia. METHODS: This study examined this proposal by testing the hypothesis that single nucleotide polymorphisms (SNPs) in the promoter region of the alpha(1A)-adrenergic gene were associated with schizophrenia by performing case-control association analysis on SNPs found in a 5' upstream region, which included the putative promoter region and 5' untranslated region. Our sample consisted of 103 schizophrenia and 14 schizoaffective disorder patients and 176 control subjects. All recruits were from a Spanish population isolate of Basque origin that is characterized by low heterogeneity, which was selected with the intent that it might facilitate the identification of disease-related polymorphisms. RESULTS: A total of eight SNPs (-9625 G/A, -7255 A/G, -6274 C/T, -4884 A/G, -4155 C/G, -2760 A/C, -1873 G/A, and -563 C/T) were confirmed at a rare allele frequency of >5%. Association with schizophrenia and schizoaffective disorder was found for the -563 C/T SNP (p = .0005 for allele and p = .007 for genotype, Bonferroni corrected) and -9625 G/A SNP (p = .02 for allele and p = .03 for genotype, Bonferroni corrected). Significant differences in the 54 haplotypes formed by these eight SNPs were also found between patients and control subjects (p = .008, Bonferroni corrected). CONCLUSIONS: Because of the strength of these results and the location of these SNPs in the regulatory region of this gene, functional studies investigating the possible influence of these SNPs on receptor expression levels in schizophrenia are warranted.

Confidence Intervals↗

Performance of floating absolute risks.

A recent investigation of hormone replacement therapy and breast cancer risk used a method called "floating absolute risks" (FARs) to compute confidence intervals for relative hazards. This method has been used in other medical studies and has received controversy. This controversy stems from the correct implementation of this method. However, there has been no direct comparison of the FAR method, as it is sometimes incorrectly applied and reported, with the conventional approach for computing confidence intervals from proportional hazards regression. In this paper, the author reports simulation results comparing these two methods and demonstrates that the FAR method, when applied incorrectly, can produce confidence intervals that are substantially too narrow.

Confidence Intervals↗