PubMed Health⌕ Search

PubMed · 12771550

A comprehensive method for genome scans.

Abstract

In applications involving the use of genome scans the problem of correcting for multiple testing figures prominently. A frequently used approach is the Bonferroni adjustment, but this is known to be often severely conservative. As an alternative we use the method of importance sampling to accurately and efficiently obtain required exceedance probabilities. This method is comprehensive in the sense that it has application to exceedance probabilities for other classes of test statistics, such as those for linkage disequilibrium or Hardy-Weinberg equilibrium at multiple loci. We illustrate the importance sampling technique by focusing on affected sib pair tests done at a large number of fully informative markers. We demonstrate how our approach can be used to obtain exceedance probabilities for arbitrary marker spacings, and we compare our approach with that of Feingold et al. [1993], which uses the method of large deviations and does not provide the means for adjusting for unequal marker spacing.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

James D Malley, Daniel Q Naiman, Joan E Bailey-Wilson. 2002. A comprehensive method for genome scans.. https://doi.org/10.1159/000070663

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Statistical test to compare the linkage model and the admixture model based on central limit results.

In the Admixture Model, the probability that an individual carries a certain allele at a specific marker depends on the allele frequencies in K ancestral populations and the proportion of the individual's genome originating from these populations. The markers are assumed to be independent. The Linkage Model is a Hidden Markov Model that extends the Admixture Model by incorporating linkage between neighboring loci. We prove consistency and asymptotic normality of maximum likelihood estimators for the ancestry of individuals in the Linkage Model, complementing earlier results by (Pfaff et al., 2004; Pfaffelhuber and Rohde, 2022; Heinzel, 2025) for the Admixture Model. These results are used to prove that a statistical test that allows for model selection between the Admixture Model and the Linkage Model is an asymptotic level-α-test. Finally, we demonstrate the practical relevance of our results by applying the test to real-world data from The 1000 Genomes Project Consortium (2015).

Genetic Linkage↗

Multivariate and multilocus variance components method, based on structural relationships to assess quantitative trait linkage via SEGPATH.

A general-purpose modeling framework for performing path and segregation analysis jointly, called SEGPATH (Province and Rao [1995] Stat. Med. 7:185-198), has been extended to cover "model-free" robust, variance-components linkage analysis, based on identity-by-descent (IBD) sharing. These extended models can be used to analyze linkage to a single marker or to perform multipoint linkage analysis, with a single phenotype or multivariate vector of phenotypes, in pedigrees. Within a single, consistent approach, SEGPATH models can perform segregation analysis, path analysis, linkage analysis, or combinations thereof. SEGPATH models can incorporate environmental or other measured covariate fixed effects (including measured genotypes), genotype-specific covariate effects, population heterogeneity models, repeated-measures models, longitudinal models, autoregressive models, developmental models, gene-by-environment interaction models, etc., with or without linkage components. The data analyzed can have any missing value structure (assumed missing at random), with entire individuals missing, or missing on one or more measurements. Corrections for ascertainment can be made on a vector of phenotypes and/or other measures. Because of the flexibility of the class of models, the SEGPATH approach can also be used in nongenetic applications where there is a hierarchical structure, such as longitudinal, repeated-measures, time series, or nested models. A variety of specific models are provided, as well as some comparisons with other linkage analysis models. Particular applications demonstrate the importance of correctly accounting for the extraneous sources of familial resemblance, as can be done easily with these SEGPATH models, so as to give added power to detect linkage as well as to protect against spuriously inferring linkage.

Genetic Linkage↗