PubMed Health⌕ Search

PubMed · 12548674

Multivariate and multilocus variance components method, based on structural relationships to assess quantitative trait linkage via SEGPATH.

Abstract

A general-purpose modeling framework for performing path and segregation analysis jointly, called SEGPATH (Province and Rao [1995] Stat. Med. 7:185-198), has been extended to cover "model-free" robust, variance-components linkage analysis, based on identity-by-descent (IBD) sharing. These extended models can be used to analyze linkage to a single marker or to perform multipoint linkage analysis, with a single phenotype or multivariate vector of phenotypes, in pedigrees. Within a single, consistent approach, SEGPATH models can perform segregation analysis, path analysis, linkage analysis, or combinations thereof. SEGPATH models can incorporate environmental or other measured covariate fixed effects (including measured genotypes), genotype-specific covariate effects, population heterogeneity models, repeated-measures models, longitudinal models, autoregressive models, developmental models, gene-by-environment interaction models, etc., with or without linkage components. The data analyzed can have any missing value structure (assumed missing at random), with entire individuals missing, or missing on one or more measurements. Corrections for ascertainment can be made on a vector of phenotypes and/or other measures. Because of the flexibility of the class of models, the SEGPATH approach can also be used in nongenetic applications where there is a hierarchical structure, such as longitudinal, repeated-measures, time series, or nested models. A variety of specific models are provided, as well as some comparisons with other linkage analysis models. Particular applications demonstrate the importance of correctly accounting for the extraneous sources of familial resemblance, as can be done easily with these SEGPATH models, so as to give added power to detect linkage as well as to protect against spuriously inferring linkage.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

M A Province, T K Rice, I B Borecki, C Gu, A Kraja, D C Rao. 2003. Multivariate and multilocus variance components method, based on structural relationships to assess quantitative trait linkage via SEGPATH.. https://doi.org/10.1002/gepi.10208

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Statistical test to compare the linkage model and the admixture model based on central limit results.

In the Admixture Model, the probability that an individual carries a certain allele at a specific marker depends on the allele frequencies in K ancestral populations and the proportion of the individual's genome originating from these populations. The markers are assumed to be independent. The Linkage Model is a Hidden Markov Model that extends the Admixture Model by incorporating linkage between neighboring loci. We prove consistency and asymptotic normality of maximum likelihood estimators for the ancestry of individuals in the Linkage Model, complementing earlier results by (Pfaff et al., 2004; Pfaffelhuber and Rohde, 2022; Heinzel, 2025) for the Admixture Model. These results are used to prove that a statistical test that allows for model selection between the Admixture Model and the Linkage Model is an asymptotic level-α-test. Finally, we demonstrate the practical relevance of our results by applying the test to real-world data from The 1000 Genomes Project Consortium (2015).

Genetic Linkage↗

Linkage analysis with sequential imputation.

Multilocus calculations, using all available information on all pedigree members, are important for linkage analysis. Exact calculation methods in linkage analysis are limited in either the number of loci or the number of pedigree members they can handle. In this article, we propose a Monte Carlo method for linkage analysis based on sequential imputation. Unlike exact methods, sequential imputation can handle large pedigrees with a moderate number of loci in its current implementation. This Monte Carlo method is an application of importance sampling, in which we sequentially impute ordered genotypes locus by locus, and then impute inheritance vectors conditioned on these genotypes. The resulting inheritance vectors, together with the importance sampling weights, are used to derive a consistent estimator of any linkage statistic of interest. The linkage statistic can be parametric or nonparametric; we focus on nonparametric linkage statistics. We demonstrate that accurate estimates can be achieved within a reasonable computing time. A simulation study illustrates the potential gain in power using our method for multilocus linkage analysis with large pedigrees. We simulated data at six markers under three models. We analyzed them using both sequential imputation and GENEHUNTER. GENEHUNTER had to drop between 38-54% of pedigree members, whereas our method was able to use all pedigree members. The power gains of using all pedigree members were substantial under 2 of the 3 models. We implemented sequential imputation for multilocus linkage analysis in a user-friendly software package called SIMPLE.

Genetic Linkage↗