PubMed Health⌕ Search

PubMed · 16859555

A strategy for extracting and analyzing large-scale quantitative epistatic interaction data.

Abstract

Recently, approaches have been developed for high-throughput identification of synthetic sick/lethal gene pairs. However, these are only a specific example of the broader phenomenon of epistasis, wherein the presence of one mutation modulates the phenotype of another. We present analysis techniques for generating high-confidence quantitative epistasis scores from measurements made using synthetic genetic array and epistatic miniarray profile (E-MAP) technology, as well as several tools for higher-level analysis of the resulting data that are greatly enhanced by the quantitative score and detection of alleviating interactions.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Sean R Collins, Maya Schuldiner, Nevan J Krogan, Jonathan S Weissman. 2006. A strategy for extracting and analyzing large-scale quantitative epistatic interaction data.. https://doi.org/10.1186/gb-2006-7-7-r63

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Evaluation of epistasis detection methods for quantitative phenotypes.

MOTIVATION: Epistasis, or genetic interaction, plays a crucial role in shaping complex traits and has been increasingly recognized for its widespread influence in genetic architectures. While epistasis detection has been extensively evaluated in case-control studies, its performance with quantitative phenotypes remains comparatively understudied. RESULTS: We identified and evaluated six epistasis detection methods applicable to quantitative trait analysis: EpiSNP, Matrix Epistasis, MIDESP, PLINK Epistasis, QMDR, and REMMA. Using the EpiGEN simulator, we generated synthetic datasets modeling four classes of pairwise SNP interactions-dominant, multiplicative, recessive, and XOR. We also assessed BOOST and MDR algorithms using discretized (case-control) versions of the same datasets. Performance varied notably by interaction type: REMMA achieved the highest overall detection rate (55%), particularly excelling with dominant interactions (100%). MDR excelled with multiplicative (57%) and XOR (69%) interactions. Meanwhile, EpiSNP attained the best performance for recessive interactions (67%). All methods except BOOST produced F1 scores below 0.05 for most interaction types. We further evaluated the methods using a real-world dataset. When applied to the Adolescent Brain Cognitive Development dataset to analyse the externalizing behavior phenotype, both PLINK Epistasis and PLINK BOOST identified SNPs within the DRD2 and DRD4 genes, consistent with previously reported genetic associations. Given the variability in tool performance across interaction types, no single method provides optimal detection across all scenarios. Leveraging multiple detection algorithms may therefore yield more comprehensive insights into epistatic effects in quantitative trait analyses. AVAILABILITY AND IMPLEMENTATION: All relevant code and simulated datasets can be found at github.com/staslist/Epistasis_Review repository.

Epistasis, Genetic↗

Testing for Genetic Interactions in Complex Disease With Distance Correlation.

Understanding epistasis (genetic interaction) may shed some light on the genomic basis of common diseases, including disorders of maximum interest due to their high socioeconomic burden, like schizophrenia. Distance correlation is an association measure that characterizes general statistical independence between random variables, not only the linear one. Here, we propose distance correlation as a novel tool for the detection of epistasis from case-control data of single-nucleotide polymorphisms. On the methodological side, we highlight the derivation of the explicit asymptotic null distribution of the test statistic. We show that this is the only way to obtain enough computational speed for the method to be used in practice, in a scenario where the resampling techniques found in the literature are impractical. Our simulations show satisfactory calibration of significance, as well as comparable or better power than existing methodology. We conclude with the application of our technique to a schizophrenia genetics dataset, obtaining biologically sound insights.

Epistasis, Genetic↗

Robustness-epistasis link shapes the fitness landscape of a randomly drifting protein.

The distribution of fitness effects of protein mutations is still unknown. Of particular interest is whether accumulating deleterious mutations interact, and how the resulting epistatic effects shape the protein's fitness landscape. Here we apply a model system in which bacterial fitness correlates with the enzymatic activity of TEM-1 beta-lactamase (antibiotic degradation). Subjecting TEM-1 to random mutational drift and purifying selection (to purge deleterious mutations) produced changes in its fitness landscape indicative of negative epistasis; that is, the combined deleterious effects of mutations were, on average, larger than expected from the multiplication of their individual effects. As observed in computational systems, negative epistasis was tightly associated with higher tolerance to mutations (robustness). Thus, under a low selection pressure, a large fraction of mutations was initially tolerated (high robustness), but as mutations accumulated, their fitness toll increased, resulting in the observed negative epistasis. These findings, supported by FoldX stability computations of the mutational effects, prompt a new model in which the mutational robustness (or neutrality) observed in proteins, and other biological systems, is due primarily to a stability margin, or threshold, that buffers the deleterious physico-chemical effects of mutations on fitness. Threshold robustness is inherently epistatic-once the stability threshold is exhausted, the deleterious effects of mutations become fully pronounced, thereby making proteins far less robust than generally assumed.

Epistasis, Genetic↗