PubMed Health⌕ Search

PubMed · 14735843

[Regression methods and causal inference: structural equations models].

Abstract

The estimate of correlations among observed outcomes is crucial in biomedical research, especially when the aim of the study is to infer, from the magnitude of these correlations, the causal influence of certain, sometimes latent, factors. In such situations, a typical regression approach, known as "structural equation models" (SEM), which was introduced in the 1970s, becomes significant. These models allow hypotheses to be formulated quite clearly, thanks to some explicit and rigorous graphical representations, on which the "path analysis" is based. SEM, which were initially used in economics, have in the past decade been applied in a wide variety of fields, especially in genetic epidemiology. It's in this field that SEM are extraordinarily effective, representing a simple yet powerful means of estimating the contribution of genes and the environment to the phenotypic expression of a given disease. To this end, data on twins are particularly useful, and in this case the correlation between the outcomes describes the extent of similarity of the twin phenotypes. From this standpoint, SEM undoubtedly constitute one of the most promising statistical tools for family studies and quantitative genetic research. The method can be easily extended to traditional epidemiology, and some interesting applications have already been developed in occupational and social epidemiology. In this paper, we describe in detail the SEM approach and discuss the use of these models in genetic epidemiology, using twin studies as an example. We also discuss the application of SEM in fields other than genetic research.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Corrado Fagnani, Rodolfo Cotichini, M Antonietta Stazi. [Regression methods and causal inference: structural equations models].. https://pubmed.ncbi.nlm.nih.gov/14735843/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The causal relationship between multiple cardiovascular diseases and glioblastoma: A Mendelian randomization study.

Observational studies suggest an association between glioblastoma (GBM) and cardiovascular diseases (CVDs), but a causal relationship remains unestablished. This study aimed to investigate the causal link between multiple CVDs and GBM risk. The inverse variance weighted method indicated that all 18 CVDs had significant causal associations with GBM (P&#x2005;<&#x2005;.05). Genetically predicted CVDs were uniformly associated with a lower risk of GBM (odds ratio&#x2005;<&#x2005;1), identifying them as potential protective factors. Sensitivity analyses confirmed the absence of significant heterogeneity or horizontal pleiotropy, and the MR-Steiger test validated the correct causal direction. This Mendelian randomization (MR) study provides evidence that a range of CVDs are causally associated with a decreased risk of developing GBM. These findings suggest shared biological pathways and offer new insights for understanding GBM etiology. We conducted a 2-sample MR analysis using publicly available genome-wide association study data. GBM was the outcome, and 18 cardiovascular-related traits (including coronary artery disease, myocardial infarction, and venous thromboembolism) were exposures. Instrumental variables were single-nucleotide polymorphisms significantly associated with exposures (P&#x2005;<&#x2005;5&#x2005;&#xd7;&#x2005;10-8). The primary analysis used the inverse variance weighted method, supplemented with MR-Egger, weighted median, and weighted mode methods. Sensitivity analyses, including Cochran Q test, MR-Egger intercept test, leave-one-out analysis, and MR-Steiger directionality test, were performed to ensure robustness.

Causality↗

Analysing and interpreting competing risk data.

When competing risks are present, two types of analysis can be performed: modelling the cause specific hazard and modelling the hazard of the subdistribution. This paper contrasts these two methods and presents the benefits of each. The interpretation is specific to the analysis performed. When modelling the cause specific hazard, one performs the analysis under the assumption that the competing risks do not exist. This could be beneficial when, for example, the main interest is whether the treatment works in general. In modelling the hazard of the subdistribution, one incorporates the competing risks in the analysis. This analysis compares the observed incidence of the event of interest between groups. The latter analysis is specific to the structure of the observed data and it can be generalized only to another population with similar competing risks.

Causality↗

Evaluating candidate agents of selective pressure for cystic fibrosis.

Cystic fibrosis is the most common lethal single-gene mutation in people of European descent, with a carrier frequency upwards of 2%. Based upon molecular research, resistances in the heterozygote to cholera and typhoid fever have been proposed to explain the persistence of the mutation. Using a population genetic model parameterized with historical demographic and epidemiological data, we show that neither cholera nor typhoid fever provided enough historical selective pressure to produce the modern incidence of cystic fibrosis. However, we demonstrate that the European tuberculosis pandemic beginning in the seventeenth century would have provided sufficient historical, geographically appropriate selective pressure under conservative assumptions. Tuberculosis has been underappreciated as a possible selective agent in producing cystic fibrosis but has clinical, molecular and now historical, geographical and epidemiological support. Implications for the future trajectory of cystic fibrosis are discussed. Our result supports the importance of novel investigations into the role of arylsulphatase B deficiency in cystic fibrosis and tuberculosis.

Causality↗