PubMed HealthSearch

SEARCH · PubMed Health

Results for “structural variations”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 recordsLinked to original sources

Integrative Genotyping and Analysis of Canine Structural Variation Using Long-read and Short-read Data.

Structural variation makes an important contribution to canine evolution and phenotypic differences. Although recent advances in long-read sequencing have enabled the generation of multiple canine genome assemblies, most prior analyses of structural variation have relied on short-read sequencing. To offer a more complete assessment of structural variation in canines, we performed an integrative analysis of structural variants present in 12 canine samples with available long-read and short-read sequencing data along with genome assemblies. Use of long-reads permits the discovery of heterozygous variation that is absent in existing haploid assembly representations while offering a marked increase in the ability to identify insertion variants relative to short-read approaches. Examination of the size spectrum of structural variants shows that dimorphic LINE-1 and SINE variants account for over 45% of all deletions and identified 1,410 LINE-1s with intact open reading frames that show presence-absence dimorphism. Using a graph-based approach, we genotype newly discovered structural variants in an existing collection of 1,879 resequenced dogs and wolves, generating a variant catalog containing a 56.5% increase in the number of deletions and 705% increase in the number of insertions previously found in the analyzed samples. Examination of allele frequencies across admixture components present across breed clades identified 283 structural variants evolving with a signature of selection.

Animals

Structural variation detection and association analysis of whole-genome-sequence data from 16,543 Alzheimer's disease sequencing project subjects.

INTRODUCTION: The role of structural variations (SVs) in Alzheimer's disease (AD) remains understudied. METHODS: We analyzed whole-genome sequencing data from the Alzheimer's Disease Sequencing Project (N&#xa0;=&#xa0;16,543) and identified 400,234 (168,223 high-quality) SVs. Laboratory validation yielded a sensitivity of 82% (85% for high-quality). RESULTS: We found a burden of singletons (odds ratio [OR]&#xa0;=&#xa0;1.07, p&#xa0;=&#xa0;0.0017) and homozygous deletions (OR&#xa0;=&#xa0;1.14, p&#xa0;<&#xa0;0.0001) in cases. On AD genes, we observed the ultra-rare SVs associated with the disease, including protein-altering SVs in ABCA7, APP, PLCG2, and SORL1. Twenty-one SVs are in linkage disequilibrium (LD) with known AD-risk variants, exemplified by a 5k deletion in LD (R2&#xa0;=&#xa0;0.99) with rs143080277 in NCK2. We identified a rare deletion near RNA5SP293 associated with AD (OR&#xa0;=&#xa0;1.99, p&#xa0;=&#xa0;1.3&#xa0;&#xd7;&#xa0;10-5), which was replicated using an independent dataset. DISCUSSION: This study highlights the pivotal role of SVs in AD genetics. HIGHLIGHTS: Observed a significant burden of singletons and homozygous deletions in Alzheimer's disease (AD) patients. Identified rare protein-altering structural variations (SVs) in ABCA7, APP, PLCG2, and SORL1. Established linkages between SVs and AD risk-associated single nucleotide variants (SNVs). Discovered a novel deletion near RNA5SP293 linked to AD, replicated independently. Uncovered over-representation of SVs in neuronal function pathways.

Humans

Comparative genomics reveals lineage-associated structural variation and diversification in a barley fungal pathogen.

Leaf rust, caused by Puccinia hordei, is a major barley disease worldwide. Despite repeated shifts in virulence, contrasting reproductive histories, and emerging fungicide insensitivity, the genomic basis of its diversification and adaptation remains poorly understood. In this study, we generated haplotype-resolved, chromosome-level genome assemblies for two isolates with contrasting virulence and analyzed 41 Australian isolates collected over 54&#x2009;yr (1966-2020), integrating comparative and population genomics, mating-type gene phylogenies, chromosome-specific k-mer profiling, genome-wide copy-number variation (CNV) analysis, and gene-expression analysis. We identified a structurally dynamic chromosome characterized by repeat-associated rearrangements, structural variation, and lineage-associated CNV, representing the first evidence in a rust fungus of chromosome-scale structural diversification of this extent. Population analyses distinguished clonally expanded lineages from recombination-associated lineages, with mating-type gene phylogenies providing further support for lineage differentiation. More recently collected isolates showed increased duplication-associated variation, and CNV boundaries were associated with structural-variant breakpoints. We also identified lineage-associated amplification of Cyp51, with increased copy number associated with higher transcript abundance, supporting a potential role in fungicide adaptation. Overall, our findings highlight structural variation, contrasting reproductive histories, and lineage-associated CNV as important contributors to diversification in P. hordei, providing insights for future rust pathogen surveillance and management strategies.

Cyp51 gene

Charting host structural variations in cervical cancer by long-read sequencing pinpoints a functional deletion in PIAS1.

Host structural variations (SVs) are critical in cancer development but their landscape and interaction with HPV integration in cervical carcinogenesis remain unclear. In this study, we performed Nanopore long-read sequencing on five HPV-positive cervical cancer tissues and two cell lines to profile host SVs. We identified thousands of SVs and statistically demonstrated their significant enrichment in genomic windows &#xb1;25 to &#xb1;50&#xa0;kb from HPV integration sites. Cross-sample analysis revealed 60 shared SVs, including a recurrent deletion within the PIAS1 gene. Multi-omics integration (Hi-C, H3K27ac ChIP-seq, and TCGA data) showed that this deletion is associated with reduced PIAS1 expression, disruption of local topologically associating domains, advanced pathological tumor stage, and poorer overall survival. Functional assays confirmed that PIAS1 deficiency inhibits cervical cancer cell proliferation and migration. Our findings identify a PIAS1 deletion as a candidate driver event, and underscore the pivotal role of host genomic instability in HPV-associated oncogenesis.

Cervical cancer

Severus detects somatic structural variation and complex rearrangements in cancer genomes using long-read sequencing.

For the detection of somatic structural variation (SV) in cancer genomes, long-read sequencing is advantageous over short-read sequencing with respect to mappability and variant phasing. However, most current long-read SV detection methods are not developed for the analysis of tumor genomes characterized by complex rearrangements and heterogeneity. Here, we present Severus, a breakpoint graph-based algorithm for somatic SV calling from long-read cancer sequencing. Severus works with matching normal samples, supports unbalanced cancer karyotypes, can characterize complex multibreak SV patterns and produces haplotype-specific calls. On a comprehensive multitechnology cell line panel, Severus consistently outperforms other long-read and short-read methods in terms of SV detection F1 score (harmonic mean of the precision and recall). We also illustrate that compared to long-read methods, short-read sequencing systematically misses certain classes of somatic SVs, such as insertions or clustered rearrangements. We apply Severus to several clinical cases of pediatric leukemia/lymphoma, revealing clinically relevant cryptic rearrangements missed by standard genomic panels.

Humans

The landscape of structural variation in pediatric cancer.

Structural variants (SVs) account for over 60% of the driver variants in pediatric cancer, and in many cases act as the cancer initiating event. To study SVs from a pan-cancer perspective, we analyzed 1,616 pediatric cancer genomes in 16 major cancer types of hematological malignancies (n = 908), brain tumors (n = 183), and solid tumors (n = 525) and compared their profiles to those of 2,203 adult cancers. The SV burden varied ~100-fold across pediatric cancer types and demonstrated an 8- to 16-fold reduction compared to adult brain and solid tumors but was comparable in pediatric versus adult hematological malignancies. Recurrent SV hotspots occurred uniquely in pediatric acute lymphoblastic leukemias (ALLs) in proximity to RAG-mediated recombination signal sequences (RSS) and disrupted multiple immune-related loci as well as 69 genes, which often involved cryptic RSS sites. By contrast, such hotspots affected only immune-related loci but not driver genes in adult lymphoid cancers. Eight SV signatures extracted from the cohort had varying distributions across cancer types, with clustered translocations reflecting templated insertions in osteosarcoma, and medium-sized deletions (10 kb to 1 Mb) enriched in cancers with RAG-mediated deletions. Intra-patient evolutionary analysis in 13 patients with multiple spatiotemporally distinct samples revealed that RAG-mediated recombination in leukemia and complex rearrangements in solid tumors occurred both early in disease initiation and continuously during later diversification, contributing to clonal heterogeneity. Finally, we found that both driver genes and fragile sites were the two genomic regions most frequently disrupted by SVs. The unique and diverse SV landscapes that emerged from this comprehensive analysis expand the scope of RSS-mediated mutagenesis in pediatric ALL and will be a valuable resource for guiding future functional studies and the design of clinical genomic testing in pediatric cancer.

Journal Article

Analysis of deep-resequencing data of 984 soybean accessions reveals structural variations underlying agronomic traits.

Genomic structural variants (SVs) are major sources of genetic variation and have profound impacts on phenotypic traits. However, their functional effects remain largely unexplored in soybean. Here, we resequence 940 soybean accessions. Together with 44 publicly available datasets, we identify 602,281 SVs. Using a graph-based genome, we detect&#xa0;an additional 58,760 presence/absence variations (PAVs) that broadly affect gene expression. Population genomic analyses reveal that SVs serve as a core driving force for soybean domestication and improvement. Integrating SVs with QTLs for oil and protein&#xa0;content, and performing GWAS on 27 traits, we identify key functional SVs. These include transposable element insertions altering seed coat color, multiple insertions within a cytochrome P450 gene modifying flower and hypocotyl color, and a GmMATE1 deletion enhancing seed size. Together, our study establishes a comprehensive SV map of soybean, offering a valuable resource for dissecting the genetic basis of complex traits to accelerate molecular breeding.

Glycine max

SVbyEye: a visual tool to characterize structural variation among whole-genome assemblies.

MOTIVATION: We are now in the era of being able to routinely generate highly contiguous (near telomere-to-telomere) genome assemblies of human and nonhuman species. Complex structural variation and regions of rapid evolutionary turnover are being discovered for the first time. Thus, efficient and informative visualization tools are needed to evaluate and directly observe structural differences between two or more genomes. RESULTS: We developed SVbyEye, an open-source R package to visualize and annotate sequence-to-sequence alignments along with various functionalities to process these alignments. The tool facilitates the characterization of complex structural variants in the context of sequence homology helping resolve the mechanisms underlying their formation. AVAILABILITY AND IMPLEMENTATION: SVbyEye is available on GitHub (https://github.com/daewoooo/SVbyEye) and via Zenodo (https://doi.org/10.5281/zenodo.15303553).

Software

Complex structural variation, phylogeny, and disease associations of the mucin pangenome.

Mucins are large glycoproteins that provide hydration and barrier function to epithelial tissues. Although genetically heterogeneous, all mucins harbor a large exon composed of variable number tandem repeats (VNTRs). Short-read sequencing has limited our understanding of mucin VNTR diversity and makes disease association studies challenging. We leverage 296 long-read phased genome assemblies to characterize 14 mucin family members, achieving &#x2265;97% accuracy across 572 haplotypes. Phylogenetic haplogroup analysis reveals extraordinary structural heterozygosity, with MUC4 harboring the greatest allelic diversity (n=240 distinct lengths) and MUC12 the greatest size range (&#x394; = 55,233 bp; 23,080 amino acids). Ten mucins show significant population stratification (pFDR < 0.05). At the MUC4/MUC20 locus, we characterize higher-order structural variation, including a recurrent inversion, copy number variation, and interlocus gene conversion. Optimized genotyping achieves &#x2265;95% haplogroup concordance across 10 loci. We apply this to 4,637 deeply phenotyped cystic fibrosis patients and identify a significant association between short MUC1 VNTRs and severe disease (p=0.0056), demonstrating the pangenome's utility for complex locus genotyping and disease discovery.

Journal Article

Pan-genome-based resequencing of 2,320 accessions reveals structural variations and accelerates breeding advances in cultivated peanut.

The cultivated peanut is a crucial global legume crop that is essential for food security and nutrition, particularly in developing regions. However, its limited genetic variation hampers breeding progress and yield improvement. Here we constructed a graph-based pan-genome for peanut, incorporating 14 genomes that represent all 6 peanut varieties. Using this pan-genome, we genotyped 2,320 accessions, covering 88.03% of ICRISAT and 59.21% of USDA core germplasm, enriching valuable resources for genomic studies and breeding. We cataloged genomic structural variations and investigated the role of homoeologous exchanges in population divergence. Through our pan-genome approach, we overcame the challenges of genotyping posed by homoeologous exchanges and identified key genes associated with flowering and dwarfism in peanut. By integrating superior haplotypes and germplasm resources guided by the pan-genome, we further developed high-yield dwarf lines. This work provides essential genomic resources to accelerate functional gene discovery and modern peanut breeding.

Journal Article

Substantial non-homologous recombination and structural variation results from Brassica AABC and CCAB hybrid meiosis.

Meiotic crossovers contribute to genetic diversity and play a crucial role in homologous chromosome segregation. Non-homologous crossovers in Brassica, involving the exchange of genetic material between genomes, can be valuable for transferring novel traits or characteristics between Brassica species. However, there are a limited number of studies that specifically investigate crossover frequencies in populations of interspecific hybrids. We investigated the distribution and frequency of homologous crossover events, as well as non-homologous recombination and structural variation, in hybrids between B. juncea (AABB)&#x2009;&#xd7;&#x2009;B. napus (AACC) (resulting in AABC hybrids; 5 genotypes) and B. napus (AACC)&#x2009;&#xd7;&#x2009;B. carinata (BBCC) (resulting in CCAB hybrids; 4 genotypes). The analysis was performed on individuals derived from microspore culture of both unreduced and reduced gametes produced by the AABC and CCAB hybrids. All AABC and almost all CCAB unreduced gamete-derived individuals and most AABC and CCAB reduced gamete-derived individuals showed copy number variation indicative of non-homologous (A-C) recombination. Additionally, a higher frequency of homologous crossovers, also in centromeric and pericentromic regions, was observed in the diploid genomes of the AABC and CCAB hybrids. Overall, these hybrid types show high frequencies of A-C introgressions, which may be useful in B. juncea or B. carinata introgression breeding, and this increased recombination frequency may help break up existing linkage disequilibrium blocks in the Brassica A and C genomes.

Meiosis

European ash pangenome reveals widespread structural variation and genetic basis of low ash dieback susceptibility.

European Ash (Fraxinus excelsior) is a keystone tree species, whose populations are being decimated by ash dieback disease (ADB) - better characterisation of genetic variants associated with low susceptibility to the disease is needed. Here, we develop a F. excelsior pangenome to more fully capture sequence variability within this species compared with a linear reference genome, using a geographically diverse set of fifty F. excelsior samples. We identify 362,965 structural variants (SVs), including 174&#x2009;Mb of sequence absent from the linear reference genome (22% of the linear reference size), and identify 3,412 high-confidence dispensable genes (those present only in some individuals). We use the pangenome to analyse existing genomic data from over 1,200 individuals, revealing 220 single nucleotide polymorphisms (SNPs) showing consistent allele frequency shifts between healthy individuals and those highly damaged by ADB, across UK seed sources, explicitly demonstrating the existence of a shared genetic component to low ADB susceptibility.

Polymorphism, Single Nucleotide

Recurrent structural variation and recent turnover at the 17q21.31 locus in humans and great apes.

The 17q21.31 locus in humans harbors several complex structural haplotypes including a ~970kb inversion. Different inversion haplotypes have been associated with susceptibility to microdeletions causing Koolen-de Vries syndrome and variation in fecundity and recombination rates. Here, using 210 haplotype-resolved human genome assemblies and pangenome graph-based approaches we characterize 11 distinct structural haplotypes, several of which have not been previously described. Extending our analyses to a set of haplotype-resolved great-ape genomes, we characterize the structure of an independent inversion in chimpanzees which extends an additional 650kb, encompasses 5 additional genes, and is ~2 million years younger than the human inversion. We further determine that gorillas exhibit an independent duplication of the KANSL1 gene which may predispose them to Koolen-de Vries syndrome causing microdeletions. Using short read sequencing data we characterize 17q21.31 haplotype diversity worldwide in ~5174 individuals from 107 populations finding increased frequencies of KANSL1 duplication-containing haplotypes in both European and South Asian populations as well as 8 double recombination events between inverted and non-inverted haplotypes ranging in size from 20-180kb. Finally, using 626 ancient Eurasian human genomes we show the frequency of haplotypes containing KANSL1 duplications has increased ~6-fold over the past 12 thousand years in Europe. Together, our results highlight the dynamics, complexity, and recurrent, independent evolution of a medically relevant locus across humans and great apes.

Journal Article

RAG-mediated structural variation and its impact on relapse risk in acute lymphoblastic leukemia.

Relapse during treatment of B-cell acute lymphoblastic leukemia (B-ALL) is a harbinger of poor outcomes. Identifying biomarkers for subsequent relapse risk which are detectable at B-ALL diagnosis remains a priority. Off-target recombination-activating gene (RAG)-mediated structural variants (SVs) generate genomic instability that drives leukemogenesis and may underlie treatment resistance. Leveraging sequencing data in 1,496 pediatric B-ALL patients enriched for relapse status (relapse n=532; non-relapse n=964), we characterized RAG-mediated SVs across B-ALL molecular subtypes and examined their association with patient characteristics and their impact on clinical outcomes. Off-target RAG-mediated SVs were overall frequent, particularly in ETV6::RUNX1, ETV6::RUNX1-like, and Ph-like B-ALL subtypes, while increasing age-at-diagnosis was positively associated with burden of off-target RAG-mediated SVs (P<.001). Off-target RAG-mediated SVs with a recombination signal sequence (RSS) at one breakpoint, a hallmark of off-target RAG activity, were significantly more frequent at diagnosis in patients who subsequently relapsed (P=.001). This association remained significant in multivariable regression analysis (per SV odds ratio [OR]:1.08, 95%CI:1.04-1.12), in minimal residual disease (MRD)-negative patients (OR:1.09, 95%CI:1.04-1.14) and across subtypes. Excluding deletions, MRD-negative ETV6::RUNX1 patients with &#x2265;3 off-target RAG-mediated SVs had a >3-fold risk of relapse (hazard ratio:3.47, 95% CI:1.86-6.49). RAG-mediated SVs were also associated with relapse risk in T-cell ALL patients. Off-target RAG-mediated SV burden at diagnosis is a risk factor of relapse in pediatric ALL across molecular subtypes and independent of MRD status.

Journal Article

[Morphological and structural variations of the human inguinal region (author's transl)].

In the inguinal region, numerous muscular and fibrous alterations are described. They are related to the unconstant position of the pubic tubercle in relation io the interspinous diameter (linea bi-spinalis). The pubic tubercle can be observed in two different locations: either high or low. The high location is characterized by the presence of the pubic tubercle at a distance of 5 to 7.5 cm below the interspinous diameter. It must be considered as normal and is found in 65% of the subjects. In the low locations, the distance between spinous tubercle and interspinous diameter reaches 7,5 to 12 cm. It is an important abnormality which interests 35% of the subjects. The lower the pubic tubercle are located, the more often morphological alterations are to be found in the following structures: obliquus externus, obliquus internus, transversus and cremaster muscles as well as fascia transversalis. Nevertheless, the pyramidalis muscle as well as the inguinal ligamentary formation, Hesselbach's interfoveolar ligament and Thompson's iliopubic tract do not follow that rule, since the important morphological variations of these deep fibrous components can never be related to the distance between pubic tubercle and interspinous diameter. The functional signification of the inguinal region and especially of the inguinal canal is modified by those ostelogical, muscular and ligamentary variations.

Humans

Effects of structural variation in beta-monoglycerides and other lipids on ordering in synthetic membranes.

Studies of beta-monoglyceride multilayers were carried out using a variety of spin probes. Effects of variables such as chain length, unsaturation, and branching on organization of acyl chains in lipids of model membranes were assessed. In addition, effects of added cholesterol on membrane order were determined. Results indicated that pure beta-monolaurin yields highly ordered films, whereas, unsaturated glycerides such as beta-monoolein, beta-monolinolein, and analogous lecithins yield fluid films. Branched monoglycerides behaved similarly to beta-monoolein, suggesting that branching in acyl chains is an effective substitute for unsaturation in maintaining membrane integrity. Multilayers of beta-monoglycerides exhibited similar properties to those of more complex lipids such as phospholipids. beta-Monoglycerides, by virtue of the presence of a single acyl chain, provided a relatively simple and effective alternative to the use of phospholipids in studies of membrane architecture.

Binding Sites