PubMed HealthSearch

Biomedical subjects

Wei Zhao

Publications and source records attributed to Wei Zhao.

At least 19 recordsLinked to original sources

The Wild Soybean C3HC4-Type RING Zinc-Finger Protein ZFP4 Enhances Resistance to Soybean Mosaic Virus.

Soybean [Glycine max (L.) Merr.] is a globally important source of protein and edible oil, but is severely threatened by soybean mosaic virus (SMV). Wild soybean [Glycine soja Sieb. & Zucc.], the wild ancestor of cultivated soybean, exhibits high genetic diversity and strong resistance to pathogens. In this study, we identified a novel SMV resistance locus RSC7-4 and its candidate gene ZFP4 from wild soybean, encoding a C3HC4-type RING zinc-finger protein. The knockout mutants of ZFP4 showed enhanced susceptibility to SMV strains SC7 and SC3, while its overexpressing lines conferred resistance without yield penalty; ZFP4 mediates resistance by inhibiting GSTT1 to increase glutathione and reduce excessive reactive oxygen species accumulation. Domestication analysis revealed reduced genetic diversity of ZFP4 in cultivated soybean, with the resistant ZFP4Hap1 underutilized in breeding. In summary, this study provides not only excellent genetic resources for SMV-resistant soybean breeding but also new insights into the regulatory mechanisms of soybean resistance to SMV.

ZFP4

Association of Enterocytozoon bieneusi Infection with chronic/persistent diarrhea and ITS genotypic diversity: a hospital-based case-control study in Suburban Shanghai, China.

Enterocytozoon bieneusi is a globally distributed zoonotic enteric pathogen that remains largely overlooked in routine diarrheal disease surveillance. Although previous studies in Shanghai, China, have reported elevated prevalence in diarrheal populations, case-control data from suburban areas at the peri&#x2011;urban interface and the strength of the association between E. bieneusi infection and chronic diarrhea in non-immunocompromised individuals remain poorly characterized. We performed a hospital-based case-control study in suburban Shanghai, enrolling 286 diarrheal outpatients without documented immunodeficiency and 138 asymptomatic controls frequency-matched for age and sex. Fecal specimens were collected and subjected to genomic DNA extraction. E. bieneusi was detected via nested PCR amplification of the ribosomal internal transcribed spacer (ITS) region. Factors associated with infection were identified using multivariate logistic regression. Genotypic diversity and zoonotic potential were assessed by Sanger sequencing and phylogenetic analysis. The overall prevalence of E. bieneusi was 12.2% (35/286) in diarrheal patients, significantly higher than the 2.2% (3/138) observed in asymptomatic controls (P < 0.001). E. bieneusi positivity was independently associated with chronic/persistent diarrhea (adjusted odds ratio = 2.63, 95% confidence interval: 1.25-5.54, P = 0.011). Fourteen distinct ITS genotypes were identified, comprising five known genotypes (D, EbpD, SHW7, Henan-III, and CHG5) and nine novel genotypes (designated SHH2 to SHH10). Thirteen genotypes clustered within Group 1, and one genotype (CHG5) fell within Group 2, two phylogenetic groups that contain genotypes with documented zoonotic potential in global surveillance. E. bieneusi was detected at a relatively high prevalence among diarrheal patients in suburban Shanghai, and its detection was associated with chronic/persistent diarrhea. The predominance of zoonotic genotypes and the identification of nine novel Group 1 genotypes indicate phylogenetic similarity to known zoonotic lineages and warrant further investigation of local zoonotic transmission; no animal or environmental samples were analyzed in this study. These findings suggest that E. bieneusi testing may be considered as part of the differential diagnosis for patients with unexplained chronic/persistent diarrhea and highlight the need for One Health surveillance in the surveyed area.

Diarrhea

NS2A V89F mutation in a DENV1 clinical isolate enhances neurotropism and neuroinvasion.

INTRODUCTION: Dengue virus (DENV) neurological complications are increasingly reported, yet the viral genetic determinants of neurotropism remain poorly characterized. METHODS: We screened 25 DENV1 clinical isolates from the 2014 outbreak in Guangdong, China, for neurotropism in suckling mice, and integrated comparative genomics, pre-expression functional assays, population-scale sequence analysis, and OpenFold3 structural modeling to identify mutations associated with enhanced neuroinvasion. RESULTS: We found that only strain P1253 induced neurological symptoms and mortality via subcutaneous inoculation, producing cortical-selective lesions distinct from the diffuse encephalitic damage observed after intracranial inoculation, and P1253 replicated preferentially in human brain microvascular endothelial cells (HBMEC) compared to contemporaneous strains. Comparative genomics identified three unique mutations in P1253 (NS1 175Y&#x2192;H, NS2A 89V&#x2192;F, NS4A 2V&#x2192;I), and pre-expression assays demonstrated that only NS2A 89V&#x2192;F significantly enhanced viral replication and cytopathic effect in HBMEC. Analysis of 1,990 complete DENV1 genomes revealed five natural mutant types in the NS2A 89 -96 residue region, with P1253 representing the FIPI quadruple-mutant type, and OpenFold3 structural prediction showed that 89V&#x2192;F introduced on the VIPI background induced the most significant distal domain reorientation (RMSD 1.605 &#xc5;), increasing the centroid-to-centroid distance between residues 89 -96 and 185 -218 from 18.221 &#xc5; to 27.462 &#xc5;. DISCUSSION: These findings identify NS2A 89V&#x2192;F as a candidate adaptive mutation associated with enhanced neurotropism in DENV1 and provide a framework for monitoring neurovirulent variants.

Dengue Virus

CARS1 as a Prognostic Biomarker and Candidate Therapeutic Vulnerability in Hepatocellular Carcinoma: Insights Into Tumor Progression and the Immune Microenvironment.

BACKGROUND: Cysteinyl-tRNA synthetase 1 (CARS1) has been included in ferroptosis-related prognostic signatures, but its clinicopathological relevance, cellular functions, and relationship with the immune microenvironment in hepatocellular carcinoma (HCC) remain incompletely characterized. METHODS: Transcriptomic and clinical data from The Cancer Genome Atlas Liver Hepatocellular Carcinoma (TCGA-LIHC) dataset were integrated with corresponding data from an institutional HCC tissue cohort of 60 patients. CARS1 expression was evaluated by immunohistochemistry, and immune infiltration was examined using single-sample gene-set enrichment analysis (ssGSEA) and multiplex immunofluorescence, as well as by analyzing public single-cell datasets. The effects of CARS1 depletion were evaluated in MHCC97H and Hep3B cells using Cell Counting Kit-8 (CCK-8) assays, cell-cycle profiling, wound-healing assays, Transwell migration assays, western blotting, and erlotinib-sensitivity assays. RESULTS: CARS1 expression was elevated in HCC and was associated with adverse clinicopathological features and poor overall survival. Quantitative immunohistochemistry confirmed elevated CARS1 protein expression in tumor tissues. CARS1 depletion inhibited cell proliferation, altered cell-cycle distribution, impaired migration, and enhanced in vitro sensitivity to erlotinib. High CARS1 expression was also associated with increased infiltration of Th2-like immune cells. CONCLUSIONS: Elevated CARS1 expression is associated with an adverse biological and immune phenotype in HCC. These clinical, histopathological, and loss-of-function findings support further investigation of CARS1 as a prognostic marker and candidate therapeutic target in HCC, although additional mechanistic and in vivo validation is required.

Humans

Carboxyl group number and acidity of organic acids regulate structural reorganization and low glycemic index in cassava pyrodextrins via molecular interactions.

Transforming high-glycemic cassava starch into functional dietary fiber via pyrodextrinization is a promising way to valorize tuber crops, yet the molecular mechanisms catalyzed by organic acids with different carboxyl numbers and acidity remain unclear. This study investigates how carboxyl number and acidity of acetic acid (AA), tartaric acid (TA), and citric acid (CA) affect structural reorganization and low glycemic properties of cassava pyrodextrins. Compared with AA, TA, and CA with stronger acidity and more carboxyl groups promoted more extensive hydrolysis, transglycosylation, repolymerization, and esterification. These changes increased indigestible glycosidic linkages and the branching degree, while reducing molecular weight. Molecular docking confirmed stronger hydrogen-bonding interactions between TA/CA and starch chains. Furthermore, TA- and CA-catalyzed pyrodextrins exhibited superior anti-digestive properties with resistant starch up to 54.26% and an estimated glycemic index as low as 42.46, highlighting the critical role of carboxyl numbers and acidities in modulating the functionality of pyrodextrins.

Manihot

Facilitators and Barriers to Volunteers' Involvement in Palliative Care: A Qualitative Meta-Synthesis.

OBJECTIVE: This study aims to systematically synthesize qualitative evidence on facilitators and barriers to volunteer involvement in palliative care services, providing insights to inform strategies for strengthening volunteer support systems. METHODS: PubMed, Web of Science, Embase, Cochrane Library, Medline, EBSCO, ProQuest, China National Knowledge Infrastructure, Wanfang, VIP, and Sinomed were searched from inception to December 2025 to identify qualitative studies examining factors influencing volunteer participation in palliative care. Methodological quality was assessed using the Joanna Briggs Institute Critical Appraisal Checklist for Qualitative Research. Data were analyzed using Thomas and Harden's thematic synthesis approach and managed using NVivo 12.0 software, following the Enhancing Transparency in Reporting the Synthesis of Qualitative Research (ENTREQ) guidelines. RESULTS: Thirty-one studies involving 1042 participants were included, yielding 68 findings. Facilitators included intrinsic motivation and meaning-making at the individual level; supportive relationships and teamwork at the interpersonal level; structured support and professional recognition at the organizational level; social recognition and resource integration at the community level; and institutional safeguards and governmental incentives at the policy level. Barriers included emotional burden and limited competencies at the individual level; relationship conflicts and insufficient collaboration at the interpersonal level; management deficiencies at the organizational level; community resource imbalances at the community level; and inadequate regulations and incentives at the policy level. CONCLUSION: Volunteer participation in palliative care is influenced by multiple interacting factors. Strengthening training and support systems, enhancing team collaboration, and improving institutional frameworks may help sustain volunteer engagement and improve the quality of palliative care services.

Palliative Care

Synergistic HMGN1 and VP64 Fusions Potentiate High-Precision and PAM-Flexible Base Editing.

RNA-guided CRISPR-derived base editors (BEs) have revolutionized genome editing by enabling targeted base substitutions. However, their application is frequently constrained by the stringent requirement for PAM sequences and low editing precision (bystander editing). Here, we present a robust strategy to overcome these limitations by coupling SpRY, a near-PAM-less Cas9 variant, with truncated CDA1 cytidine deaminases. While this combination enables precise editing of virtually any cytosine in the genome, it initially exhibited suboptimal efficiency. To address this, we systematically screened a diverse panel of candidate DNA-binding proteins and identified that the synergistic fusion of HMGN1 and VP64 substantially enhances editing activity without compromising precision. Importantly, this enhanced editing efficiency was achieved without markedly increasing off-target effects. Our new BEs demonstrated robust performance not only in yeast but also in rice, suggesting broad applicability in gene therapy, precision breeding, and fundamental research.

Gene Editing

Maraviroc alleviates neuropathic pain symptoms in a mouse model of spared nerve injury.

Chronic pain represents a major health problem in the health care system. According to the CDC data brief in 2020, 20.4% of adults have chronic pain. There has been no promising therapy for chronic pain. Currently available treatments include medications such as nonsteroidal anti-inflammatory drugs, antiepileptic drugs, tricyclic antidepressants, corticosteroids, opioids, and cannabinoids, all of which may cause various negative side effects. Thus, there is an urgent need to develop novel, efficacious, and safe interventions for treating pain. Studies have shown that proinflammatory cytokines and chemokines make important contributions to the initiation and persistence of pain. We have found that C-C motif chemokine ligand 5 levels increased at day 14 post-spared nerve injury (SNI). This study was designed to investigate the effect of maraviroc (MVC), an FDA-approved CCR5 antagonist, on neuropathic pain in a mouse model of SNI. We found that MVC alleviated SNI-induced mechanical allodynia at 3, 7, and 14 days postinjury. MVC treatment also prevented SNI-mediated thermal hypersensitivity at 7 and 14 days postinjury in both male and female cohorts. SNI resulted in weight-bearing deficits, which were corrected by MVC administration in male mice. RNA sequencing analysis revealed that MVC rescued SNI-induced dysregulation of sex-specific canonical pathways in the spinal cord. Collectively, our findings showed that MVC could reduce neuropathic pain following peripheral nerve injury, providing a base for the repurposing of this FDA-approved human immunodeficiency virus drug as a pain reducer in clinical applications. SIGNIFICANCE STATEMENT: Spared nerve injury-induced neuropathic pain is associated with upregulation of the C-C motif chemokine ligand 5. Targeting the C-C motif chemokine ligand 5-CCR5 axis with FDA-approved maraviroc alleviated pain phenotype through modulating different pathways in male and female mice.

Animals

Genetic predisposition and mediating pathways in ischemic stroke-induced cardiac arrhythmias: a genome-wide analysis.

INTRODUCTION: The clinical presentation of stroke-heart syndrome (SHS) underscores the interplay between the central nervous system and the cardiovascular system. While cardiac arrhythmia is the prevalent form of cardiac injury in SHS patients, the causal link between ischemic stroke and cardiac arrhythmia is still unclear. METHODS: Mendelian randomization analyses and genome-wide association studies data were used to investigate the causal role of ischemic stroke on cardiac complications. Mediation and colocalization analyses were used to identify potential pathways and shared genetic variants. Single nucleotide polymorphisms (SNPs) associated with arrhythmias and ischemic stroke were used for Gene Ontology and Kyoto Encyclopedia of Genes and Genomes (KEGG) analyses. Gene expression omnibus (GEO) database from atrial fibrillation patients were used for validation. RESULTS: Mendelian randomization analyses showed a strong correlation between arrhythmias, including ventricular tachyarrhythmias and atrial fibrillation, with ischemic stroke. Diabetic microvascular (nephropathy, retinopathy) and macrovascular (cardiomyopathy, peripheral arterial disease) complications significantly mediated the effect of ischemic stroke on cardiac arrhythmias and atrial fibrillation, explaining 28.69&#xa0;% and 20.48&#xa0;% of the indirect effect, respectively. Colocalization analyses identified a shared causal variant in the Phosphodiesterase 3A (PDE3A) gene (rs11045239), providing genetic evidence for a shared pathogenic pathway between ischemic stroke and cardiac arrhythmias. Moreover, KEGG pathway enrichment analyses identified a role of the cyclic adenosine monophosphate (cAMP) signaling pathway in both ischemic stroke and arrhythmias. Validation using the GEO database confirmed a significant upregulation of the PDE3A gene expression in atrial fibrillation patients. CONCLUSION: This study demonstrated a causal link between ischemic stroke and cardiac arrhythmias, with diabetic complications as one mediating factor. The identification of a shared causal variant in the PDE3A gene and the role of the cAMP signaling pathway have the potential to improve prediction and management of SHS patients.

Humans

Population pharmacokinetics and dosing optimization of cefoselis in paediatric patients with haematological malignancies.

BACKGROUND: Cefoselis is a fourth-generation cephalosporin primarily indicated for infections caused by susceptible bacteria. The pharmacokinetic (PK) characteristics, efficacy and safety of cefoselis in paediatric patients with haematological malignancies remain unclear, posing a risk of suboptimal exposure and associated therapeutic failure or toxicity. Therefore, we studied cefoselis pharmacokinetics (PK) to optimize dosing in paediatric patients with haematological malignancies. METHODS: Blood samples were collected from paediatric patients with haematological malignancies. A population PK (PopPK) analysis was performed using NONMEM (v7.4). Monte Carlo simulations were used to evaluate current dosing regimens by calculating the PTA. Pharmacodynamic target was defined as unbound plasma concentrations above the MIC throughout the entire dosing interval. Clinical efficacy and safety data were collected. RESULTS: A total of 96 samples from 53 patients were collected. A two-compartment model with zero-order input and first-order elimination best described the PK of cefoselis after IV administration. Weight was the only covariate that affected PK. Monte Carlo simulations showed that the PTA was more than 96.7% for susceptible pathogens (MIC&#x200a;=&#x200a;0.25&#x2005;mg/L) at 40&#x2005;mg/kg, and less than 30.5% for Pseudomonas aeruginosa (MIC&#x200a;=&#x200a;32&#x2005;mg/L) at 80&#x2005;mg/kg. A total of 39 patients had body temperatures below 37.3&#xb0;C after 3&#x202f;&#xb1;&#x202f;1&#x2005;days of cefoselis treatment (with a median baseline temperature of 38.5&#xb0;C). There were no adverse events leading to discontinuation. CONCLUSIONS: A PopPK model of cefoselis in paediatric patients with haematological malignancies was established and the dosing regimens were evaluated.

Humans

Genome-wide gene-sleep interaction study identifies novel lipid loci in 732,564 participants.

BACKGROUND AND AIMS: Deviations from the population mean in sleep duration have been associated with increased risk for developing dyslipidemia and atherosclerotic cardiovascular disease, but the mechanism of effect is poorly characterized. We performed large-scale genome-wide gene-sleep interaction analyses of lipid levels to identify genetic variants underpinning the biomolecular pathways of sleep-associated lipid disturbances and to suggest possible druggable targets. METHODS: We collected data from 55 cohorts with a combined sample size of 732,564 participants (87&#xa0;% European ancestry) with data on lipid traits (high-density lipoprotein [HDL-c] and low-density lipoprotein [LDL-c] cholesterol and triglycerides [TG]). Short (STST) and long (LTST) total sleep time were defined by the extreme 20&#xa0;% of the age- and sex-standardized values within each cohort. Based on cohort-level summary statistics data, we performed meta-analyses for one-degree of freedom tests of interaction and two-degree of freedom joint tests of the SNP-main and -interaction effect on lipid levels. RESULTS: The one-degree of freedom variant-sleep interaction test identified 10 novel loci (Pint<5.0e-9), and we additionally identify 7 loci within the two-degree of freedom analyses (Pjoint<5.0e-9 in combination with Pint<6.6e-6). Multiple loci, including those mapped to APSH (target for aspartic and succinic acid) and SLC8A1 showed biological plausibility and druggability potential based on literature. CONCLUSIONS: Collectively, the 17 (9 with short and 8 with long sleep) loci provided evidence into the biomolecular mechanisms underlying sleep-associated lipid changes, including potential involvement of the vitamin D receptor pathway. Collectively, these findings may contribute developing novel interventions for treating dyslipidemia in people with sleep disturbances.

Humans

SMART-RNA-Metavirome: a practical RNA metavirome platform compatible with high-throughput sequencing of both short and long reads.

BACKGROUND: The RNA virosphere's extensive diversity and its role in emerging infectious diseases underscore the importance of non-targeted sequencing for identifying unknown or rare pathogens, including co-infections. However, enriching low-abundance viral sequences in RNA metaviromics, particularly in&#xa0;the preparation of cDNA libraries and their compatibility with next-generation sequencing (NGS) and third-generation sequencing (TGS), remains challenging. Therefore, our objective is to develop and systematically assess a practical RNA metavirome methodology specifically tailored for the enrichment of low-abundance viral sequences within samples. METHODS: We developed the SMART-RNA-Metavirome platform, integrating SMART-9n library preparation with NGS and TGS technologies. Total RNA was extracted from two field-collected wild Aedes albopictus pools, along with one laboratory-infected Ae. albopictus pool harboring dengue virus (DENV). This RNA was subjected to reverse transcription using both this optimized protocol and random primer-based methods, followed by high-throughput sequencing on Illumina, Oxford Nanopore, and QitanTech Nanopore technologies. Welch's t-test was employed for comparative analysis of the subsequent RNA metavirome data, specifically to evaluate differences in viral species composition and abundance of viral reads between experimental groups. Furthermore, the effectiveness of this platform was systematically validated via RT-qPCR and SMART-RNA-Metavirome-based Oxford&#xa0;Nanopore sequencing across multiple sample types, including mosquito specimens from DENV-infected Ae. albopictus, serum samples from dengue patients and viral isolates of Japanese encephalitis virus (JEV) and Zika virus (ZIKV). RESULTS: The SMART-RNA-Metavirome platform has been systematically validated to excel in enriching the composition and diversity of the RNA virome (P&#x2009;=&#x2009;0.04), providing sufficient coverage for the complete reconstruction of viral genomes. When employed in the detection of DENV-infected Ae. albopictus, clinical serum samples, and viral isolates of JEV and ZIKV, this technique exhibits a robust correlation with RT-qPCR (r2&#x2009;>&#x2009;0.95). Notably, it demonstrates exceptional sensitivity, ensuring sufficient coverage even in samples of DENV-infected Ae. albopictus with a Ct-value of 35.3, attaining an impressive 99.88% genome coverage. Furthermore, this platform possesses the capability to identify virus species and determine their serotypes. CONCLUSIONS: In our study, the SMART-RNA-Metavirome platform outperforms traditional methods, enriching RNA virome composition and diversity, enabling practical compatibility with both NGS and TGS technologies. It demonstrates significant proficiency in detecting both known and unknown arboviruses, even in low-titer samples such as those from wild mosquitoes and clinical sera. This platform facilitates comprehensive monitoring, risk assessment, and early warning of RNA virus transmissions, enhancing our understanding of RNA virome diversity and ecological patterns.

High-Throughput Nucleotide Sequencing

Panorama of Chromosomal Instability in Lung Cancer.

Lung cancer is a highly heterogeneous disease primarily driven by tobacco smoking. About 20% of lung cancers occur among patients who have never smoked (LCINS) with differences in patient ancestry, sex, tumor histology, and clinical features. Our understanding of chromosomal instability in lung cancer, especially LCINS, is still limited. Here, we perform a comprehensive study of 182,429 somatic structural variations (SVs) detected in 1,209 whole-genome sequenced lung cancers, of which 864 LCINS. SVs are more abundant in tumors from patients who have smoked (LCSS); however, they are more complex and play more important roles in tumorigenesis in LCINS. EGFR mutations and KRAS mutations profoundly and independently shape the SV landscape. EGFR-mutant tumors have higher SV burden and more cancer-driving SVs. In contrast, KRAS mutations are associated with lower SV burden and less driver SVs. We decompose 16 SV signatures for both complex and simple SVs that likely represent divergent molecular mechanisms. The SV breakpoints have distinct distributions across the genome depending on the signatures due to mutagenic mechanisms and positive selection. Many established cancer-driving genes are recurrently rearranged by multiple SV signatures suggesting functional convergence of these genome instability mechanisms.

Journal Article

Acetyl-CoA synthetase mutations affect the susceptibility of Plasmodium falciparum to antimalarial drugs.

Plasmodium falciparum acetyl-CoA synthetase (PfAcAS) is an important source of acetyl-CoA. We detected mutations S868G and V950I in PfAcAS by whole-genome sequencing analysis in certain recrudescent parasites after treatment with artesunate and dihydroartemisinin-piperaquine. Using CRISPR/Cas9 technology, we engineered parasite lines to carry the PfAcAS S868G and V950I mutations in two genetic backgrounds and evaluated their susceptibilities to antimalarial drugs in vitro. The results demonstrated that PfAcAS S868G and V950I mutations alone or in combination affected the susceptibility of P. falciparum to several antimalarial drugs, including the artemisinin derivatives (dihydroartemisinin, artesunate, and artemether) and chloroquine, although absolute changes in susceptibilities were modest.IMPORTANCEMalaria, an infectious disease caused by Plasmodium parasites and transmitted by mosquitoes, continues to be one of the most pressing public health challenges worldwide. P. falciparum has demonstrated reduced sensitivity to artemisinin-based combination therapies (ACTs), thereby intensifying the difficulties associated with malaria management. Currently, only a limited number of molecular markers exist for identifying drug resistance in P. falciparum, and these markers do not fully elucidate the mechanisms behind this resistance. In this study, we performed whole-genome sequencing analysis on P. falciparum strains that reemerged following ACT treatment. We aim to identify molecules potentially associated with drug resistance, which may provide new molecular markers for monitoring drug resistance in P. falciparum.

Plasmodium falciparum

Comparing artificial and convolutional neural networks with traditional models for Genomic prediction in wheat.

With the rapid development of sequencing technology, the application of genomic prediction has become more and more common in breeding schemes of livestocks and crops. Selecting an appropriate statistical model is of central importance to achieve high prediction accuracy. Recently, machine learning models have been expected to upgrade genomic prediction into a new era. However, the perspective still suffers from lack of evidence that machine learning models can generally outperform the traditional ones on empirical data sets. In this study, we compared two machine learning models based on artificial neural network (ANN) and convolutional neural network (CNN) with four traditional models, including genomic best linear unbiased prediction (GBLUP), Bayesian ridge regression (BRR), BayesA and BayesB, using three published data sets for grain yield in wheat. For each model, we considered two variants: modeling and ignoring the genotype-by-environment ([Formula: see text]) interaction. In the comparison, we considered two strategies of cross-validation: predicting genotypes that have not been evaluated in any environment (CV1) and predicting genotypes that have been tested in other environments (CV2). Our results showed that traditional Bayesian models (BayesA, BayesB, and BRR) outperformed GBLUP, ANN and CNN when considering [Formula: see text] interaction. The accuracies of ANN and CNN were higher than traditional models only in CV1 and when [Formula: see text] interaction was ignored. It was also found that the performance of the two machine learning models was significantly affected by the interaction between the CV strategy and the way of treating the [Formula: see text] interaction, while that of the four traditional models was only influenced by whether the [Formula: see text] interaction was considered or not. Thus, machine learning models can be a powerful complementary to the traditional ones and their superiority may depend on the prediction scenario. Among the two machine learning models, we observed that the accuracy of ANN was higher than CNN in most cases, indicating that it is still challenging to adapt complex machine learning models such as CNN to genomic prediction.

ANN

Whole genome sequence analysis of low-density lipoprotein cholesterol across 246&#xa0;K individuals.

BACKGROUND: Rare genetic variation provided by whole genome sequence datasets has been relatively less explored for its contributions to human traits. Meta-analysis of sequencing data offers advantages by integrating larger sample sizes from diverse cohorts, thereby increasing the likelihood of discovering novel insights into complex traits. Furthermore, emerging methods in genome-wide rare variant association testing further improve power and interpretability. RESULTS: Here, we conduct the largest meta-analysis of whole genome sequencing for low-density lipoprotein cholesterol (LDL-C), a therapeutic target for coronary artery disease, analyzing data from 246&#xa0;K participants and integrating 1.23B variants from the UK Biobank and the Trans-Omics for Precision Medicine (TOPMed) program. We identify numerous rare coding and non-coding gene associations related to LDL-C, with replication across 86&#xa0;K participants in All of Us. Our findings are based on single-variant analyses, rare coding and non-coding variant aggregation tests, and sliding window approaches. Through this comprehensive analysis, we identify 704 novel single-variant associations, 25 novel rare coding variant aggregates, 28 novel rare non-coding variant aggregates, and one novel sliding window aggregate. CONCLUSIONS: This study provides a meta-analysis framework for large-scale whole genome sequence association analyses from diverse population groups, yielding novel rare non-coding variant associations.

Humans

A prognostic signature for lung adenocarcinoma in people who have never smoked.

Knowledge of tumor cell dynamics can inform prognosis and treatment yet is largely lacking for lung adenocarcinoma in people who have never smoked (NS-LUAD). With RNA-seq data from 684 NS-LUAD and validation in an independent dataset, we identified three subtypes with distinct phenotypic traits and cell compositions. Additional genomic and histological data further characterized the subtypes. 'Steady', marked by low proliferation, high alveolar cell fraction, moderate-to-well differentiation, and fewer driver genes' alterations, is linked to prolonged survival and low immune evasion. 'Proliferative' shows high proliferation markers, TP53 mutations, and gene fusions. 'Chaotic', with high epithelial-to-mesenchymal transition markers, has the worst prognosis even within stage I tumors. Lacking known molecular or histological characteristics, this aggressive subtype is solely identified by transcriptomic data. A 60-gene signature recapitulates the overall classification and strongly predicts survival even within subgroups based on tumor stage or known genomic features, emphasizing its potential for improving NS-LUAD prognostication in clinical settings.

Journal Article

Epigenetic mechanisms underlying variation of IL-6, a well-established inflammation biomarker and risk factor for cardiovascular disease.

BACKGROUND AND AIMS: Cardiovascular disease (CVD) is one of the leading causes of morbidity and mortality worldwide, yet the underlying molecular mechanisms remain less understood. Chronic low-grade inflammation is a complex immune response contributing to the pathophysiology of cardiovascular disease. This response is signaled in part by interleukin-6 (IL-6), a pleiotropic, pro-inflammatory cytokine. Phenotypic variance in circulating IL-6 level may be explained in part by DNA methylation which is increasingly being associated with cardiovascular effects. METHODS: In this study we evaluated methylated DNA (CpG sites) associated with blood IL-6 levels across &#x223c;4,400 ancestrally diverse individuals (81&#xa0;% self-reported White; 9&#xa0;% Black or African American, 8&#xa0;% Hispanic or Latino/a, and 2&#xa0;% Chinese American). RESULTS: We identified 178 CpG sites associated with IL-6 (p<0.05/&#x223c;395,000). Among the sites, cg04437762 is located within the transcription unit of IL6R, a current therapeutic target for inflammatory disease, and cg26692003 and cg00464927 were significant for IL6 and IL6ST trans-CpG-gene transcripts. Functional gene expression downstream of methylation identified cellular response to IL-6 and B-cell regulation and activation pathways. Four genes were linked with both a genetic component of cardiovascular disease and an IL-6 associated CpG site. Three CpG sites identified through Mendelian randomization analyses supported inference of a causal effect on IL-6 levels, including the LYN gene that regulates immune cell signaling and has been previously associated with atherosclerosis. CONCLUSIONS: Overall, we identified several novel IL-6-CpG sites and downstream pathways affected by methylation. Follow-up functional studies including the regulation of IL-6 would complement current knowledge of CVD pathophysiology and potential therapeutic targets.

Humans