PubMed Health⌕ Search

SEARCH · PubMed Health

Results for “functional annotations”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7Linked to original sources

Complex Genetics and Regulatory Drivers of Hypermobile Ehlers-Danlos Syndrome: Insights from Genome-Wide Association Study Meta-analysis.

BACKGROUND: Hypermobile Ehlers-Danlos syndrome (hEDS) is the most common subtype of EDS, a group of heritable connective tissue disorders. Clinically, hEDS is defined by generalized joint hypermobility and chronic musculoskeletal pain, but its impact extends beyond the musculoskeletal system. Affected individuals frequently experience autonomic, gastrointestinal, immune, and neuropsychiatric involvement, highlighting both the multisystemic nature of the condition and challenges of diagnosis. In contrast to other EDS subtypes with defined genetic causes, the molecular basis of hEDS has remained elusive. METHODS: We conducted a genome-wide association study (GWAS) of hEDS across three case controls studies, including 1,815 cases and 5,008 ancestry-matched controls. Fixed-effects meta-analysis of 6.2 million variants was complemented with LDAK gene-based association testing, transcriptome-wide association studies, and integrative annotation across multiple tissues and cell types including eQTLs, enhancer marks and open chromatin accessibility profiles, supported by luciferase assays on one candidate variant. LD-score genetic correlations were assessed between hEDS and 19 frequently reported comorbid conditions. RESULTS: Two loci reached genome-wide significance, including a regulatory region near the atypical chemokine receptor 3 gene (ACKR3) on chromosome 2. Functional annotation supports ACKR3 risk alleles colocalize with eQTLs in tibial nerve, alter enhancer activity, and generate a de novo AHR transcription factor regulatory site, implicating neuroimmune and pain signaling pathways. Gene-based and transcriptome-wide analyses identified common variants in a locus containing multiple candidates, including SLC39A13, a zinc transporter critical for connective tissue development previously implicated in a rare form of EDS, and PSMC3, a gene involved in central nervous system development. LD-score regression revealed significant genetic correlations between hEDS and joint hypermobility, myalgic encephalomyelitis/chronic fatigue syndrome, fibromyalgia, depression, anxiety, autism spectrum disorder, migraine, and gastrointestinal diseases. CONCLUSIONS: These results establish the first evidence of common variant contributions to hEDS, supporting a complex, multisystem model involving neuroimmune-stromal dysregulation. Our findings add novel indications to hEDS pathogenesis and provide solid foundations for future molecular definition and therapeutic discovery.

Genome-wide association study↗

Metagenomic Analysis of Gut Microbiome of Persistent Pulmonary Hypertension of the Newborn.

Persistent pulmonary hypertension of the newborn (PPHN) is one of the most common diseases in the neonatal intensive care unit which severely affects neonatal survival. Gut microbes play an increasingly important role in human health, but there are rarely reported how gut microbiota contribute to PPHN. In our study, the metagenomic sequencing of feces from 12 PPHN's neonates and 8 controls were performed to expose the relation between neonatal gut microbes and PPHN disease. Firstly, we found that the abundance of Actinobacteria, Proteobacteria, Bacteroidetes were significantly increased in PPHN compared with controls, but the Firmicutes components was reduced. And some pathogenic strains (like Vibrio metschnikovii) were significantly enriched in the PPHN compared with controls. Secondly, functional annotation of genes found that PPHN up-regulated transmembrane transport, but down-regulated ribosome and ATP binding. Lastly, microbial metabolic pathway enrichment analysis indicated that some metabolic pathway in PPHN were conflicting and contradictory, showed that an abnormally increased metabolism, disturbed protein synthesis and genomic instability in the PPHN neonate. Our results contribute to understanding the changes in the species and function of gut microbiota in PPHN, thus providing a theoretical basis for the explanation and treatment of PPHN.

Gastrointestinal Microbiome↗

Mutation accumulation in a hybrid parthenogenetic vertebrate.

Asexual lineages are thought to experience elevated extinction rates compared with sexual species, yet direct evidence for the underlying genetic causes remains scarce. Muller's ratchet predicts that the absence of recombination in asexual organisms facilitates the accumulation of deleterious mutations, thereby reducing long-term fitness. Here, we test this hypothesis in the hybrid-origin, parthenogenetic whiptail lizard Aspidoscelis tesselatus by integrating short-read RNAseq and long-read IsoSeq data from both the asexual lineage and its parental sexual species. We reconstructed phased transcripts for A. tesselatus to quantify mutation accumulation relative to the parental sexual species. Comparative analyses revealed elevated ω ratios in both parental genomic complements (subgenomes) of the parthenogenetic lineage, consistent with accelerated accumulation of nonsynonymous mutations. Structural variant analyses identified multiple indels in expressed transcripts predicted to disrupt protein domains. Functional annotation indicated that genes affected by both single-nucleotide variants and indels were enriched for roles in chromatin organization, apoptosis regulation, and transcriptional control. While both parental subgenomes showed similar evolutionary patterns, the maternal complement exhibited more structural and missense mutations than the paternal complement. Together, these results provide evidence that mutations accumulate in asexual A. tesselatus in genes involved in core cellular functions, supporting theoretical predictions that Muller's ratchet contributes to mutation accumulation in asexual lineages.

Animals↗

Genome-resolved analysis reveals disruption of gut microbial vitamin B and K2 biosynthesis during Toxoplasma gondii infection in mice.

UNLABELLED: Toxoplasma gondii infection remodels the gut microbiome, yet its impact on microbial vitamin biosynthetic potential and host redox metabolism remains unclear. Here, we integrated mouse gut metagenomes with publicly available metagenome-assembled genomes (MAGs) to construct a genome-resolved atlas of B-vitamin and vitamin K2 biosynthesis. From 45,697 MAGs, we curated 4,771 representative genomes, of which 2,682 met high-quality criteria (completeness &#x2265;90%, contamination <5%). Functional annotation identified 229,717 vitamin-related genes corresponding to 177 Kyoto Encyclopedia of Genes and Genomes (KEGG) orthologs across de novo pathways for eight B vitamins, thiamine (B1), riboflavin (B2), niacin (B3), pantothenate (B5), pyridoxine (B6), biotin (B7), folate (B9), cobalamin (B12), and vitamin K2. Among the high-quality genomes, 1,665 encoded complete de novo pathways for at least one vitamin, highlighting functional specialization and community-level complementarity. Transcripts per million-normalized metagenomic read counts revealed significant differences in KEGG ortholog abundances across six of the nine vitamin pathways. Reanalysis of metagenomic data from infected mice (acute, chronic, and control; n = 10 per group) revealed a stage-dependent reduction in &#x3b1;-diversity of vitamin biosynthesis pathways during acute infection, and a clear &#x3b2;-diversity separation from chronic and control groups. Core niacin biosynthesis genes (nadB, nadA, nadC) displayed phylum-specific redistribution, indicating selective remodeling of microbial NAD+ precursor production under infection-induced metabolic stress. These results suggest that T. gondii infection disrupts cooperative vitamin biosynthetic networks while specifically modulating niacin pathways linked to host NAD+ metabolism. IMPORTANCE: Gut microbes can synthesize essential vitamins, but how infection alters this function is poorly understood. By integrating mouse gut metagenomes with genome-resolved microbial data, we show that Toxoplasma gondii infection reshapes the vitamin biosynthetic potential of the gut microbiome in a stage-dependent manner. Acute infection reduces the diversity of vitamin biosynthesis pathways and shifts the taxonomic distribution of key niacin biosynthesis genes involved in microbial NAD+ precursor production. These findings identify vitamin metabolism, especially niacin-related pathways, as a sensitive functional axis of microbiome remodeling during infection. Our work links microbial taxonomic changes to functional metabolic consequences and suggests that microbiome-mediated regulation of NAD+-related metabolism may contribute to host redox adaptation during T. gondii infection.

B vitamins↗

Genomes of the ex-type strains of Elsino&#xeb; mangiferae and E. perseae, the causal agents of scab on mango and avocado.

Elsino&#xeb; species are slow-growing, hemibiotrophic to necrotrophic fungi that cause scab diseases on economically important fruit crops. Genome resources for many host-specific species remain limited. We report high-quality draft genome assemblies for the ex-type strains of Elsino&#xeb; mangiferae (CBS 226.50) and E. perseae (CBS 406.34), causal agents of mango and avocado scab, respectively. Among 5 approaches tested, a Nanopore-only NextDenovo assembly produced the most contiguous genomes, yielding 24.5 Mb (E. mangiferae) and 25.1 Mb (E. perseae) assemblies with 13 and 18 contigs, respectively, BUSCO completeness scores of &#x223c;94%, and multiple putative telomere-to-telomere chromosomes. Gene prediction identified 9,134 and 9,243 genes, respectively. Functional annotation revealed enrichment of metabolic and regulatory pathways, including those involved in posttranslational modification, protein transport, and secondary metabolism. Carbohydrate-active enzyme repertoires were small but conserved, consistent with stealth pathogenicity strategies and low plant cell wall degradation. Both genomes encoded large secretomes (>850 proteins), diverse protease repertoires (>300 proteins), Ecp2-like effector proteins, and multiple biosynthetic gene clusters, including clusters with similarity to those associated with elsinochrome and ACT-toxin II biosynthesis, some of which may contribute to host-pathogen interactions and disease development. A large fraction of genes lacked functional characterization, suggesting incomplete databases and/or the presence of lineage-specific genes potentially involved in virulence or host adaptation. These genome resources fill critical gaps for underrepresented Elsino&#xeb; species and provide taxonomically anchored references essential for diagnostics, comparative genomics, and research into the molecular basis of host specificity and pathogenicity in scab-causing fungi.

Persea↗

B cell pathways implicate shared genetic architecture between schizophrenia and immune-mediated diseases.

BACKGROUND: Schizophrenia and immune-mediated diseases are globally prevalent and highly heritable conditions that frequently co-occur, posing major public health burdens. However, their shared genetic architecture remains poorly understood. METHODS: We applied the bivariate causal mixture model (MiXeR) to investigate the polygenic overlap between schizophrenia and eight common immune-mediated diseases, using genome-wide association study summary statistics comprising 2,489 to 67,323 cases and 9,066 to 497,622 controls. Shared loci were identified through conditional/conjunctional false discovery rate (cond/conjFDR), local genetic correlation (LAVA), and colocalization analyses. Subsequently, gene mapping, functional annotation, expression-trait association, and drug-gene interaction analyses were performed to explore shared genes and enriched pathways, and genetic risk scores (GRS) from the UK Biobank were used to validate the findings. RESULTS: MiXeR estimated substantial polygenic overlap between schizophrenia and immune-mediated diseases, and conjFDR identified 133 shared loci, with eight prioritized through local genetic correlation and colocalization signals. These eight loci were mapped to 85 protein-coding genes enriched in pathways essential for B cell function. Among them, S-PrediXcan analyses identified 14 genes whose expression in brain tissues or blood was associated with both diseases. These genes also interact with immunomodulatory or antihypertensive drugs. Additionally, 11 of the 14 genes were linked to innate immunity and/or cognitive traits. Using UK Biobank data, we further confirmed that overall, shared gene, and B cell activation and receptor signaling pathway&#x2013;specific genetic risk for schizophrenia is associated with immune-mediated disease susceptibility. CONCLUSIONS: These findings underscore the shared genetic architecture of schizophrenia and immune-mediated diseases, advancing insights at the interface of psychiatric genetics and immunology.

Schizophrenia↗

Identifying fundamental gaps in functional metagenomics: a step towards unlocking microbiome research potential.

Incomplete functional annotation limits biological interpretation in microbiome studies and their translational potential. Poor annotation arises from multiple causes, with incomplete gene-protein-reaction mapping being one tractable yet under-examined contributor. We address this gap by developing a comprehensive hierarchical framework that systematically integrates gene families in UniRef, proteins in UniProt, and metabolic reactions in MetaCyc and BioCyc through UniProtKB accession, EC number, and Pfam-domain matching. Applied to a human gut metagenome dataset via HUMAnN3, our MetaCyc-based mapping recovers up to 2.3-fold more unique reaction identifiers than the default pipeline and increases reaction prevalence across samples from &#x2248;32% to 52% core reactions, addressing the data sparsity that limits statistical and machine-learning applications in microbiome research. Biological plausibility for the tested functions was supported by positive and negative controls: gut-microbial hormone-metabolism reactions previously linked to this dataset were recovered, while vertebrate-specific hormone-metabolism reactions remained correctly undetected. These gains derive from systematic database integration alone, without predictive algorithms, indicating that a tractable, mapping-related component of functional dark matter and data sparsity in microbiome studies is directly addressable. Because Pfam- and BioCyc-derived mappings trade specificity for coverage, confidence in any individual reaction assignment depends on the supporting evidence tier and source database.

Humans↗

Chromosomal-level genome assembly of minute pirate bug Orius nagaii Yasunaga, 1993 (Hemiptera: Anthocoridae).

Species of the genus Orius, diminutive predatory insects that act as natural enemies of other arthropods, are frequently employed in agricultural pest management for controlling various pests, such as thrips, mites, aphids, whiteflies, etc. However, the scarcity of high-quality genomic resources for these predators hinders our comprehension of their population evolution and predation ecology. Consequently, we assembled and annotated a chromosomal-scale genome of Orius nagaii by collating PacBio and Illumina sequencing and Hi-C genomic analysis techniques. The final genome assembly size 152.62&#x2009;Mb, with scaffold and contig N50 lengths of 11.53 and 2.39&#x2009;Mb, respectively. It is organized into 12 pairs of autosomes and a pair of XY sex chromosomes. The quality assessment of the genomic data with BUSCO revealed a completeness of 98.5% (n&#x2009;=&#x2009;1,367). Also, 11,917 protein-coding genes were discovered, with 94.28% of them having functional annotations. The high-quality genome of O. nagaii produced serves as a valuable resource for comprehending the interactions between predatory natural enemies and hosts, along with their evolutionary trajectories.

Animals↗

EucaMOD: a comprehensive multi-omics database for functional genomics research and molecular breeding of fast-growing eucalyptus trees.

Eucalyptus, one of the most widely planted plantation tree species globally, is primarily found in tropical and subtropical regions and contributes significantly to economic and social benefits. With advances in sequencing technologies, there is an increasing demand for the systematic analysis of multi-omics data among Eucalyptus species to enhance genetic breeding efforts. Although several early genomic databases have been established for eucalyptus, they have not been updated in a timely manner and lack recent multi-omics data, rendering them insufficient for current research needs. To address this gap, we developed the eucalyptus multi-omics database (EucaMOD, http://eucalyptusggd.net/eucamod), a comprehensive resource for cross-omics studies. In this study, we functionally annotated 45 eucalyptus genomes and structurally annotated 15, conducting comparative genomics and pan-proteomics analyses across all genomes. Additionally, we analyzed eucalyptus transcriptome, epigenome, and variome data through standardized workflows, enabling the in-depth mining and reanalysis of multi-omics datasets. EucaMOD is the most comprehensive multi-omics database for eucalyptus to date and includes data from 45 genomes (39 species), 870 mRNA-seq samples, 17 miRNA-seq samples, 52 epigenomic datasets (histone modifications and transcription factor binding), and genetic variation data from 1219 samples. To support functional genomics and molecular breeding research, the database is organized into the following 11 modules: Home, Species, Genomics, Comparative genomics, Pan-proteomics, Transcriptomics, Epigenetics, Variomics, Tools, Download, and Help. EucaMOD also offers online analysis tools for data mining, providing free public services to aid eucalyptus gene function and genetic engineering studies.

Eucalyptus↗

Microbial and functional shifts between flare and remission in a single-center cohort of children with inflammatory bowel disease.

BACKGROUND: Gut microbial dysbiosis is central to the pathogenesis of inflammatory bowel disease (IBD). While gut microbiome differences between patients with and without IBD are well established, microbiome changes associated with disease activity and remission remain limited, particularly in paediatric populations. AIM: To examine intra-individual taxonomic and functional gut microbiome changes during transition from active flare to remission under maintenance immunosuppression in a pilot single-center Singapore cohort of children with IBD. METHODS: Paired stool samples and clinical data were collected from seven patients with paediatric IBD [5 Crohn's disease (CD), 2 ulcerative colitis; &#x2264; 18 years] during active disease/flare (visit 1; Pediatric CD Activity Index/Pediatric Ulcerative Colitis Activity Index &#x2265; 10) and subsequent clinical remission (visit 2; Pediatric CD Activity Index/Pediatric Ulcerative Colitis Activity Index < 10). Samples underwent shotgun metagenomic sequencing for high-resolution taxonomic profiling and functional annotation of Kyoto Encyclopaedia of Genes and Genomes pathways. RESULTS: Gut microbial diversity was reduced during flare compared to remission, with Actinobacteria abundance significantly higher in remission. Two distinct microbial clusters differentiated flare and remission states: The remission cluster was enriched with Bifidobacterium adolescentis, Bifidobacterium dentium, Lactobacillus gasseri, Faecalibacterium prausnitzii, while the flare state showed increased Klebsiella pneumoniae. Remission was further characterized by a downregulation of pathogenic microbes and an upregulation of beneficial microbes including a higher abundance of the butyrate producer Anaerostipes hadrus (P = 0.046). Microbial functional genes enriched in remission were predominantly associated with metabolic pathways including vitamin and cofactor biosynthesis, as well as carbohydrate, amino acid, and lipid metabolism. CONCLUSION: The transition from flare to remission in Singaporean children with IBD is characterized by functional remodeling of the gut microbiome, which may contribute to recovery processes related to intestinal barrier integrity, cellular maintenance, and tissue repair. Targeted modulation of the gut microbiome may help sustain remission in paediatric IBD.

Functional shift↗

Application of a Translational Research Platform to Unveil Efficacy Signals and Mechanisms of Resistance of FGFR Inhibitors in Multiple FGFR-Altered Solid Tumors.

PURPOSE: The predictive value of fibroblast growth factor receptor (FGFR) amplifications (amp) and the role of FGFR mutations (mut) beyond known activating variants remain unclear. We aimed to establish a translational research platform to characterize FGFR alterations (alt) and explore their potential as predictive biomarkers for FGFR-targeted agents. EXPERIMENTAL DESIGN: This ambispective study included a retrospective analysis of patients with FGFR-alt tumors treated with selective FGFR inhibitors (FGFRi) and a prospective collection of longitudinal tumor samples. Patient-derived xenografts (PDX) were generated to investigate FGFRi mechanisms of action and resistance. Molecular characterization included genomic, transcriptomic, proteomic, and functional analyses using the Functional Annotation for Cancer Treatment (FACT) assay. RESULTS: Among 36 retrospectively analyzed patients, clinical benefit from FGFRis was observed in cases with FGFR mRNA overexpression or FGFR2/11q co-amp, but no association was found with the amplification levels. In archival tumor samples, exploratory proteomic analysis showed FGFR1-4 protein expression in 78% of FGFR1/2-amp tumors detected by fluorescence in situ hybridization. RNA sequencing identified a higher prevalence of FGFR mRNA overexpression than proteomic analysis. Among patients harboring FGFR-mut, only one bladder cancer with an FGFR3-mut S249C derived benefit. FACT assay supported the functional activity of selected variants, including FGFR3 T689M, and suggested potential resistance mechanisms involving PI3K/PTEN and MAPK pathway co-alterations. A prospective FGFR-alt PDX biorepository enabled exploratory biomarker analyses, supporting the hypothesis that FGFR1-4 mRNA expression may better reflect FGFR dependency than genomic alterations alone. CONCLUSIONS: These findings highlight the complexity of FGFR-driven oncogenesis and support integrative molecular approaches to refine patient selection for FGFR-targeted therapies.

Humans↗

Chromosome-level genome assembly and annotation of the porcupine fish (Diodon hystrix).

The porcupinefish (Diodon hystrix), a coral reef teleost, is widely distributed in tropical/subtropical waters of the Pacific, Atlantic, Indian Oceans, and Mediterranean Sea. It shares easily recognizable features with pufferfish, such as body inflation and spines. Additionally, its culinary value makes D. hystrix a highly desirable species in many tropical coastal regions, with considerable market potential. However, lack of a high-quality genome hindered further studies on its reproduction, molecular biology, and genomic improvement. Here, we assembled the chromosome-scale genome using PacBio HiFi, ultra-long reads, and Hi-C. Of the 713.62&#x2009;Mb genome, 98.63% anchored to 23 chromosomes (scaffold N50: 31.52&#x2009;Mb) with 39.82% repetitive sequences. The assembled genome achieved a BUSCO completeness score of 97.7%, with 23,171 protein-coding genes predicted, 22,221 of which were functionally annotated. Phylogenetic analysis identified D. hystrix's evolutionary relationships with other species in the Tetraodontiformes. In summary, the high-quality genome of D. hystrix sheds light on valuable insights into genome size evolution, and provides a valuable resource for exploiting genomic study and breeding applications in this species.

Animals↗

Chromosome-level genome assembly and annotation of Spinibarbus caldwelli.

Spinibarbus caldwelli is an economically important freshwater species within the Cyprinidae family, abundant in the middle and lower reaches of the Yangtze River and its adjacent basins. As a promising species suitable for aquaculture in southern China, the lack of genomic resources has hampered the genetic breeding and conservation. Here, we release a chromosome-level genome assembly for S. caldwelli using PacBio HiFi long-reads, Illumina short-reads, and Hi-C sequencing data. The final genome assembly is 1.77&#x2009;Gb in size, with a contig N50 of 24.27&#x2009;Mb. Using Hi-C scaffolding, 99.14% of the contigs were successfully anchored to 50 chromosomes, resulting in a scaffold N50 of 35.29&#x2009;Mb. The final genome assembly shows a BUSCO completeness of 98.27%. The assembled genome contains 49.41% repetitive sequences and 51,505 predicted genes, 90.83% of which have been functionally annotated. This genome provides a genetic basis for S. caldwelli, facilitating the exploration of Cyprinid phylogeny, genetic improvement, and conservation efforts.

Animals↗

Transcriptomic insights into the coordinated regulation of signaling, apoptosis, immunity, and metabolism during Sinonovacula constricta larval metamorphosis.

Metamorphosis is a critical ontogenetic transition for marine bivalves, marking the shift from planktonic to benthic lifestyles, where successful transformation dictates survival. The razor clam Sinonovacula constricta is economically important; however, low larval metamorphosis rates remain a major bottleneck in seedling production. To elucidate the mechanisms governing this process, we performed a comparative transcriptome analysis of S. constricta larvae at pre- and post-metamorphosis stages using Illumina sequencing. A total of 3701 differentially expressed genes (DEGs) were identified, including 3254 up-regulated and 447 down-regulated genes. Functional annotation of the respective top 20 significantly up-regulated and down-regulated DEGs indicated their potential pivotal roles in signal transduction (e.g., up-regulated: CAV1, CHRNA2; down-regulated: APP, NOTCH1), cellular proliferation and differentiation (e.g., up-regulated: TUBA, EGF1; down-regulated: KIF23, TTC25), transcriptional and epigenetic regulation (e.g., up-regulated: NFIL3; down-regulated: OVO, HMX1), substance transport (e.g., up-regulated: LRP2, LRP1B; down-regulated: SLC51A, Slc33a1), substance metabolism (e.g., up-regulated: CPK3, CYP26A1; down-regulated: RDMT1, ADAC), immunomodulation (e.g., up-regulated: CPN2, CRISP2), and protein homeostasis (e.g., up-regulated: HSP27, NAS-27). Functional enrichment analysis further revealed that DEGs were significantly enriched in pathways related to signal transduction and developmental regulation (e.g., Ras, TNF), cell death and homeostasis (e.g., apoptosis), immune responses (e.g., Toll-like receptor), energy metabolism (e.g., lipid), cardiovascular related (e.g., Fluid shear stress), cell junction and architecture (e.g., Tight junction), and infectious disease (e.g., measles). These results suggest a synergistic interplay between signaling, apoptosis, immunity, and metabolism during S. constricta metamorphosis. This study advances our understanding of marine bivalve metamorphosis and offers candidate genes for further mechanistic studies.

Animals↗

Meta-QTL Analysis Reveals Consensus Genomic Regions and Candidate Genes for Resistance to Sudden Death Syndrome in Soybean.

Sudden death syndrome (SDS), caused by Fusarium virguliforme, is one of the most economically important diseases limiting soybean production worldwide. Although numerous quantitative trait loci (QTL) associated with SDS resistance have been reported, inconsistencies among mapping populations, marker systems, and experimental conditions have hindered the identification of robust resistance loci for soybean improvement. In this study, a comprehensive meta-analysis was conducted to integrate published QTL and identify stable consensus genomic regions associated with SDS resistance. After a systematic literature survey and data curation, 153 QTL derived from 14 linkage-mapping studies were analyzed using a custom R-based workflow, resulting in the identification of 23 consensus meta-QTL (MQTL) distributed across 17 chromosomes. Several MQTL, particularly those located on chromosomes 6, 8, 18, and 20, were supported by multiple independent studies and represented major genomic hotspots for SDS resistance. Physical localization and functional annotation of these MQTL identified 217 candidate genes, including genes predicted to be involved in plant defense, signal transduction, transcriptional regulation, and secondary metabolism. Gene Ontology enrichment analysis identified response to salicylic acid as the only biological process that remained significant after FDR correction, whereas Kyoto Encyclopedia of Genes and Genomes pathway analysis did not identify significantly enriched pathways. Independent support using five published genome-wide association studies further supported several MQTL, especially those on chromosomes 6, 18, and 20, thereby increasing confidence in these genomic regions. The identified MQTL and prioritized candidate genes provide potential genomic resources for future marker development, improvement applications, and functional validation aimed at improving soybean resistance to SDS.

Fusarium virguliforme↗

MetagenomicKG: a knowledge graph for metagenomic applications.

MOTIVATION: The sheer volume and variety of genomic content within microbial communities makes metagenomics a field rich in biomedical knowledge. To traverse these complex communities and their vast unknowns, metagenomic studies often depend on distinct reference databases, such as the Genome Taxonomy Database (GTDB), the Kyoto Encyclopedia of Genes and Genomes (KEGG), and the Bacterial and Viral Bioinformatics Resource Center (BV-BRC), for various analytical purposes. These databases are crucial for the genetic and functional annotation of microbial communities. Nevertheless, the inconsistent nomenclature or identifiers of these databases present challenges for effective integration, representation, and utilization. Knowledge graphs (KGs) offer an appropriate solution by organizing biological entities from different databases to standardized identifiers, allowing their interrelations to be captured into a cohesive network regardless of the naming conventions used in each source. The graph structure not only facilitates the unveiling of hidden patterns but also enriches our biological understanding with deeper insights. Despite KGs having shown potential in various biomedical fields, their application in metagenomics remains underexplored. RESULTS: We present MetagenomicKG, a novel knowledge graph specifically tailored for metagenomic analysis. MetagenomicKG integrates taxonomic, functional, and pathogenesis-related information on the human microbiome sourced from various databases, and further connects these with existing biomedical KGs to expand the biological network. Through various case studies involving the human microbiome, we demonstrate its utility in enabling hypothesis generation regarding the relationships between microbes and diseases, generating sample-specific graph embeddings, and providing robust pathogen prediction. CODE AVAILABILITY: The source code and technical details for constructing the MetagenomicKG and reproducing all analyses are available on GitHub at https://github.com/KoslickiLab/MetagenomicKG. The data used in this manuscript, including the pre-built files and use case input data, are archived on Zenodo with DOI: 10.5281/zenodo.17546861.

Metagenomics↗

Whole genome sequence-based association analysis of African American individuals with bipolar disorder and schizophrenia.

In studies of individuals of primarily European genetic ancestry, common and low-frequency variants and rare coding variants have been found to be associated with the risk of bipolar disorder (BD) and schizophrenia (SZ). However, less is known for individuals of other genetic ancestries or the role of rare non-coding variants in BD and SZ risk. We performed whole genome sequencing of African American individuals: 1,598 with BD, 3,295 with SZ, and 2,651 unaffected controls (InPSYght study). We increased power by incorporating 14,812 jointly called psychiatrically unscreened ancestry-matched controls from the Trans-Omics for Precision Medicine (TOPMed) Program for a total of 17,463 controls. To identify variants and sets of variants associated with BD and/or SZ, we performed single-variant tests, gene-based tests for singleton protein truncating variants, and rare and low-frequency variant annotation-based tests with conservation and universal chromatin states and sliding windows. We found suggestive evidence of BD association with single-variants on chromosome 18 and of lower BD risk associated with rare and low-frequency variants on chromosome 11 in a region with multiple BD GWAS loci, using a sliding window approach. We also found that chromatin and conservation state tests can be used to detect differential calling of variants in controls sequenced at different centers and to assess the effectiveness of sequencing metric covariate adjustments. Our findings reinforce the need for continued whole genome sequencing in additional samples of African American individuals and more comprehensive functional annotation of non-coding variants.

Journal Article↗

A high-quality chromosome-level genome assembly and annotation of the giant freshwater prawn (Macrobrachium rosenbergii).

The giant freshwater prawn, Macrobrachium rosenbergii, is native to Southeast Asia and is used in aquacultural practices worldwide. It is considered advantageous because of its rapid growth, high nutritional value, and economic benefits. As one of the three major freshwater aquaculture shrimp sources in China, a high-quality genome resource is of great significance for promoting the germplasm improvement of varieties. This study presents a high-quality chromosome-level genome assembly of M. rosenbergii that was generated by combining PacBio, MGI, and Hi-C reads. The assembled genome was 2.96&#x2009;Gb in size, with a contig N50 of 0.64&#x2009;Mb and a scaffold N50 of 55.76&#x2009;Mb, which was positioned on 59 pseudo-chromosomes. The Benchmarking Universal Single-Copy Orthologs (BUSCO) analysis for genome assembly reached 94.37%. In total, 27,111 protein-coding genes were identified, of which 25,470 were functionally annotated. These results provide a foundation for future research into adaptive evolution, genomics, and molecular breeding in M. rosenbergii.

Animals↗