PubMed HealthSearch

PubMed · 41601190

X-intNMF: a cross- and intra-omics regularized NMF framework for multi-omics integration.

Abstract

MOTIVATION: The rapid accumulation of multi-omics data presents a valuable opportunity to advance our understanding of complex diseases and biological systems, driving the development of integrative computational methods. However, the complexity of biological processes, spanning multiple molecular layers and involving intricate regulatory interactions, requires models that can capture both intra- and cross-omics relationships. Most existing integration methods primarily focus on sample-level similarities or intra-omics feature interactions, often neglecting the interactions across different omics layers. This limitation can result in the loss of critical biological information and suboptimal performance. To address this gap, we propose X-intNMF, a network-regularized non-negative matrix factorization (NMF) framework that simultaneously integrates intra- and cross-omics feature interactions into a shared low-dimensional representation (see Fig. 1). By modeling these multi-layered relationships, X-intNMF enhances the representation of biological interactions and improves integration quality and prediction accuracy. RESULTS: For evaluation, we applied X-intNMF to predict breast cancer phenotypes and classify clinical outcomes in lung and ovarian cancers using mRNA expression, microRNA expression, and DNA methylation data from TCGA. The results show that X-intNMF consistently outperforms state-of-the-art methods. Ablation studies confirm that incorporating both cross-omics and intra-omics interactions contributes significantly to the model's improved performance. Additionally, survival analysis on 25 TCGA cancer datasets demonstrates that the integrated multi-omics representation offers strong prognostic value for both overall survival and disease-free status. These findings highlight X-intNMF's ability to effectively model multi-layered molecular interactions while maintaining interpretability, robustness, and scalability within the NMF framework. AVAILABILITY AND IMPLEMENTATION: The source code and datasets supporting this study are publicly available at GitHub (https://github.com/compbiolabucf/X-intNMF) and archived on Zenodo (https://doi.org/10.5281/zenodo.18238385).

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Tien-Thanh Bui, Rui Xie, Wei Zhang. 2026-02-03. X-intNMF: a cross- and intra-omics regularized NMF framework for multi-omics integration.. https://doi.org/10.1093/bioinformatics%2Fbtag046

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Advancing the Deciphering of Host-Microbe Crosstalk with Spatial Omics: A Mini-Review.

Host-microbe crosstalk refers to the reciprocal influences between a host and its resident or invading microorganisms. This crosstalk plays important roles in maintaining host health, regulating physiological functions, and coordinating responses to infection. The rapid rise of spatial omics is transforming how this crosstalk is studied in both animals and plants. Unlike traditional bulk omics, which homogenize tissues and erase spatial context, spatial methods preserve in situ organization and can simultaneously capture molecular information from hosts and microbes. As a result, researchers can characterize the spatial organization of colonization and infection, identify spatial associations between microbial niches and host cell states, and visualize local host response gradients across intact tissues. Current spatial omics technologies encompass sequencing-based, imaging-based, and hybrid platforms. Spatial multi-omics approaches enable the joint measurement or integration of gene expression, protein abundance, and metabolite distributions. Although spatial association alone does not establish causality, spatial omics provides a high-resolution framework for characterizing host-microbe relationships within intact tissues and generating spatially constrained, testable hypotheses. When combined with perturbation experiments and complementary experimental evidence, these hypotheses can contribute to mechanistic interpretation of host-microbe crosstalk. Here, we review spatial omics technologies, compare their suitability and major trade-offs for host-microbe studies, and discuss computational strategies, analytical challenges, and future prospects.

Multiomics

Integrating multi-omics technologies to decipher microbiome functions.

Multi-omics approaches have revolutionized our understanding of microbial communities by enabling simultaneous interrogation of genomic, transcriptomic, proteomic, and metabolomic data. The systematic integration and analysis of these deep datasets help decipher the functional roles of microbiomes, providing critical insights into microbial activities, interactions, and dynamics across diverse environments. Biological complexity makes multi-omics analysis of a single, isolated organism demanding but highly informative, yet this complexity increases further when samples comprise hundreds to thousands of individual species. As microbiome research continues to expand into clinical, environmental, and engineered systems, standardized workflows, benchmarked datasets, and community-driven initiatives are essential to ensure reproducibility, standardization and interpretability. Establishing and disseminating best practices for experimental design, data processing, and integrative analyses will be critical for maximizing comparability and scientific rigor across studies. This perspective highlights recent advances in multi-omics microbiome research, outlines key obstacles in data integration and metadata harmonization, and proposes a collaborative roadmap for scalable, FAIR-compliant multi-omics investigations and potentially disruptive Artificial Intelligence (AI) advances comparable to those of AlphaFold in the field of microbiome science.

Multiomics

Scalable, generalizable and uncertainty-aware integration of spatial multiomics across diverse modalities and platforms with SCIGMA.

Recent advances in spatial omics technologies have enabled simultaneous profiling of transcriptomic, proteomic, epigenomic, metabolomic and imaging data at high spatial resolution, offering unprecedented opportunities to dissect tissue complexity. However, integrating these diverse and large-scale spatial multimodal datasets remains a major computational challenge. We present SCIGMA, a scalable and generalizable deep learning framework for spatial multiomics integration. SCIGMA introduces an uncertainty-aware contrastive learning objective and multiview graph neural networks to preserve modality-specific signals while learning biologically meaningful joint representations. Unlike previous methods, SCIGMA provides spatially resolved uncertainty estimates, interpretably identifying regions of biological or technical heterogeneity. SCIGMA supports integration of up to five modalities, and its modular framework is extensible to future technologies with even more modalities. It also scales to more than 1 million spatial locations, enabling analysis of high-resolution datasets such as Visium HD and Xenium Prime. We evaluated SCIGMA across 19 datasets spanning 8 modalities, 10 tissues and 9 platforms. On benchmarkable datasets, SCIGMA outperformed other methods in spatial domain detection, modality preservation, feature reconstruction and reproducibility. SCIGMA identifies biologically meaningful structures, refined spatial domains and modality-specific regulatory programs, providing a robust, flexible and future-ready solution for scalable spatial multimodal integration.

Multiomics