PubMed HealthSearch

PubMed · 42234838

Network methods for diagonal integration of unpaired single-cell multiomics data: a review.

Abstract

MOTIVATION: Advances in single-cell sequencing have enabled multiomics profiling at unprecedented resolution; however, mass spectrometry-based single-cell proteomics (scMS) remains inherently destructive, precluding simultaneous transcriptomic capture. Unlike antibody-based methods such as CITE-seq, which permit paired profiling but are restricted to targeted protein panels, scMS provides unbiased, genome-scale coverage of the intracellular proteome yet necessitates post hoc integration of unpaired datasets. This diagonal integration challenge, where transcriptomes and proteomes are measured in separate cells lacking shared anchors, remains underserved by existing reviews, which focus predominantly on vertical integration strategies enabled by non-destructive assays. RESULTS: We survey the complete computational pipeline for constructing mechanistic proteogenomic networks from unpaired single-cell data, covering: (i) unimodal network inference such as knowledge-based approaches, probabilistic graphical models, temporal directionality inference, and generative and foundation model strategies that establish the transcriptomic scaffold; (ii) cross-modal integration architectures such as network propagation, graph neural networks (scMRDR, scmFormer, scCotag), and consensus frameworks designed explicitly for the unpaired proteomics setting; and (iii) benchmarking paradigms spanning network reconstruction (BEELINE, GRETA, CausalBench) and multi-task integration evaluation (scMultiBench, SCMMIB), with guidance on metric selection under network sparsity and class imbalance. We identify three principal axes of future development: generative proteomic translation from transcriptomic precursors, inductive prior embedding in next-generation architectures, and perturbation-based causal benchmarking. AVAILABILITY AND IMPLEMENTATION: This is a review article; no novel software is distributed. A curated benchmark resource table, methods starter guide, and per-method bottleneck annotations are provided in the Supplementary Material.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Marcello Barylli, Joyaditya Saha, Tineke E Buffart, Jan Koster, Kristiaan J Lenos, Louis Vermeulen, Roland V Bumbuc, Vivek M Sheraton. 2026-06-01. Network methods for diagonal integration of unpaired single-cell multiomics data: a review.. https://doi.org/10.1093/bioinformatics%2Fbtag353

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Advancing the Deciphering of Host-Microbe Crosstalk with Spatial Omics: A Mini-Review.

Host-microbe crosstalk refers to the reciprocal influences between a host and its resident or invading microorganisms. This crosstalk plays important roles in maintaining host health, regulating physiological functions, and coordinating responses to infection. The rapid rise of spatial omics is transforming how this crosstalk is studied in both animals and plants. Unlike traditional bulk omics, which homogenize tissues and erase spatial context, spatial methods preserve in situ organization and can simultaneously capture molecular information from hosts and microbes. As a result, researchers can characterize the spatial organization of colonization and infection, identify spatial associations between microbial niches and host cell states, and visualize local host response gradients across intact tissues. Current spatial omics technologies encompass sequencing-based, imaging-based, and hybrid platforms. Spatial multi-omics approaches enable the joint measurement or integration of gene expression, protein abundance, and metabolite distributions. Although spatial association alone does not establish causality, spatial omics provides a high-resolution framework for characterizing host-microbe relationships within intact tissues and generating spatially constrained, testable hypotheses. When combined with perturbation experiments and complementary experimental evidence, these hypotheses can contribute to mechanistic interpretation of host-microbe crosstalk. Here, we review spatial omics technologies, compare their suitability and major trade-offs for host-microbe studies, and discuss computational strategies, analytical challenges, and future prospects.

Multiomics

Integrating multi-omics technologies to decipher microbiome functions.

Multi-omics approaches have revolutionized our understanding of microbial communities by enabling simultaneous interrogation of genomic, transcriptomic, proteomic, and metabolomic data. The systematic integration and analysis of these deep datasets help decipher the functional roles of microbiomes, providing critical insights into microbial activities, interactions, and dynamics across diverse environments. Biological complexity makes multi-omics analysis of a single, isolated organism demanding but highly informative, yet this complexity increases further when samples comprise hundreds to thousands of individual species. As microbiome research continues to expand into clinical, environmental, and engineered systems, standardized workflows, benchmarked datasets, and community-driven initiatives are essential to ensure reproducibility, standardization and interpretability. Establishing and disseminating best practices for experimental design, data processing, and integrative analyses will be critical for maximizing comparability and scientific rigor across studies. This perspective highlights recent advances in multi-omics microbiome research, outlines key obstacles in data integration and metadata harmonization, and proposes a collaborative roadmap for scalable, FAIR-compliant multi-omics investigations and potentially disruptive Artificial Intelligence (AI) advances comparable to those of AlphaFold in the field of microbiome science.

Multiomics

Scalable, generalizable and uncertainty-aware integration of spatial multiomics across diverse modalities and platforms with SCIGMA.

Recent advances in spatial omics technologies have enabled simultaneous profiling of transcriptomic, proteomic, epigenomic, metabolomic and imaging data at high spatial resolution, offering unprecedented opportunities to dissect tissue complexity. However, integrating these diverse and large-scale spatial multimodal datasets remains a major computational challenge. We present SCIGMA, a scalable and generalizable deep learning framework for spatial multiomics integration. SCIGMA introduces an uncertainty-aware contrastive learning objective and multiview graph neural networks to preserve modality-specific signals while learning biologically meaningful joint representations. Unlike previous methods, SCIGMA provides spatially resolved uncertainty estimates, interpretably identifying regions of biological or technical heterogeneity. SCIGMA supports integration of up to five modalities, and its modular framework is extensible to future technologies with even more modalities. It also scales to more than 1 million spatial locations, enabling analysis of high-resolution datasets such as Visium HD and Xenium Prime. We evaluated SCIGMA across 19 datasets spanning 8 modalities, 10 tissues and 9 platforms. On benchmarkable datasets, SCIGMA outperformed other methods in spatial domain detection, modality preservation, feature reconstruction and reproducibility. SCIGMA identifies biologically meaningful structures, refined spatial domains and modality-specific regulatory programs, providing a robust, flexible and future-ready solution for scalable spatial multimodal integration.

Multiomics