PubMed Health⌕ Search

PubMed · 38105974

Inferring Metabolic States from Single Cell Transcriptomic Data via Geometric Deep Learning.

Abstract

The ability to measure gene expression at single-cell resolution has elevated our understanding of how biological features emerge from complex and interdependent networks at molecular, cellular, and tissue scales. As technologies have evolved that complement scRNAseq measurements with things like single-cell proteomic, epigenomic, and genomic information, it becomes increasingly apparent how much biology exists as a product of multimodal regulation. Biological processes such as transcription, translation, and post-translational or epigenetic modification impose both energetic and specific molecular demands on a cell and are therefore implicitly constrained by the metabolic state of the cell. While metabolomics is crucial for defining a holistic model of any biological process, the chemical heterogeneity of the metabolome makes it particularly difficult to measure, and technologies capable of doing this at single-cell resolution are far behind other multiomics modalities. To address these challenges, we present GEFMAP (Gene Expression-based Flux Mapping and Metabolic Pathway Prediction), a method based on geometric deep learning for predicting flux through reactions in a global metabolic network using transcriptomics data, which we ultimately apply to scRNAseq. GEFMAP leverages the natural graph structure of metabolic networks to learn both a biological objective for each cell and estimate a mass-balanced relative flux rate for each reaction in each cell using novel deep learning models.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Holly Steach, Siddharth Viswanath, Yixuan He, Xitong Zhang, Natalia Ivanova, Matthew Hirn, Michael Perlmutter, Smita Krishnaswamy. 2023-12-07. Inferring Metabolic States from Single Cell Transcriptomic Data via Geometric Deep Learning.. https://doi.org/10.1101/2023.12.05.570153

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

AAV-mediated genome editing is influenced by the formation of R-loops.

Recombinant adeno-associated viral vectors (rAAV) hold an intrinsic ability to stimulate homologous recombination (AAV-HR) and are the most used in clinical settings for in vivo gene therapy. However, rAAVs also integrate throughout the genome. Here, we describe DNA-RNA immunoprecipitation sequencing (DRIP-seq) in murine HEPA1-6 hepatoma cells and whole murine liver to establish the similarities and differences in genomic R-loop formation in a transformed cell line and intact tissue. We show enhanced AAV-HR in mice upon genetic and pharmacological upregulation of R-loops. Selecting the highly expressed Albumin gene as a model locus for genome editing in both in vitro and in vivo experiments showed that the R-loop prone, 3' end of Albumin was efficiently edited by AAV-HR, whereas the upstream R-loop-deficient region did not result in detectable vector integration. In addition, we found a positive correlation between previously reported off-target rAAV integration sites and R-loop enriched genomic regions. Thus, we conclude that high levels of R-loops, present in highly transcribed genes, promote rAAV vector genome integration. These findings may shed light on potential mechanisms for improving the safety and efficacy of genome editing by modulating R-loops and may enhance our ability to predict regions most susceptible to off-target insertional mutagenesis by rAAV vectors.

Preprint↗

An essential and highly selective protein import pathway encoded by nucleus-forming phage.

UNLABELLED: Targeting proteins to specific subcellular destinations is essential in prokaryotes, eukaryotes, and the viruses that infect them. Chimalliviridae phages encapsulate their genomes in a nucleus-like replication compartment composed of the protein chimallin (ChmA) that excludes ribosomes and decouples transcription from translation. These phages selectively partition proteins between the phage nucleus and the bacterial cytoplasm. Currently, the genes and signals that govern selective protein import into the phage nucleus are unknown. Here we identify two components of this novel protein import pathway: a species-specific surface-exposed region of a phage intranuclear protein required for nuclear entry and a conserved protein, PicA, that facilitates cargo protein trafficking across the phage nuclear shell. We also identify a defective cargo protein that is targeted to PicA on the nuclear periphery but fails to enter the nucleus, providing insight into the mechanism of nuclear protein trafficking. Using CRISPRi-ART protein expression knockdown of PicA, we show that PicA is essential early in the chimallivirus replication cycle. Together our results allow us to propose a multistep model for the Protein Import Chimallivirus (PIC) pathway, where proteins are targeted to PicA by amino acids on their surface, and then licensed by PicA for nuclear entry. The divergence in the selectivity of this pathway between closely-related chimalliviruses implicates its role as a key player in the evolutionary arms race between competing phages and their hosts. SIGNIFICANCE STATEMENT: The phage nucleus is an enclosed replication compartment built by Chimalliviridae phages that, similar to the eukaryotic nucleus, separates transcription from translation and selectively imports certain proteins. This allows the phage to concentrate proteins required for DNA replication and transcription while excluding DNA-targeting host defense proteins. However, the mechanism of selective trafficking into the phage nucleus is currently unknown. Here we determine the region of a phage nuclear protein that targets it for nuclear import and identify a conserved, essential nuclear shell-associated protein that plays a key role in this process. This work provides the first mechanistic model of selective import into the phage nucleus.

Preprint↗

Comprehensive analyses of a large human gut Bacteroidales culture collection reveal species and strain level diversity and evolution.

Species of the Bacteroidales order are among the most abundant and stable bacterial members of the human gut microbiome with diverse impacts on human health. While Bacteroidales strains and species are genomically and functionally diverse, order-wide comparative analyses are lacking. We cultured and sequenced the genomes of 408 Bacteroidales isolates from healthy human donors representing nine genera and 35 species and performed comparative genomic, gene-specific, mobile gene, and metabolomic analyses. Families, genera, and species could be grouped based on many distinctive features. However, we also show extensive DNA transfer between diverse families, allowing for shared traits and strain evolution. Inter- and intra-specific diversity is also apparent in the metabolomic profiling studies. This highly characterized and diverse Bacteroidales culture collection with strain-resolved genomic and metabolomic analyses can serve as a resource to facilitate informed selection of strains for microbiome reconstitution.

Preprint↗