PubMed HealthSearch

PubMed · 42331253

Limited Impact of Column Chemistry and Length on Proteome Coverage Under High-Speed DIA.

Abstract

The evolution of mass spectrometry (MS)-based proteomics has been driven by continuous technological advances in sample preparation, liquid-phase separations, instrumentation, and data acquisition. Chromatographic performance has been recognized as a contributing factor to identification depth, particularly on earlier-generation MS platforms. Recent advances in MS sampling speed and sensitivity now raise the question of how strongly chromatographic quality continues to determine overall proteome coverage. We investigate how column chemistry and length influence proteome coverage and chromatographic selectivity under modern data-independent acquisition conditions, and whether traditional optimization priorities still apply. Spanning a matrix of experiments with five distinct stationary phases, including C18 chemistries, C8, and Phenyl-Hexyl, across eight column lengths (40-140 mm), we evaluate protein identification performance using data-independent acquisition on the Orbitrap Astral mass spectrometer. Despite differences in stationary-phase chemistry and column length, we observed remarkably convergent proteome coverage metrics. All C18 and C8 phases consistently achieved over 150,000 precursor- and approximately 9000 protein group identifications, regardless of column length variations. While retention fingerprints persisted across chemistries, these chromatographic differences did not translate into meaningful variations in proteome coverage under high-speed acquisition conditions at 200 Hz. Within the range of modern sub-2 μm reversed-phase materials tested, identification depth showed limited dependence on column chemistry and length, suggesting that for state-of-the-art stationary phases, method development priorities may increasingly favor operational robustness, throughput, and reproducibility over traditional separation optimization.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Alicia-Sophie Schebesta, Kathrin Korff, Ericka C M Itang, Vincent Albrecht, Philipp E Geyer, Johannes B Mueller-Reif. 2026-06-23. Limited Impact of Column Chemistry and Length on Proteome Coverage Under High-Speed DIA.. https://doi.org/10.1016/j.mcpro.2026.101609

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Selective Enrichment of Newly Synthesized Proteins Using Phos-Tag Click Tip Enables Nascent Proteome Analysis in Influenza A Virus Infection.

Profiling of newly synthesized proteins (NSPs) provides access to dynamic changes in protein production that accompany acute cellular responses. Bioorthogonal noncanonical amino acid tagging (BONCAT)-based approaches enable selective labeling of NSPs; however, their broader application remains constrained by labor-intensive enrichment workflows and limited sensitivity for direct peptide-level analysis. Here, we developed a workflow termed "Phos-tag Click Tip" by integrating a phosphorylated variant of bicyclononyne (pBCN) with Phos-tag affinity purification to selectively capture azidohomoalanine (AHA)-labeled peptides for newly synthesized proteome analysis (NSProteomics). This approach overcomes key limitations of conventional proteomics and BONCAT-based strategies by enabling efficient enrichment and sensitive detection of NSP-derived peptides. Using this workflow, we performed comprehensive NSP profiling of host cells during influenza A virus infection. We identified dynamic changes in distinct NSP profiles associated with viral replication, host restriction, and immune responses, many of which were not readily detected with conventional whole-cell- or phospho-proteomic analyses. Overall, the Phos-tag Click Tip workflow provides a complementary approach for stimulus-responsive NSP profiling, offering functionally relevant insights into host-virus interactions and cellular response mechanisms.

Proteome

CoMR: an integrative scoring pipeline for comprehensive mitochondrial proteome reconstruction across eukaryotes.

Mitochondrial proteome reconstruction from eukaryotic sequence data typically relies on prediction of mitochondrial targeting signals (MTSs). However, MTS predictors are primarily trained on model organisms and may perform poorly in phylogenetically divergent lineages or in organisms with atypical or reduced targeting sequences. Accurate reconstruction therefore requires integration of complementary sources of evidence beyond targeting prediction alone. We developed Comprehensive Mitochondrial Reconstructor (CoMR), an integrative workflow that combines targeting prediction, curated homology searches, large-scale similarity searches, and automated phylogenetic analysis within a unified scoring framework. Benchmarking on the model yeast Saccharomyces cerevisiae yielded strong discriminatory performance [receiver operating characteristic (ROC)-area under the curve (AUC) = 0.92], exceeding standalone prediction with TargetP2, a predictor of N-terminal targeting peptides (ROC-AUC = 0.72). In the divergent anaerobic protist Paratrimastix pyriformis, CoMR maintained robust performance (ROC-AUC = 0.86) validated with an experimental proteome despite extreme class imbalance, achieving a precision-recall AUC of 0.183 (~78-fold enrichment over random expectation and ~10-fold improvement over TargetP2). Ablation analyses demonstrate that predictive performance is robust to individual evidence-layer removal, while overlap analyses showed that homology-based searches recovered candidates missed by targeting predictors, particularly in P. pyriformis. Overall, CoMR improves mitochondrial proteome reconstruction over targeting prediction alone and provides a reproducible workflow for predicting mitochondrial and mitochondrion-related organelle protein repertoires across eukaryotes to aid investigations of organelle evolution and proteome reduction.

Proteome

ECLIPSE: exploring the dark proteome of ESKAPE pathogens through the sequence similarity network of the Protein Universe Atlas.

MOTIVATION: The accelerating crisis of antimicrobial resistance among the critical so-called ESKAPE pathogens demands the urgent identification of novel molecular targets. However, a substantial fraction of ESKAPE proteomes remains functionally uncharacterized, with many genes annotated as encoding hypothetical proteins. These protein sequences often lack significant similarity to known protein families when conventional homology-based annotation methods are used and thus remain "dark". This limits our ability to explore their roles in pathogenicity, and it is thus crucial to bridge this substantial gap in pathogen biology by developing new strategies to illuminate these "dark" regions of the ESKAPE pan-proteome. RESULTS: We introduce ECLIPSE (ESKAPE Connectome Linkage and Inference for Proteome Sequence Exploration), a network-based computational framework that systematically identifies and prioritizes functionally dark protein families in ESKAPE pan-proteomes. ECLIPSE embeds target ESKAPE pathogen proteomes within the global sequence similarity network of the Protein Universe Atlas. It detects connected components composed entirely of unannotated proteins, called the "dark proteome." As a case study, we applied ECLIPSE to a pan-proteome of 3 460 657 protein sequences from 635 strains of Pseudomonas aeruginosa (PA). ECLIPSE identified 120 985 proteins (4%) residing in completely dark connected components. Furthermore, we have performed a taxonomic diversity analysis using normalized Shannon indices to characterize each dark component by its enrichment in ESKAPE pathogens. The analysis utilized the evenness (E) value (see Methods 2.1), which distinguishes Pseudomonas-specific (target-specific) from ESKAPE-enriched dark components. We then developed the Dark Proteome Prioritization Score (DPPS), a composite multidimensional scoring framework (see Methods 2.5). It ranks these dark components by biological relevance across four orthogonal axes: (i) functional darkness, (ii) P. aeruginosa proportion in the Atlas, (iii) AMR-clade taxonomic restriction, and (iv) conservation across the 635 P. aeruginosa strains. This framework outputs a robust four-tier scoring system; the prioritized Tier I components were validated by weight sensitivity analysis and remained stable across 500 Monte Carlo weight perturbations. Structural characterization of one of the top-ranked ESKAPE-enriched dark components revealed that it belongs to the beta-barrel fold DUF1302 (PF06980) family, for which no experimentally solved three-dimensional structure exists in the PDB. The genomic context analysis indicates that it is co-localized with a LuxR-type transcriptional regulator. Collectively, ECLIPSE identifies evolutionarily conserved, structurally defined, and functionally dark proteins enriched across ESKAPE pathogens; these dark proteins can further be utilized as alternative antimicrobial targets for experimental characterization. AVAILABILITY AND IMPLEMENTATION: The source code and dataset are available for free at: Github: https://github.com/surabhilata/ECLIPSE.git, Zenodo: DOI: 10.5281/zenodo.21064323.

Proteome