PubMed Health⌕ Search

PubMed · 15584737

Refinement of NMR structures using implicit solvent and advanced sampling techniques.

Abstract

NMR biomolecular structure calculations exploit simulated annealing methods for conformational sampling and require a relatively high level of redundancy in the experimental restraints to determine quality three-dimensional structures. Recent advances in generalized Born (GB) implicit solvent models should make it possible to combine information from both experimental measurements and accurate empirical force fields to improve the quality of NMR-derived structures. In this paper, we study the influence of implicit solvent on the refinement of protein NMR structures and identify an optimal protocol of utilizing these improved force fields. To do so, we carry out structure refinement experiments for model proteins with published NMR structures using full NMR restraints and subsets of them. We also investigate the application of advanced sampling techniques to NMR structure refinement. Similar to the observations of Xia et al. (J.Biomol. NMR 2002, 22, 317-331), we find that the impact of implicit solvent is rather small when there is a sufficient number of experimental restraints (such as in the final stage of NMR structure determination), whether implicit solvent is used throughout the calculation or only in the final refinement step. The application of advanced sampling techniques also seems to have minimal impact in this case. However, when the experimental data are limited, we demonstrate that refinement with implicit solvent can substantially improve the quality of the structures. In particular, when combined with an advanced sampling technique, the replica exchange (REX) method, near-native structures can be rapidly moved toward the native basin. The REX method provides both enhanced sampling and automatic selection of the most native-like (lowest energy) structures. An optimal protocol based on our studies first generates an ensemble of initial structures that maximally satisfy the available experimental data with conventional NMR software using a simplified force field and then refines these structures with implicit solvent using the REX method. We systematically examine the reliability and efficacy of this protocol using four proteins of various sizes ranging from the 56-residue B1 domain of Streptococcal protein G to the 370-residue Maltose-binding protein. Significant improvement in the structures was observed in all cases when refinement was based on low-redundancy restraint data. The proposed protocol is anticipated to be particularly useful in early stages of NMR structure determination where a reliable estimate of the native fold from limited data can significantly expedite the overall process. This refinement procedure is also expected to be useful when redundant experimental data are not readily available, such as for large multidomain biomolecules and in solid-state NMR structure determination.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jianhan Chen, Wonpil Im, Charles L Brooks. 2004-12-15. Refinement of NMR structures using implicit solvent and advanced sampling techniques.. https://doi.org/10.1021/ja047624f

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Prediction of bacterial protein-compound interactions with only positive samples.

MOTIVATION: Prediction of Compound-Protein Interactions (CPI) in bacteria is crucial to advance various pharmaceutical and chemical engineering fields, including biocatalysis, drug discovery, and industrial processing. However, current CPI models cannot be applied for bacterial CPI prediction due to the lack of curated negative interaction samples. RESULTS: We propose a novel Positive-Unlabeled (PU) learning framework, named BIN-PU, to address this limitation. BIN-PU generates pseudo positive and negative labels from known positive interaction data, enabling effective training of deep learning models for CPI prediction. We also propose a weighted positive loss function that weights to truly positive samples. We have validated BIN-PU coupled with multiple CPI backbone models, comparing the performance with the existing PU models using bacterial cytochrome P450 (CYP) data. Extensive experiments demonstrate the superiority of BIN-PU over the benchmark models in predicting CPIs with only truly positive samples. Furthermore, we have validated BIN-PU on additional bacterial proteins obtained from literature review, human CYP datasets, and uncurated data for its reproducibility. We have also validated the CPI prediction for the uncurated CYP data with biological and biophysical experiments. BIN-PU represents a significant advancement in CPI prediction for bacterial proteins, opening new possibilities for improving predictive models in related biological interaction tasks. AVAILABILITY AND IMPLEMENTATION: The source code and data are available at https://github.com/datax-lab/CYP.

Bacterial Proteins↗

ComFB, a widespread family of c-di-NMP receptor proteins.

Cyclic dimeric-GMP (c-di-GMP) is a ubiquitous bacterial second messenger that regulates a variety of cellular processes, including motility, biofilm formation, secretion, cell cycle progression, and development, and also contributes to the virulence of many bacterial pathogens. While the genes encoding c-di-GMP cyclases and hydrolases are readily identifiable in microbial genomes, known c-di-GMP receptor domains are quite few, with only PilZ and MshEN broadly distributed across bacterial phyla. Recently, a new c-di-GMP receptor, named CdgR or ComFB, has been identified in cyanobacteria and shown to regulate cell size and natural competence. We demonstrated that CdgR proteins exhibit sequence and structural similarity to the Bacillus subtilis late competence development protein ComFB, a conserved protein of unknown function associated with bacterial competence. This prompted us to hypothesize that ComFB and ComFB-like proteins could also serve as c-di-GMP receptors. Here, we comprehensively investigated the ComFB protein family and demonstrated that ComFB proteins are evolutionarily widespread among bacteria and function as a novel family of c-di-GMP receptors. We showed that ComFB proteins from Gram-positive bacteria (B. subtilis, Thermoanaerobacter brockii) and Gram-negative pathogens (Vibrio cholerae, Treponema denticola) bind c-di-GMP with high affinity. Several ComFB proteins also bind cyclic di-adenosine monophosphate (c-di-AMP), suggesting that ComFB represents a widely distributed bacterial protein family with dual specificity for c-di-GMP and c-di-AMP. Our physiological studies further showed that ComFB plays vital roles in controlling motility in a c-di-GMP-dependent manner in two phylogenetically distant bacteria, B. subtilis and the gram-negative Shewanella oneidensis, attesting to the biological relevance of ComFB as a c-di-GMP binding protein.

Bacterial Proteins↗

Dynamic structural determinants in bacterial microcompartment shells.

Bacterial microcompartments (BMCs) are polyhedral structures that segregate enzymatic cargo from the cytosol via encapsulation within a protein shell. Unlike other biological polyhedra, such as viral capsids and encapsulins, BMC shells can exhibit a highly advantageous structural and functional plasticity, conforming to a variety of anabolic (CO2 fixation in carboxysomes) and catabolic (nutrient assimilation in metabolosomes) roles. Consequently, understanding the subunit properties and associated protein-protein interaction processes that guide shell assembly and function is a necessary step to fully harness BMCs as modular, biotechnological nanomachines. Here, we describe the recent insights into the dynamics of structural features of the key BMC domain (Pfam00936)-containing proteins, which serve as a structural template for BMC-H and BMC-T shell building blocks.

Bacterial Proteins↗