PubMed Health⌕ Search

PubMed · 16225676

ProfNet, a method to derive profile-profile alignment scoring functions that improves the alignments of distantly related proteins.

Abstract

BACKGROUND: Profile-profile methods have been used for some years now to detect and align homologous proteins. The best such methods use information from the background distribution of amino acids and substitution tables either when constructing the profiles or in the scoring. This makes the methods dependent on the quality and choice of substitution table as well as the construction of the profiles. Here, we introduce a novel method called ProfNet that is used to derive a profile-profile scoring function. The method optimizes the discrimination between scores of related and unrelated residues and it is fast and straightforward to use. This new method derives a scoring function that is mainly dependent on the actual alignment of residues from a training set, and it does not use any additional information about the background distribution. RESULTS: It is shown that ProfNet improves the discrimination of related and unrelated residues. Further it can be used to improve the alignment of distantly related proteins. CONCLUSION: The best performance is obtained using superfamily related proteins in the training of ProfNet, and a classifier that is related to the distance between the structurally aligned residues. The main difference between the new scoring function and a traditional profile-profile scoring function is that conserved residues on average score higher with the new function.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Tomas Ohlson, Arne Elofsson. 2005-10-14. ProfNet, a method to derive profile-profile alignment scoring functions that improves the alignments of distantly related proteins.. https://doi.org/10.1186/1471-2105-6-253

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Comparing ARG Inference Methods Under Transmission of Reproductive Success: Tree Imbalance Matters.

Inferring coalescent trees from genomic data has become a major subject in population genetics, particularly with the recent advances in tree sequence reconstruction methods. However, it remains unclear how well these methods perform for imbalanced genealogies. Such imbalances can arise from processes such as cultural transmission of reproductive success (CTRS) or positive selection. Using simulated genomic data, we benchmarked three major software packages, SINGER, Relate, and tsinfer, by comparing the imbalance of reconstructed trees by these methods with that of the true simulated trees, for three indices that quantify this imbalance. The three methods performed well under scenarios yielding balanced trees. However, their accuracy declined as imbalance increased. Performances also varied with mutation rate, recombination rate, and sample size. This study opens possibilities for applying these methods to infer CTRS or positive selection in large-scale genomic datasets, using simulation-based inference such as approximate Bayesian computation.

Models, Genetic↗

Diffusive Noise Controls Early Stages of Genetic Demixing.

Theoretical descriptions of the stepping-stone model, a cornerstone of spatial population genetics, have long overlooked diffusive noise arising from migration dynamics. We derive a fluctuating hydrodynamic description of this model from microscopic rules, which we then use to demonstrate that diffusive noise significantly alters early-time genetic demixing, which we characterize through heterozygosity, a key measure of diversity. Combining macroscopic fluctuation theory and microscopic simulations, we demonstrate that the scaling of allele-number fluctuations in a spatial domain displays an early-time behavior dominated by diffusive noise. Our results underscore the need for additional terms in existing continuum theories and highlight the necessity of including diffusive noise in models of spatially structured populations.

Models, Genetic↗

The relationship between the pleiotropic phenotypic effects of a mutation fixed by selection.

A pleiotropic model of mutation is presented that allows for correlations between the effects of a new mutation and for the distribution of mutational effects to vary from being leptokurtic to normally distributed. Using this model I quantify how selection transforms the correlation between the effects of a new (random) mutation into the correlation between the effects of a mutation that is fixed by selection and contributes to an adaptation. Results suggest that under most conditions the correlation between the effects of a fixed mutation is less than the correlation between the effects of a new mutation. I also generalize previous results that quantified the expected size of a fixed mutation's effect on a character given an observed effect of that mutation on another character. In agreement with previous results, work here suggests that as the observed effect becomes large and beneficial the expected effect on another character approaches the expected effect of a new (random) mutation given the observed effect. Lastly, these theoretical results are related to recent empirical work that found beneficial mutations had a positive correlation in their pleiotropic effects.

Models, Genetic↗