PubMed Health⌕ Search

Biomedical subjects

A P Cootes

Publications and source records attributed to A P Cootes.

2 recordsLinked to original sources

Automated protein design and sequence optimisation: scoring functions and the search problem.

Advances in molecular biology may mean that almost any protein sequence can be synthesised, but perhaps this has served to highlight the inadequacy of theoretical work. For a given protein fold, it is probably not possible to reliably predict an "ideal" sequence. We identify and survey several aspects of the problem. Firstly, it is not clear what is the best way to score a sequence-structure pair. Secondly, there is no consensus as to what the score function should represent (free energy or some abstract measure of sequence-structure compatibility). Finally, the number of possible sequences is astronomical and searching this space poses a daunting optimisation problem. These problems are discussed in the light of recent experimental successes.

Algorithms↗

The dependence of amino acid pair correlations on structural environment.

A statistical analysis was performed to determine to what extent an amino acid determines the identity of its neighbors and to what extent this is determined by the structural environment. Log-linear analysis was used to discriminate chance occurrence from statistically meaningful correlations. The classification of structures was arbitrary, but was also tested for significance. A list of statistically significant interaction types was selected and then ranked according to apparent importance for applications such as protein design. This showed that, in general, nonlocal, through-space interactions were more important than those between residues near in the protein sequence. The highest ranked nonlocal interactions involved residues in beta-sheet structures. Of the local interactions, those between residues i and i + 2 were the most important in both alpha-helices and beta-strands. Some surprisingly strong correlations were discovered within beta-sheets between residues and sites sequentially near to their bridging partners. The results have a clear bearing on protein engineering studies, but also have implications for the construction of knowledge-based force fields.

Amino Acids↗