PubMed Health⌕ Search

PubMed · 16603205

Model identification for DNA sequence-structure relationships.

Abstract

We investigate the use of algebraic state-space models for the sequence dependent properties of DNA. By considering the DNA sequence as an input signal, rather than using an all atom physical model, computational efficiency is achieved. A challenge in deriving this type of model is obtaining its structure and estimating its parameters. Here we present two candidate model structures for the sequence dependent structural property Slide and a method of encoding the models so that a recursive least squares algorithm can be applied for parameter estimation. These models are based on the assumption that the value of Slide at a base-step is determined by the surrounding tetranucleotide sequence. The first model takes the four bases individually as inputs and has a median root mean square deviation of 0.90 A. The second model takes the four bases pairwise and has a median root mean square deviation of 0.88 A. These values indicate that the accuracy of these models is within the useful range for structure prediction. Performance is comparable to published predictions of a more physically derived model, at significantly less computational cost.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Stephen Dwyer Hawley, Anita Chiu, Howard Jay Chizeck. 2006-04-17. Model identification for DNA sequence-structure relationships.. https://doi.org/10.1016/j.mbs.2006.02.003

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Prediction of side-chain conformations on protein surfaces.

An approach is described that improves the prediction of the conformations of surface side chains in crystal structures, given the main-chain conformation of a protein. A key element of the methodology involves the use of the colony energy. This phenomenological term favors conformations found in frequently sampled regions, thereby approximating entropic effects and serving to smooth the potential energy surface. Use of the colony energy significantly improves prediction accuracy for surface side chains with little additional computational cost. Prediction accuracy was quantified as the percentage of side-chain dihedral angles predicted to be within 40 degrees of the angles measured by X-ray diffraction. Use of the colony energy in predictions for single side chains improved the prediction accuracy for chi(1) and chi(1+2) from 65 and 40% to 74 and 59%, respectively. Several other factors that affect prediction of surface side-chain conformations were also analyzed, including the extent of conformational sampling, details of the rotamer library employed, and accounting for the crystallographic environment. The prediction of conformations for polar residues on the surface was generally found to be more difficult than those for hydrophobic residues, except for polar residues participating in hydrogen bonds with other protein groups. For surface residues with hydrogen-bonded side chains, the prediction accuracy of chi(1) and chi(1+2) was 79 and 63%, respectively. For surface polar residues, in general (all side-chain prediction), the accuracy of chi(1) and chi(1+2) was only 73 and 56%, respectively. The most accurate results were obtained using the colony energy and an all-atom description that includes neighboring molecules in the crystal (protein chains and hetero atoms). Here, the accuracy of chi(1) and chi(1+2) predictions for surface side chains was 82 and 73%, respectively. The root mean square deviations obtained for hydrogen-bonding surface side chains were 1.64 and 1.81 A, with and without consideration of crystal packing effects, respectively.

Crystallography, X-Ray↗

Double-stranded cycles: toward C84's belt region.

The reactivity of the double-stranded hydrocarbon cycle with two ether bridges (1) toward iodotrimethylsilane (TMSI) was investigated in some detail. The carbon skeleton of cycle 1 resembles the belt region of a C84 fullerene which makes it a potential precursor to the long sought after fully aromatic derivative. Upon exposure to TMSI, cycle 1 undergoes a cascade of reactions which involve different states of iodination/reduction which ultimately lead to the hydrogenated cycle 5a, whose structure was proven by single-crystal X-ray analysis. A deeper insight into mechanistic aspects of this sequence of conversions was gained by performing the reaction under dry and wet conditions, whereby the latter involved both normal and deuterated water. With the help of detailed NMR correlation studies and DFT computations, all important aspects were clarified including an unexpected selective H/D exchange at the naphthalenic moieties.

Crystallography, X-Ray↗