PubMed · 12875825
Mutually symmetric and complementary triplets: differences in their use distinguish systematically between coding and non-coding genomic sequences.
Abstract
The general property of asymmetry in word use in meaningful texts written in a variety of languages, motivates a quantification of the differences in the use of mutually symmetric triplets in genomic sequences. When this is done in the three reading frames, high values found for one of them are used as indication that the sequence is coding for a protein. Moreover, a similar quantification of the differences in the use of complementary triplets is introduced, again with predictive power of the coding character of a sequence. This method reflects the non-equivalence between sense and anti-sense strand of a coding segment. In both approaches, "linguistic asymmetry" in coding sequences is related to the form of the genetic code and to the bias in codon usage and amino acid use skews.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Christoforos Nikolaou, Yannis Almirantis. 2003-08-21. Mutually symmetric and complementary triplets: differences in their use distinguish systematically between coding and non-coding genomic sequences.. https://doi.org/10.1016/s0022-5193(03)00123-1
Cite the original work for its findings. Save a collection to share your selection of sources.