PubMed Health⌕ Search

PubMed · 3627775

Information theory and the genetic code.

Abstract

The genetic code, which directs the protein biosynthesis, is an information system. Although all its details are not known at present, its essential characteristics are elucidated, as well for the replication or transcription as for the translation of the genetic message. A coherent picture now appears, which reveals the existence of an universal structure, the most fundamental features of which seem to obey some logic. A systematic approach has been devised, which aims to their integration in a theorectical scheme: many features of the code table can thus be interpreted as resulting from a unique principle of best resistance against the effects of mutations. Any group of triplets or amino-acids can be considered along this line. It is more difficult however, to analyse the coexistence of two (or more) different groups. In this work, we propose to extend our optimization principle into a more general one, which includes the notion of information as defined by Shannon. We explore some consequences of this new principle in the most simple models that one can build for the origin and evolution of the genetic code.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

A Figureau. 1987. Information theory and the genetic code.. https://doi.org/10.1007/bf02386481

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Alternative genetic codes in bacteria and archaea identified with a fast k-mer-based algorithm.

The genetic code is conserved across all domains of life and is often described as universal. Nevertheless, many exceptions to the "universal" code have now been documented, most of these through manual or semiautomated inspection of highly conserved genes. Modern bioinformatics tools improved our ability to find alternative genetic codes but remain computationally expensive, preventing widespread use on thousands of new species identified by sequencing environmental samples. Here, I report a >100-fold accelerated method for inferring the genetic code directly from assembled genomes and apply it to thousands of previously uncharacterized assemblies from archaea and bacteria. I describe three candidate genetic code variations, one of which, an alternative genetic code used by a family of Asgard archaea, is a unique example of sense codon reassignments for this domain. Identifying genetic code variations is important for understanding evolution of the standard code and improving accuracy of protein databases and open reading frame identification.

Genetic Code↗

The genetic code at the balance point of error and demand.

The origin and organizing principles of the genetic code remain central problems in molecular evolution. The low probability of the natural codon-to-amino acid mapping arising by chance has spurred the hypothesis that its structure is optimized for robustness to mutations and translational errors. For the construction of effective molecular machines, the repertoire of encoded amino acids must also be diverse enough in physicochemical features. Here, we examine whether the standard genetic code can be understood as a near-optimal solution balancing these two objectives: minimizing error load and aligning codon assignments with the naturally occurring amino acid composition. Using simulated annealing, we explore this trade-off across a broad range of parameters. We find that the standard genetic code resides near an optimum in the fitness landscape of possible genetic codes. The degeneracy of the code plays a dual role, minimizing mistranslation errors while matching codon multiplicity to amino acid usage frequencies. As a result, uniform codon usage alone is sufficient to recover the empirical amino acid composition, without any additional bias. It is a highly effective solution that balances fidelity against resource availability constraints. A comparative analysis of natural variants also reveals a functional decoupling: error robustness acts as a rigid global constraint determined by code topology, whereas compositional alignment serves as a more flexible variable that adapts to lineage-specific demands. These results support a multi-objective optimization framework in which the genetic code reflects a balance between translational fidelity and proteomic demand.

Genetic Code↗

Reading another hidden message in the genetic code.

The genetic code determines not only the amino acid sequences of proteins but also mRNA stability. How is this hidden message read? Hia and colleagues have now identified human DHX29 as a reader of the mRNA stability code carried by codons, providing new mechanistic insights into translation-coupled gene regulation.

Genetic Code↗