PubMed Health⌕ Search

SEARCH · PubMed Health

Results for “codon optimization”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23Linked to original sources

Binding of the bacteriophage T4 regA protein to mRNA targets: an initiator AUG is required.

Bacteriophage T4 regA protein translationally represses the synthesis of a subset of early phage-induced proteins. The protein binds to the translation initiation site of at least two mRNAs and prevents formation of the initiation complex. We show here that the protein binds to the translation initiation sites of other regA-sensitive mRNAs. Analysis of mRNA binding by filtration and nuclease protection assays shows that AUG is necessary but not sufficient for specific binding of regA protein to its mRNA targets. Anticipating the need for large quantities of regA protein for structural studies to further define the regA protein-RNA ligand interaction, we also report cloning the regA gene into a T4 overexpression system. The expression of regA protein in uninfected E. coli is lethal, so in our system regA driven by a strong T7 promoter is sequestered in a T4 phage until 'induction' by phage infection is desired. We have replaced the regA sensitive wild-type ribosome binding site with a strong insensitive ribosome binding site at an optimal distance from the regA initiation codon for maximizing expression. We have obtained large amounts of regA protein.

Base Sequence↗

Complete sequence and comparative analysis of the genome of herpes B virus (Cercopithecine herpesvirus 1) from a rhesus monkey.

The complete DNA sequence of herpes B virus (Cercopithecine herpesvirus 1) strain E2490, isolated from a rhesus macaque, was determined. The total genome length is 156,789 bp, with 74.5% G+C composition and overall genome organization characteristic of alphaherpesviruses. The first and last residues of the genome were defined by sequencing the cloned genomic termini. There were six origins of DNA replication in the genome due to tandem duplication of both oriL and oriS regions. Seventy-four genes were identified, and sequence homology to proteins known in herpes simplex viruses (HSVs) was observed in all cases but one. The degree of amino acid identity between B virus and HSV proteins ranged from 26.6% (US5) to 87.7% (US15). Unexpectedly, B virus lacked a homolog of the HSV gamma(1)34.5 gene, which encodes a neurovirulence factor. Absence of this gene was verified in two low-passage clinical isolates derived from a rhesus macaque and a zoonotically infected human. This finding suggests that B virus most likely utilizes mechanisms distinct from those of HSV to sustain efficient replication in neuronal cells. Despite the considerable differences in G+C content of the macaque and B virus genes (51% and 74.2%, respectively), codons used by B virus are optimal for the tRNA population of macaque cells. Complete sequence of the B virus genome will certainly facilitate identification of the genetic basis and possible molecular mechanisms of enhanced B virus neurovirulence in humans, which results in an 80% mortality rate following zoonotic infection.

Animals↗

Avidin expressed in transgenic rice confers resistance to the stored-product insect pests Tribolium confusum and Sitotroga cerealella.

Rice (Oryza sativa var. Nipponbare) was transformed with an artificial avidin gene. The features of this construct are as follows: (1) a signal peptide sequence derived from barley alpha amylase was added at the N-terminal region, (2) codon usage of the gene was optimized for rice, and (3) the gene was driven by rice glutelin GluB-1, an endosperm-specific promoter. Avidin was produced in the grain of the transgenic rice but not in the leaves. The concentration of avidin in the kernels was about 1,800 ppm. All larvae of the confused flour beetle (Tribolium confusum) and Angoumois grain moth (Sitotroga cerealella) died when fed transgenic avidin rice powder or kernels, respectively, whereas most of the test insects developed into adults when they were fed a nontransgenic rice control diet. Avidin extracted from the transgenic rice kernel lost most biotin-binding activity after 5 min heating at 95 degrees C.

Amino Acid Sequence↗

Immunobead-PCR: a technique for the detection of circulating tumor cells using immunomagnetic beads and the polymerase chain reaction.

The presence of tumor cells in the circulation may predict disease recurrence and metastases. We have developed a sensitive technique for the detection of carcinoma cells in blood, using immunomagnetic beads to enrich for epithelial cells and the polymerase chain reaction to identify a tumor marker. The colon carcinoma cell line SW480, homozygous for a K-ras codon 12 mutation, was used to establish optimal conditions. The SW480 cells were serially diluted in normal blood and incubated with immunomagnetic beads labeled with a monoclonal antibody specific for epithelial cells. Cells bound to the beads were retrieved using a magnetic field and the presence of K-ras codon 12 mutations determined by a polymerase chain reaction based analysis. SW480 cells could be detected in dilutions up to 1 SW480 cell/10(5) leukocytes in whole blood.

Base Sequence↗

The genetic code at the balance point of error and demand.

The origin and organizing principles of the genetic code remain central problems in molecular evolution. The low probability of the natural codon-to-amino acid mapping arising by chance has spurred the hypothesis that its structure is optimized for robustness to mutations and translational errors. For the construction of effective molecular machines, the repertoire of encoded amino acids must also be diverse enough in physicochemical features. Here, we examine whether the standard genetic code can be understood as a near-optimal solution balancing these two objectives: minimizing error load and aligning codon assignments with the naturally occurring amino acid composition. Using simulated annealing, we explore this trade-off across a broad range of parameters. We find that the standard genetic code resides near an optimum in the fitness landscape of possible genetic codes. The degeneracy of the code plays a dual role, minimizing mistranslation errors while matching codon multiplicity to amino acid usage frequencies. As a result, uniform codon usage alone is sufficient to recover the empirical amino acid composition, without any additional bias. It is a highly effective solution that balances fidelity against resource availability constraints. A comparative analysis of natural variants also reveals a functional decoupling: error robustness acts as a rigid global constraint determined by code topology, whereas compositional alignment serves as a more flexible variable that adapts to lineage-specific demands. These results support a multi-objective optimization framework in which the genetic code reflects a balance between translational fidelity and proteomic demand.

Genetic Code↗

CUG as a mutant start codon for cat-86 and xylE in Bacillus subtilis.

The cat-86 gene specifies chloramphenicol acetyltransferase (CAT). The cat-86 start codon is UUG, although related genes have AUG as the start codon. Changing the start codon to AUG increased expression of cat-86 by 36% in Bacillus subtilis. Changing the start codon to GUG and CUG decreased expression to 65% and 30%, respectively, of the level obtained when AUG was the start codon. CUG has not been previously shown to function as a start codon in B. subtilis. N-terminal sequencing of purified CAT protein specified by the CUG mutant, revealed that CUG was indeed the start codon and specified methionine. The gene xylE, which specifies catechol 2,3-dioxygenase, has AUG as its start codon. Changing the start codon for xylE to CUG decreased expression by 98%. However, when the ribosome-binding site sequence for xylE was optimized and the spacing between it and the start codon was increased to 8 nucleotides, xylE activity increased to 13% of the activity observed for AUG. CUG did not function efficiently as a start codon for cat-86 in Escherichia coli. These data suggest conditions under which CUG can function, with modest efficiency, as a start codon in B. subtilis.

Bacillus subtilis↗

Overexpression and purification of Pyrococcus abyssi phosphopantetheine adenylyltransferase from an optimized synthetic gene for NMR studies.

Phosphopantetheine adenylyltransferase (PPAT) is an essential enzyme that catalyses a rate-limiting step in coenzyme A (CoA) biosynthesis in all organisms. This study was conducted to obtain a high amount of pure, soluble, and stable PPAT from the hyperthermophilic archaeon Pyrococcus abyssi with the aim of investigating its structural characterization by NMR. Production of this enzyme from its natural gene in the Escherichia coli classical expression strain (BL21(DE3)) was not possible, most likely due to the presence of a high number of E. coli rare codons. Only a low amount of P. abyssi PPAT was previously obtained in two E. coli strains encoding tRNAs that recognize these rare E. coli codons and only by using a very rich growth medium. It was not possible to use this strategy to prepare labelled samples for the NMR study, thus another solution had to be found. Therefore, a synthetic gene encoding P. abyssi PPAT was constructed for which not only the rare codons were changed but which was also optimized to avoid other expression-limiting factors such as internal ribosome entry sites, RNA secondary structures, and DNA repeats. Gene optimization strongly increased the yield of P. abyssi PPAT in E. coli BL21(DE3) and allowed us to start the structural characterization of the enzyme. Circular dichroism and 2D NMR experiments indicate the presence of a well-ordered structure for P. abyssi PPAT and also confirm the existence of this enzyme as a monomer in solution.

Amino Acid Sequence↗

Selection at the wobble position of codons read by the same tRNA in Saccharomyces cerevisiae.

The transfer RNA gene complement of Saccharomyces cerevisiae was utilized for a whole-genome analysis of the deviation from a neutral usage of pyrimidine-ending cognate codons, that is, codons read by a single tRNA species having either inosine or guanosine as the first anticodon base. Mutational pressure at the wobble position was estimated from the base composition of the noncoding portion of the yeast genome. The selective pressure for translational efficiency was inferred from the degree of codon adaptation to tRNA gene redundancy and from mRNA abundance data derived from yeast transcriptome analysis. Amino acid conservation in orthologous comparisons with wholly sequenced microbial genomes was used to estimate translational accuracy requirements. A close correspondence was observed between the usage of wobble position pyrimidines and the frequency predicted by mutational bias. However, in the case of four cognate pairs (Gly: ggu/ggc; Asn: aau/aac; Phe: uuu/uuc; Tyr: uau/ uac) all read by guanosine-starting anticodons, we found evidence for a strong selective pressure driven by translational efficiency. Only for the glycine pair, wobble pyrimidine choice also appears to fulfill a translational accuracy requirement. Wobble pyrimidine selection is strictly related to the number of hydrogen bonds formed by alternative cognate codons: whenever a different number of hydrogen bonds can be formed at the wobble position, there is selection against six- or nine-hydrogen-bonded codon-anticodon pairs. Our results indicate that an intrinsic codon preference, critically dependent on the stability of codon-anticodon interaction and mainly reflecting selection for the optimization of translational efficiency, is built into the translational apparatus.

Codon↗

Biological significance of the U residue at the -3 position of the mRNA sequences of influenza A viral segments PB1 and NA.

The levels of viral proteins in infected cells are thought to be regulated by a variety of mechanisms. The initiation codons for the PB1 and NA proteins of A/WSN/33 (H1N1) influenza virus are in a suboptimal Kozak sequence for translation. To determine the significance of these suboptimal Kozak sequences, model vRNAs, whose coding regions were replaced with the reporter SEAP gene (for secreted alkaline phosphatase) and recombinant viruses with optimal Kozak sequences for PB1 and NA were constructed. Conversion of the upstream sequence of the PB1 and NA initiation codon to an optimal Kozak sequence was reflected in the level of reporter protein expression, but not the level of PB1 and NA protein expression. The recombinant viruses that had optimal Kozak sequences for PB1, NA, or both genes had similar replicative properties, both in cell culture and in mice, to those of the wild-type virus. These results suggest that expression of the PB1 and NA proteins is regulated by a mechanism other than that controlling the initiation of translation of these proteins.

Adenine↗

Analysis of leaky viral translation termination codons in vivo by transient expression of improved beta-glucuronidase vectors.

Plant RNA viruses commonly exploit leaky translation termination signals in order to express internal protein coding regions. As a first step to elucidate the mechanism(s) by which ribosomes bypass leaky stop codons in vivo, we have devised a system in which readthrough is coupled to the transient expression of beta-glucuronidase (GUS) in tobacco protoplasts. GUS vectors that contain the stop codons and surrounding nucleotides from the readthrough regions of several different RNA viruses were constructed and the plasmids were tested for the ability to direct transient GUS expression. These studies indicated that ribosomes bypass the leaky termination sites at efficiencies ranging from essentially 0 to ca. 5% depending upon the viral sequence. The results suggest that the efficiency of readthrough is determined by the sequence surrounding the stop codon. We describe improved GUS expression vectors and optimized transfection conditions which made it possible to assay low-level translational events.

Base Sequence↗

The extracellular portion of HLA-DR alpha chain is composed of two compactly folded domains.

A truncated form of the class II antigen DR alpha chain of the human major histocompatibility complex was produced in bacteria. A cDNA clone encoding the intact chain was modified so that the segment encoding the signal sequence was replaced by an ATG codon and the 3' region downstream to the part corresponding to the third exon was replaced by a stop codon. The new construct was put under the control of the Tac promoter in a bacterial expression vector. The distance between the Shine-Delgarno sequence and the initiation codon was randomized so that clones with optimal expression of the truncated DR alpha chain could be obtained after induced expression and immunoscreening. The truncated DR alpha chain was subjected to limited proteolysis with chymotrypsin, and the resulting cleavage products were analysed by sodium dodecyl sulphate-polyacrylamide gel electrophoresis. Two fragments were visualized by western blotting. Electrophoresis in the absence and presence of reducing agents suggested that one of the proteolytic fragments contained a disulphide bridge. It is concluded that the extracellular portion of the DR alpha chain is composed of two compactly folded domains connected by an extended stretch of the polypeptide chain.

Base Sequence↗

"Cold" single-strand conformational variants for mutation analysis of the RET protooncogene.

BACKGROUND: RET protooncogene mutation analysis is a routinely performed predictive DNA test in kindreds affected by multiple endocrine neoplasia (MEN) types 2A and 2B and familial medullary thyroid carcinoma (FMTC), and is a valuable diagnostic tool in newly diagnosed cases of medullary thyroid carcinoma (MTC). METHODS: We tested the suitability of the recently introduced "cold" single-strand conformational variant (SSCV) technique, which promises rapid, simple, nonradioactive detection of sequence variants in the identification of germline and somatic RET mutations. A total of 11 different mutations in exon 10 (codons 609, 611, 618, and 620) and 6 mutations in exon 11 (codon 634) were studied. RESULTS: Conditions were optimized so that conformational variants were demonstrated for all mutations examined in a single setting for exons 10 and 11. A novel six base pair (bp) inframe deletion between cysteines 630 and 634 was detected in a sporadic MTC. This adds to the evidence that not only cysteine deletions and substitutions but also changes in the spacing between cysteine residues have a pathogenic effect. CONCLUSIONS: Our results indicate that the cold SSCV method offers the advantages of simplicity, time savings, and nonradioactive detection for screening for RET sequence variants in hereditary and sporadic MTCs.

Amino Acid Sequence↗

The effect of queuosine on tRNA structure and function.

Computational modeling was performed to determine the potential function of the queuosine modification of tRNA found in wobble position 34 of tRNAasp, tRNAasn, tRNAhis, and tRNAtyr. Using the crystal structure of tRNAasp and a tRNA-tRNA-mRNA complex model, we show that the queuosine modification serves as a structurally restrictive base for tRNA anticodon loop flexibility. An extended intraresidue and intramolecular hydrogen bonding network is established by queuosine. The quaternary amine of the 7-aminomethyl side chain hydrogen bonds with the base's carbonyl oxygen. This positions the dihydroxycyclopentenediol ring of queuosine in proper orientation for hydrogen bonding with the backbone of the neighboring uridine 33 residue. The interresidue association stabilizes the formation of a cross-loop hydrogen bond between the uridine 33 base and the phosphoribosyl backbone of the cytosine at position 36. Additional interactions between RNAs in the translation complex were studied with regard to potential codon context and codon bias effects. Neither steric nor electrostatic interaction occurs between aminoacyl- and peptidyl-site tRNA anticodon loops that are modified with queuosine. However, there is a difference in the strength of anticodon/codon associations (codon bias) based on the presence or lack of queuosine in the wobble position of the tRNA. Unmodified (guanosine-containing) tRNAasp forms a very stable association with cytosine (GAC), but is much less stable in complex with a uridine-containing codon (GAU). Queuosine-modified tRNAasp exhibits no bias for either of cognate codons GAC or GAU and demonstrates a lower binding energy similar to the wobble pairing of guanosine-containing tRNA with a GAU codon. This is proposed to be due to the inflexibility of the queuosine-modified anticodon loop to accommodate proper positioning for optimal Watson-Crick type associations. A preliminary survey of codon usage patterns in oncodevelopmental versus housekeeping gene transcripts suggests a significant difference in bias for the queuosine-associated codons. Therefore, the queuosine modification may have the potential to influence cellular growth and differentiation by codon bias-based regulation of protein synthesis for discrete mRNA transcripts.

Anticodon↗

The downstream box: an efficient and independent translation initiation signal in Escherichia coli.

The downstream box (DB) was originally described as a translational enhancer of several Escherichia coli and bacteriophage mRNAs located just downstream of the initiation codon. Here, we introduced nucleotide substitutions into the DB and Shine-Dalgarno (SD) region of the highly active bacteriophage T7 gene 10 ribosome binding site (RBS) to examine the possibility that the DB has an independent and functionally important role. Eradication of the SD sequence in the absence of a DB abolished the translational activity of RBS fragments that were fused to a dihydrofolate reductase reporter gene. In contrast, an optimized DB at various positions downstream of the initiation codon promoted highly efficient protein synthesis despite the lack of a SD region. The DB was not functional when shifted upstream of the initiation codon to the position of the SD sequence. Nucleotides 1469-1483 of 16S rRNA ('anti-downstream box') are complementary to the DB, and optimizing this complementarity strongly enhanced translation in the absence and presence of a SD region. We propose that the stimulatory interaction between the DB and the anti-DB places the start codon in close contact with the decoding region of 16S rRNA, thereby mediating independent and efficient initiation of translation.

Bacteriophage T7↗

The positive relationship between codon usage bias and translation initiation AUG context in Saccharomyces cerevisiae.

The relationship between the codon usage bias and the sequence context surrounding the AUG translation initiation codon was examined in 211 Saccharomyces cerevisiae mRNA sequences. The codon usage bias and the number of matches to optimal AUG context, (A/U)A(A/C)AA(A/C)AUGUC(U/C), for translation initiation showed a positive relationship, indicating that these two factors are evolutionally under the similar natural selection constraint at the translation level. A new index (AUGCAI = AUG Context Adaptation Index) for the measure of optimal AUG context was devised, and the importance of each position of AUG context was also examined.

Base Sequence↗

Expression and characterization of a humanized cocaine-binding antibody.

The murine immunoglobulin G (IgG) cocaine-binding monoclonal antibody (mAb), GNC92H2, is notable for its exquisite specificity for cocaine, as opposed to chemically-related cocaine metabolites, and for its moderately high affinity (K(d) approximately 200 nM) for cocaine. Recently, we described the crystal structure of a mouse/human chimeric Fab construct at 2.3 A resolution. Herein, we report the successful framework humanization of a single-chain Fv (scFv) GNC92H2 construct without loss of affinity for cocaine. In brief, we compared the mAb GNC92H2 sequence to human antibody sequences, and used structure-based design to incorporate mutations (total = 49) that would humanize the framework region without affecting the overall shape of the binding pocket or the key cocaine-contact residues. The codons of the rationally designed sequence were optimized for E. coli expression, and the gene was synthesized by a de novo PCR reaction using 14 overlapping primers. Expression of the scFv construct was significantly improved in E. coli by fusion to thioredoxin. Intriguingly, this construct apparently refolds to form soluble active antibody in the reducing environment of the cytoplasm. Competitive ELISA and equilibrium dialysis demonstrated comparable binding activity between the humanized scFv and the whole IgG. The successful humanization of mAb GNC92H2 should enhance its potential therapeutic value by reducing its overall. immunogenicity.

Amino Acid Sequence↗

Improved expression, purification, and crystallization of p38alpha MAP kinase.

p38alpha mitogen-activated protein (MAP) kinase is widely expressed in many mammalian tissues and is activated as a part of signal transduction cascades that respond to inflammatory stimuli. The activation of p38 is known to trigger various biological effects, including cell death, differentiation, and proliferation. The central role played by p38alpha in cellular signaling events, including those that control a wide range of inflammatory and autoimmune diseases, makes it an attractive drug target. To develop optimized small molecule therapeutics targeting p38alpha, different techniques must be employed for the detailed biochemical, biophysical, and structural characterization of the interactions of p38alpha with lead compounds. These methods typically require large quantities of highly purified p38alpha protein. We describe here an improved expression and purification method for recombinant p38alpha production that reproducibly yields over 70 mg of highly purified protein per liter of shake flask bacterial culture. This yield is significantly higher than that previously reported for p38alpha production in Escherichia coli. We achieved a significant increase in soluble p38alpha protein expression by using the genetically modified E. coli strain BL21 DE3 Rosetta, which is optimized for expression of eukaryotic proteins with codons rarely used in E. coli. The p38alpha protein was purified to near homogeneity using a simple two-step procedure including nickel-chelating Sepharose chromatography followed by anion-exchange chromatography using MonoQ resin. Purified p38alpha was characterized using the standard commercially available small molecule inhibitor SB-203580. The binding association and dissociation rate constants determined by Biacore are in excellent agreement with previously reported values. The purified p38alpha protein was efficiently activated by MKK6 kinase to yield phosphorylated p38alpha. Purified p38alpha protein was also successfully crystallized, producing crystals diffracting to 1.9 angstroms, exceeding the highest resolution for p38alpha reported in the Protein DataBank. The simplicity and efficiency of this approach should prove useful for many laboratories that are interested in production of p38alpha for biochemical and biophysical studies and structure-based drug design.

Animals↗

Two rat surfactant protein A isoforms arise by a novel mechanism that includes alternative translation initiation.

A single gene for rat surfactant protein A (SP-A) encodes two isoforms that are distinguished by an isoleucine-lysine-cysteine (IKC) N-terminal extension (SP-A and IKC-SP-A). Available evidence suggests that the variants are generated by alternative signal peptidase cleavage of the nascent polypeptide at a primary site (Cys(-)(1)-Asn(1)) and a secondary site (Gly(-)(4)-Ile(-)(3)). In this study, we used site-directed mutagenesis and heterologous expression in vitro and in insect cells to the examine mechanisms that may lead to alternative signal peptidase cleavage including alternative translation initiation at two in-frame AUGs (Met(-)(30) and Met(-)(20)), a suboptimal context for hydrolysis at the primary cleavage site, or cotranslational protein modifications that expose an otherwise cryptic secondary cleavage site. In vitro translation of a rat cDNA for SP-A resulted in both 28 and 29 kDa primary translation products on SDS-PAGE analysis, while translation of cDNAs encoding Met-30Ala and Met-20Ala mutations resulted in only the single 28 and 29 kDa molecular mass species, respectively. These data are consistent with translation initiation at both Met(-)(30) and Met(-)(20) during in vitro synthesis of SP-A. The Met-30Ala mutation reduced expression of the longer isoform in insect cells, indicating that the Met(-)(30) site also contributes to eucaryotic protein expression. Forcing translation initiation at Met(-)(30) by optimizing the Kozak consensus sequence surrounding that codon or by mutating the Met(-)(20) codon resulted in preferential expression of the longer SP-A isoform but reduced overall expression of the protein almost 10-fold. Both isoforms were generated to some degree whether translation was initiated at the codon for Met(-)(30) or Met(-)(20), indicating that the site of translation initiation is not the sole determinant of isoform generation and suggesting that either the context of the primary cleavage site is suboptimal or that cotranslational modifications affect cleavage. Preventing N-terminal glycosylation at Asn(1) did not affect the site of signal peptidase cleavage. Disruption of interchain disulfide formation at Cys(-)(1) by substitution with serine markedly enhanced cleavage at the Gly(-)(4)-Ile(-)(3) bond, but substitution with alanine enhanced cleavage at the Cys(-)(1)-Asn(1) bond. We conclude that rat SP-A isoforms arise by a novel mechanism that includes both alternative translation initiation at two in-frame AUGs and a suboptimal context for signal peptidase hydrolysis at the primary cleavage site.

Amino Acid Sequence↗