Search PubMedSearch

SEARCH · Search PubMed

Results for “codon usage”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14Linked to original sources

Complete nucleotide sequence and genetic organization of the bacteriocinogenic plasmid, pIP404, from Clostridium perfringens.

The complete nucleotide sequence of the bacteriocinogenic plasmid, pIP404, from Clostridium perfringens has been determined. The plasmid genome comprises 10,207 bp and has a dA + dT content of 75%. Functions have been tentatively assigned to 6 of the 10 open reading frames and an origin-like region of repeated sequence identified. The codon usage of this extremely dA + dT rich plasmid is highly unusual and displays a pronounced preference for codons with the lowest dG + dC content. Only one of the genes from pIP404 was expressed at a significant level in Escherichia coli, suggesting that the atypical codon usage could represent a major obstacle to heterologous gene expression.

Base Sequence

Preferential use of A- and U-rich codons for Mycoplasma capricolum ribosomal proteins S8 and L6.

The nucleotide sequence of the 1.3 kilobase-pair DNA segment, which contains the genes for ribosomal proteins S8 and L6, and a part of L18 of Mycoplasma capricolum, has been determined and compared with the corresponding sequence in Escherichia coli (Cerretti et al., Nucl. Acids Res. 11, 2599, 1983). Identities of the predicted amino acid sequences of S8 and L6 between the two organisms are 54% and 42%, respectively. The A + T content of the M. capricolum genes is 71%, which is much higher than that of E. coli (49%). Comparisons of codon usage between the two organisms have revealed that M. capricolum preferentially uses A- and U-rich codons. More than 90% of the codon third positions and 57% of the first positions in M. capricolum is either A or U, whereas E. coli uses A or U for the third and the first positions at a frequency of 51% and 36%, respectively. The biased choice of the A- and U-rich codons in this organism has been also observed in the codon replacements for conservative amino acid substitutions between M. capricolum and E. coli. These facts suggest that the codon usage of M. capricolum is strongly influenced by the high A + T content of the genome.

Adenine

The sequence of the chloroplast atpB gene and its flanking regions in Chlamydomonas reinhardtii.

The chloroplast (cp)-encoded CF1 ATPase beta-subunit gene (atpB) of Chlamydomonas reinhardtii and its flanking regions have been sequenced. The derived amino acid (aa) sequence is highly homologous to that of the beta-subunit gene in Escherichia coli, bovine heart mitochondria, and higher plant cp. In contrast to all other cp genomes, the CF1 epsilon subunit gene (atpE) does not lie at the 3' end of the atpB gene but maps to a position 92 kb away in the other single-copy region. Northern blots confirm that the beta subunit is not encoded as part of a dicistronic message as it is in higher plants. The region just upstream from the atpB gene in C. reinhardtii contains two small open reading frames (ORFs) and not the gene for the large subunit of ribulose-1,5-bisphosphate carboxylase/oxygenase as is found in cp genomes of higher plants. No transcripts for either ORF were detected, but the codon usage in these ORFs as well as in the atpB gene follows the unique pattern of codon usage previously seen in other cp genes in C. reinhardtii.

Amino Acid Sequence

The ribosomal protein S8 from Thermus thermophilus VK1. Sequencing of the gene, overexpression of the protein in Escherichia coli and interaction with rRNA.

The gene of the ribosomal protein S8 from Thermus thermophilus VK1 has been isolated from a genomic library by hybridization of an oligonucleotide coding for the N-terminal amino acid sequence of the protein, amplified by PCR and sequenced. Nucleotide sequence reveals an open reading frame coding for a protein of 138 amino acid residues (M(r) 15,839). The codon usage shows that 94% of the codons possess G or C in the third position, and agrees with the preferential usage of codons of high G+C content in the bacteria of the genus Thermus. The amino acid sequence of the protein shows 48% identity with the protein from Escherichia coli. Ribosomal protein S8 from T. thermophilus has been expressed in E. coli under the control of the T7 promoter and purified to homogeneity by heat treatment of the extract followed by cation-exchange chromatography. Conditions were defined in which T. thermophilus protein S8 binds specifically an homologous 16S rRNA fragment containing the putative S8 binding site with an apparent association constant of 5 x 10(7) M-1. The overexpressed protein binds the rRNA with the same affinity as that extracted from T. thermophilus, indicating that the thermophilic protein is correctly folded in E. coli. The specificity of this binding is dependent on the ionic strength. The protein S8 from T. thermophilus recognizes the E. coli rRNA binding sites as efficiently as the S8 protein from E. coli. This result agrees with sequence comparisons of the S8 binding site on the small subunit rRNA from E. coli and from T. thermophilus, showing strong similarities in the regions involved in the interaction. It suggests that the structural features responsible for the recognition are conserved in the mesophilic and thermophilic eubacteria, despite structural peculiarities in the thermophilic partners conferring thermostability.

Amino Acid Sequence

[Use of codons in plant lectins].

Codon usage in the coding region of mature lectins has been examined for 11 plant species (8 leguminoseae, 1 euphorbiaceae, 2 gramineae). The different legume lectins exhibit nearly the same codon usage pattern whereas the choice for the silent position of codons is non-random.

Codon

The Pseudomonas cepacia 249 chromosomal penicillinase is a member of the AmpC family of chromosomal beta-lactamases.

Pseudomonas cepacia 249 produces an inducible beta-lactamase with penicillinase activity. The nucleotide sequence of the penA gene, which encodes this beta-lactamase, was determined and found to include regions with a significant homology to the ampC-encoded beta-lactamases of members of the family Enterobacteriaceae and Pseudomonas aeruginosa. The predicted amino acid sequence of the PenA beta-lactamase contained 17 amino acids immediately preceding the putative active-site serine which were highly conserved among the enzymes of the AmpC family. Although the penA-coding sequence had a total GC content of 60%, the predicted codon usage was more characteristic of Escherichia coli ampC-encoded beta-lactamase, with 53% of the codons having G or C in the third position, in contrast to the values for the P. aeruginosa ampC (88.5%) or Pseudomonas cepacia (88 to 92%) metabolic genes. The inducible expression of penA can be regulated by the E. coli gene product AmpD. A putative P. cepacia AmpR homolog was associated with the positive regulation of both Enterobacter cloacae ampC and P. cepacia penA expression, as confirmed by gel retardation studies. The E. cloacae AmpR did not regulate penA expression. Thus, by homology studies, codon usage, and genetic analysis, the P. cepacia penA beta-lactamase appears to have been acquired from members of the family Enterobacteriaceae and belongs to the class C group of beta-lactamases.

Base Sequence

Ornithine decarboxylase and trypanothione reductase genes in Leishmania braziliensis guyanensis.

Ornithine decarboxylase and trypanothione reductase are the key enzymes in polyamine and trypanothione metabolism in kinetoplastids. Using a heterologous Trypanosoma brucei brucei probe for ornithine decarboxylase and a mixed synthetic probe of 29 oligonucleotides for trypanothione reductase, we have detected the putative genes for these enzymes by Southern blot hybridization using genomic DNA of Leishmania braziliensis guyanensis MHOM/SR/80/CUMC 1. The trypanothione reductase probe was constructed both from the conserved codon usage of the redox active site for other flavin oxidoreductases over a wide evolutionary scale, and the preferred codon usage for other genes in species of Leishmania.

Amino Acid Sequence

[Use of the hygromycin phosphotransferase gene as the dominant selective marker for Chlamydomonas reinhardtii transformation].

The hygromycin phosphotransferase gene (hpt) from E. coli under the control of the SV40 early promoter was used as a dominant selectable marker for transformation of Chlamydomonas reinhardtii. Cells were transformed by electroporation (pulse length, 2 ms, field strength, 1 kV/cm). The culture growth phase was a crucial parameter for transformation (optimal density approximately 10(6) cells/ml). It was possible to obtain approximately 10(3) Hyg-resistant colonies under these conditions. Foreign DNA integrated into the Chlamydomonas genome was maintained for at least 8 months but the Hyg-resistant phenotype of the transformed clones was unstable. The frequency of codon usage in the hpt gene was compared with the one in Chlamydomonas nuclear genes. It is supposed that highly biased codon usage in Chlamydomonas does not preclude expression. Advantages of this selection system for studying Chlamydomonas transformation by heterologous genes are discussed.

Animals

Contextual constraints on codon pair usage: structural and biological implications.

Complementary DNA sequence data of 278 protein coding genes from prokaryotic systems have been analysed at the level of near neighbour codon pairs. Our analysis points out that constraints exist even at the level of near neighbour codon pairs. These constraints are in addition to those which arise due to relative levels of tRNA. Codon pairs, which in the data base have different occurrence values from their expected values, neither have common secondary structure nor do have better stabilization due to high base stacking. Our study points out that there are strong interaction between constituent codons in these codon pairs. These strongly interacting codon pairs, we suggest, are involved in the formation of three dimensional structural elements of cDNA/mRNA and interact with ribosome and thus modulate translation.

Base Sequence

Evolution of tropomyosin functional domains: differential splicing and genomic constraints.

We have cloned and determined the nucleotide sequence of a complementary DNA (cDNA) encoded by a newly isolated human tropomyosin gene and expressed in liver. Using the least-square method of Fitch and Margoliash, we investigated the nucleotide divergences of this sequence and those published in the literature, which allowed us to clarify the classification and evolution of the tropomyosin genes expressed in vertebrates. Tropomyosin undergoes alternative splicing on three of its nine exons. Analysis of the exons not involved in differential splicing showed that the four human tropomyosin genes resulted from a duplication that probably occurred early, at the time of the amphibian radiation. The study of the sequences obtained from rat and chicken allowed a classification of these genes as one of the types identified for humans. The divergence of exons 6 and 9 indicates that functional pressure was exerted on these sequences, probably by an interaction with proteins in skeletal muscle and perhaps also in smooth muscle; such a constraint was not detected in the sequences obtained from nonmuscle cells. These results have led us to postulate the existence of a protein in smooth muscle that may be the counterpart of skeletal muscle troponin. We show that different kinds of functional pressure were exerted on a single gene, resulting in different evolutionary rates and different convergences in some regions of the same molecule. Codon usage analysis indicates that there is no strict relationship between tissue types (and hence the tRNA precursor pool) and codon usage. G + C content is characteristic of a gene and does not change significantly during evolution.(ABSTRACT TRUNCATED AT 250 WORDS)

Amino Acid Sequence

Horizontal transfer of accessory chromosomes in fungi - a regulated process for exchange of genetic material?

Horizontal transfer of entire chromosomes has been reported in several fungal pathogens, often significantly impacting the fitness of the recipient fungus. All documented instances of horizontal chromosome transfers (HCTs) showed a marked propensity for accessory chromosomes, consistently involving the transfer of an accessory chromosome while other chromosomes were seldom, if ever, co-transferred. The mechanisms underlying HCTs, as well as the factors regulating the specificity of HCTs for accessory chromosomes, remain unclear. In this perspective, we provide an overview of the observed propensity in reported cases of horizontal chromosome transfers. We hypothesize the existence of a signal that distinguishes mobile, i.e., horizontally transferred, accessory chromosomes from the rest of the donor genome. Recent findings in Metarhizium robertsii and Magnaporthe oryzae, suggest that a mobile accessory chromosome may contain putative histones and/or histone modifiers, which could generate such a signal. Based on this, we propose that mobile accessory chromosomes may encode the machinery required for their own horizontal transmission, implying that HCT could be a regulated process. Finally, we present evidence of substantial differences in codon usage bias between core and accessory chromosomes in 14 out of 19 analysed fungal species and strains. Such differences in codon usage bias could indicate past horizontal transfers of these accessory chromosomes. Interestingly, HCT was previously unknown for many of these species, suggesting that the horizontal transfer of accessory chromosomes may be more widespread than previously thought, and therefore an important factor in fungal genome evolution.

Gene Transfer, Horizontal

High-level expression of staphylococcal nuclease R gene in Escherichia coli.

Staphylococcal nuclease R, an analogue of nuclease A, was overproduced under the transcriptional control of the bacteriophage lambda PRPL promoters regulated by temperature sensitive repressors. The expression level reached 200-300 mg l-1 and showed little host dependence in different strains. The investigations of the recombinant nuclease R have revealed that the amino terminal formyl methionine residue of the nuclease is precisely processed, the protein consists of 155 amino acid residues. The experiment shows that the pBV221-DH5 alpha is a quite suitable vector-host system for high-level expression and precise processing of heterologous genes in Escherichia coli. The comparative studies between the codons used in the staphylococcal nuclease R gene and the optimal codon usage in E. coli indicate that high level expression of heterologous genes in E. coli may not always require a high degree of codon usage bias.

Amino Acid Sequence

Does the 'non-coding' strand code?

The hypothesis that DNA strands complementary to the coding strand contain in phase coding sequences has been investigated. Statistical analysis of the 50 genes of bacteriophage T7 shows no significant correlation between patterns of codon usage on the coding and non-coding strands. In Bacillus and yeast genes the correlation observed is not different from that expected with random synonymous codon usage, while a high correlation seen in 52 E. coli genes can be explained in terms of an excess of RNY codons. A deficiency of UUA, CUA and UCA codons (complementary to termination) seems to be restricted to the E. coli genes, and may be due to low abundance of the relevant cognate tRNA species. Thus the analysis shows that the non-coding strand has the properties expected of a sequence complementary to a coding strand, with no indications that it encodes, or may have encoded, proteins.

Bacillus

An estimate on the effect of point mutation and natural selection on the rate of amino acid replacement in proteins.

We outline a method for estimating quantitatively the influence of point mutations and selection on the frequencies of codons and amino acids. We show how the mutation rate, i.e., the rate of amino acid replacement due to point mutation, can be affected by the codon usage as well as by the rates of the involved base exchanges. A comparison of the mutation rates calculated from reliable values of codon usage and base exchange probabilities with those that would be expected on the basis of chance reveals a notable suppression of replacements leading to tryptophan, glutamate, lysine, and methionine, and particularly of those leading to the termination codons. If selection constraints are neglected and only mutations are taken into account, the best agreement between expected and observed frequencies of both codons and amino acids is obtained for alpha = 1.13-1.15, where (Formula: see text). The "selection values" of codons and amino acids derived by our method show a pattern that partially deviates from others in the literature. For example, the selection pressure on methionine and cysteine turns out to be much more pronounced than expected if only the discrepancies between their observed and expected occurrences in proteins are considered. To estimate to what extent randomly occurring amino acid replacements are accepted by selection, we constructed an "acceptability matrix" from the well-established matrix of accepted point mutations. On the basis of this matrix "acceptability values" of the amino acids can be defined that correlate with their selection values. We also examine the significance of mutations and selection of amino acids with respect to their physicochemical properties and functions in proteins. The conservatism of amino acid replacements with respect to certain properties such as polarity can be brought about by the mutational process alone, whereas the conservatism with respect to other relevant properties--among them all measures of bulkiness--obviously is the result of additional selectional constraints on the evolution of protein structures.

Amino Acid Sequence

Characterization of In0 of Pseudomonas aeruginosa plasmid pVS1, an ancestor of integrons of multiresistance plasmids and transposons of gram-negative bacteria.

Many multiresistance plasmids and transposons of gram-negative bacteria carry related DNA elements that appear to have evolved from a common ancestor by site-specific integration of discrete cassettes containing antibiotic resistance genes or sequences of unknown function. The site of integration is flanked by conserved segments coding for an integraselike protein and for sulfonamide resistance, respectively. These segments, together with the antibiotic resistance genes between them, have been termed integrons (H. W. Stokes and R. M. Hall, Mol. Microbiol. 3:1669-1683, 1989). We report here the characterization of an integron, In0, from Pseudomonas aeruginosa plasmid pVS1, which has an unoccupied integration site and hence may be an ancestor of more complex integrons. Codon usage of the integrase (int) and sulfonamide resistance (sul1) genes carried by this integron suggests a common origin. This contrasts with the codon usage of other antibiotic resistance genes that were presumably integrated later as cassettes during the evolution and spread of these DNA elements. We propose evolutionary schemes for (i) the genesis of the integrons by the site-specific integration of antibiotic resistance genes and (ii) the evolution of the integrons of multiresistance plasmids and transposons, in relation to the evolution of transposons related to Tn21.

Amino Acid Sequence

Overproduction from a cellulase gene with a high guanosine-plus-cytosine content in Escherichia coli.

A recombinant exoglucanase was expressed in Escherichia coli to a level that exceeded 20% of total cellular protein. To obtain this level of overproduction, the exoglucanase gene coding sequence was fused to a synthetic ribosome-binding site, an initiating ATG, and placed under the control of the leftward promoter of bacteriophage lambda contained on the runaway replication plasmid vector pCP3 (E. Remaut, H. Tsao, and W. Fiers, Gene 22:103-113, 1983). With the exception of an inserted asparagine adjacent to the initiating ATG, the highly expressed exoglucanase is identical to the native exoglucanase. The overproduced exoglucanase can be isolated easily in an enriched form as insoluble aggregates, and exoglucanase activity can be recovered by solubilization of the aggregates in 6 M urea or 5 M guanidine hydrochloride. Since the codon usage of the exoglucanase gene is so markedly different from that of E. coli genes, the overproduction of the exoglucanase in E. coli indicates that codon usage may not be a major barrier to heterospecific gene expression in this organism.

Actinomycetales

Fine structural features of the chloroplast genome: comparison of the sequenced chloroplast genomes.

The entire nucleotide sequences of the rice, tobacco and liverwort chloroplast genomes have been determined. We compared all the chloroplast genes, open reading frames and spacer regions in the plastid genomes of these three species in order to elucidate general structural features of the chloroplast genome. Analyses of homology, GC content and codon usage of the genes enabled us to classify them into two groups: photosynthesis genes and genetic system genes. Based on comparisons of homology, GC content and codon usage, unidentified ORFs can also be assigned to each of these groups such that it is possible to speculate about the functions of products which may be produced by these ORFs. The spacer regions and intron sequences were compared and found to have no obvious homology between rice and liverwort or between tobacco and liverwort.

Base Composition

Variation in G + C-content and codon choice: differences among synonymous codon groups in vertebrate genes.

The relationship between G + C-content and codon usage in genes of human, mus, rat, bovine and chicken nuclear genomes was investigated. Correlation and lineal regression analyses were carried out on plots that related the frequency of each codon within each synonymous codon group to the G + C-content of the coding sequence as a whole. Under GC pressure, in most of the quartet codon groups there is a preferential choice of the C-ending codon, except in leucine and valine codon groups where the choice of the G-ending codon is preferred. Among ducts, the choice of codons specifying phenylalanine and glutamate shows the strongest dependence on G + C-content. The relationship found between G + C-content and codon usage in these genomes correlate with taxonomic distance.

Animals