Unknown

Dataset Information

0

Amino acid usage is asymmetrically biased in AT- and GC-rich microbial genomes.


ABSTRACT:

Introduction

Genomic base composition ranges from less than 25% AT to more than 85% AT in prokaryotes. Since only a small fraction of prokaryotic genomes is not protein coding even a minor change in genomic base composition will induce profound protein changes. We examined how amino acid and codon frequencies were distributed in over 2000 microbial genomes and how these distributions were affected by base compositional changes. In addition, we wanted to know how genome-wide amino acid usage was biased in the different genomes and how changes to base composition and mutations affected this bias. To carry this out, we used a Generalized Additive Mixed-effects Model (GAMM) to explore non-linear associations and strong data dependences in closely related microbes; principal component analysis (PCA) was used to examine genomic amino acid- and codon frequencies, while the concept of relative entropy was used to analyze genomic mutation rates.

Results

We found that genomic amino acid frequencies carried a stronger phylogenetic signal than codon frequencies, but that this signal was weak compared to that of genomic %AT. Further, in contrast to codon usage bias (CUB), amino acid usage bias (AAUB) was differently distributed in AT- and GC-rich genomes in the sense that AT-rich genomes did not prefer specific amino acids over others to the same extent as GC-rich genomes. AAUB was also associated with relative entropy; genomes with low AAUB contained more random mutations as a consequence of relaxed purifying selection than genomes with higher AAUB.

Conclusion

Genomic base composition has a substantial effect on both amino acid- and codon frequencies in bacterial genomes. While phylogeny influenced amino acid usage more in GC-rich genomes, AT-content was driving amino acid usage in AT-rich genomes. We found the GAMM model to be an excellent tool to analyze the genomic data used in this study.

SUBMITTER: Bohlin J 

PROVIDER: S-EPMC3724673 | biostudies-literature | 2013

REPOSITORIES: biostudies-literature

altmetric image

Publications

Amino acid usage is asymmetrically biased in AT- and GC-rich microbial genomes.

Bohlin Jon J   Brynildsrud Ola O   Vesth Tammi T   Skjerve Eystein E   Ussery David W DW  

PloS one 20130726 7


<h4>Introduction</h4>Genomic base composition ranges from less than 25% AT to more than 85% AT in prokaryotes. Since only a small fraction of prokaryotic genomes is not protein coding even a minor change in genomic base composition will induce profound protein changes. We examined how amino acid and codon frequencies were distributed in over 2000 microbial genomes and how these distributions were affected by base compositional changes. In addition, we wanted to know how genome-wide amino acid us  ...[more]

Similar Datasets

| S-EPMC3053387 | biostudies-literature
| S-EPMC5902838 | biostudies-other
| S-EPMC5054167 | biostudies-literature
| S-EPMC6292993 | biostudies-literature
| S-EPMC4177787 | biostudies-literature
| S-EPMC4450053 | biostudies-literature
| S-EPMC16270 | biostudies-literature
| S-EPMC6707462 | biostudies-literature
| S-EPMC3744432 | biostudies-literature
| S-EPMC6322265 | biostudies-literature