Unknown

Dataset Information

0

Universal and taxon-specific trends in protein sequences as a function of age.


ABSTRACT: Extant protein-coding sequences span a huge range of ages, from those that emerged only recently to those present in the last universal common ancestor. Because evolution has had less time to act on young sequences, there might be 'phylostratigraphy' trends in any properties that evolve slowly with age. A long-term reduction in hydrophobicity and hydrophobic clustering was found in previous, taxonomically restricted studies. Here we perform integrated phylostratigraphy across 435 fully sequenced species, using sensitive HMM methods to detect protein domain homology. We find that the reduction in hydrophobic clustering is universal across lineages. However, only young animal domains have a tendency to have higher structural disorder. Among ancient domains, trends in amino acid composition reflect the order of recruitment into the genetic code, suggesting that the composition of the contemporary descendants of ancient sequences reflects amino acid availability during the earliest stages of life, when these sequences first emerged.

SUBMITTER: James JE 

PROVIDER: S-EPMC7819706 | biostudies-literature | 2021 Jan

REPOSITORIES: biostudies-literature

altmetric image

Publications

Universal and taxon-specific trends in protein sequences as a function of age.

James Jennifer E JE   Willis Sara M SM   Nelson Paul G PG   Weibel Catherine C   Kosinski Luke J LJ   Masel Joanna J  

eLife 20210108


Extant protein-coding sequences span a huge range of ages, from those that emerged only recently to those present in the last universal common ancestor. Because evolution has had less time to act on young sequences, there might be 'phylostratigraphy' trends in any properties that evolve slowly with age. A long-term reduction in hydrophobicity and hydrophobic clustering was found in previous, taxonomically restricted studies. Here we perform integrated phylostratigraphy across 435 fully sequenced  ...[more]

Similar Datasets

| S-EPMC5898727 | biostudies-literature
| S-EPMC5633385 | biostudies-literature
| S-EPMC9292372 | biostudies-literature
| S-EPMC2880847 | biostudies-literature
| S-EPMC4574749 | biostudies-literature
| S-EPMC8649550 | biostudies-literature
| S-EPMC6690939 | biostudies-literature
| S-EPMC8098674 | biostudies-literature
| S-EPMC8187186 | biostudies-literature
| S-EPMC8647222 | biostudies-literature