Unknown

Dataset Information

0

Genome-Wide Co-Expression Distributions as a Metric to Prioritize Genes of Functional Importance.


ABSTRACT: Genome-wide gene expression analysis are routinely used to gain a systems-level understanding of complex processes, including network connectivity. Network connectivity tends to be built on a small subset of extremely high co-expression signals that are deemed significant, but this overlooks the vast majority of pairwise signals. Here, we developed a computational pipeline to assign to every gene its pair-wise genome-wide co-expression distribution to one of 8 template distributions shapes varying between unimodal, bimodal, skewed, or symmetrical, representing different proportions of positive and negative correlations. We then used a hypergeometric test to determine if specific genes (regulators versus non-regulators) and properties (differentially expressed or not) are associated with a particular distribution shape. We applied our methodology to five publicly available RNA sequencing (RNA-seq) datasets from four organisms in different physiological conditions and tissues. Our results suggest that genes can be assigned consistently to pre-defined distribution shapes, regarding the enrichment of differential expression and regulatory genes, in situations involving contrasting phenotypes, time-series, or physiological baseline data. There is indeed a striking additional biological signal present in the genome-wide distribution of co-expression values which would be overlooked by currently adopted approaches. Our method can be applied to extract further information from transcriptomic data and help uncover the molecular mechanisms involved in the regulation of complex biological process and phenotypes.

SUBMITTER: Alexandre PA 

PROVIDER: S-EPMC7593939 | biostudies-literature | 2020 Oct

REPOSITORIES: biostudies-literature

altmetric image

Publications

Genome-Wide Co-Expression Distributions as a Metric to Prioritize Genes of Functional Importance.

Alexandre Pâmela A PA   Hudson Nicholas J NJ   Lehnert Sigrid A SA   Fortes Marina R S MRS   Naval-Sánchez Marina M   Nguyen Loan T LT   Porto-Neto Laercio R LR   Reverter Antonio A  

Genes 20201020 10


Genome-wide gene expression analysis are routinely used to gain a systems-level understanding of complex processes, including network connectivity. Network connectivity tends to be built on a small subset of extremely high co-expression signals that are deemed significant, but this overlooks the vast majority of pairwise signals. Here, we developed a computational pipeline to assign to every gene its pair-wise genome-wide co-expression distribution to one of 8 template distributions shapes varyi  ...[more]

Similar Datasets

| S-EPMC3916258 | biostudies-literature
| S-EPMC2603584 | biostudies-literature
| S-EPMC2892466 | biostudies-literature
| S-EPMC4496047 | biostudies-literature
| S-EPMC7229696 | biostudies-literature
2022-02-16 | PXD023558 | Pride
| S-EPMC8079172 | biostudies-literature
| S-EPMC3168369 | biostudies-literature
| S-EPMC4090166 | biostudies-literature
| S-EPMC9496853 | biostudies-literature