Unknown

Dataset Information

0

Data Integration for Microarrays: Enhanced Inference for Gene Regulatory Networks.


ABSTRACT: Microarray technologies have been the basis of numerous important findings regarding gene expression in the few last decades. Studies have generated large amounts of data describing various processes, which, due to the existence of public databases, are widely available for further analysis. Given their lower cost and higher maturity compared to newer sequencing technologies, these data continue to be produced, even though data quality has been the subject of some debate. However, given the large volume of data generated, integration can help overcome some issues related, e.g., to noise or reduced time resolution, while providing additional insight on features not directly addressed by sequencing methods. Here, we present an integration test case based on public Drosophila melanogaster datasets (gene expression, binding site affinities, known interactions). Using an evolutionary computation framework, we show how integration can enhance the ability to recover transcriptional gene regulatory networks from these data, as well as indicating which data types are more important for quantitative and qualitative network inference. Our results show a clear improvement in performance when multiple datasets are integrated, indicating that microarray data will remain a valuable and viable resource for some time to come.

SUBMITTER: Sirbu A 

PROVIDER: S-EPMC4996389 | biostudies-literature | 2015 May

REPOSITORIES: biostudies-literature

altmetric image

Publications

Data Integration for Microarrays: Enhanced Inference for Gene Regulatory Networks.

Sîrbu Alina A   Crane Martin M   Ruskin Heather J HJ  

Microarrays (Basel, Switzerland) 20150514 2


Microarray technologies have been the basis of numerous important findings regarding gene expression in the few last decades. Studies have generated large amounts of data describing various processes, which, due to the existence of public databases, are widely available for further analysis. Given their lower cost and higher maturity compared to newer sequencing technologies, these data continue to be produced, even though data quality has been the subject of some debate. However, given the larg  ...[more]

Similar Datasets

| S-EPMC3743784 | biostudies-literature
| S-EPMC3481449 | biostudies-literature
| S-EPMC7676633 | biostudies-literature
| S-EPMC5286517 | biostudies-literature
| S-EPMC4245971 | biostudies-literature
| S-EPMC2266705 | biostudies-literature
| S-EPMC6956787 | biostudies-literature
| S-EPMC2788929 | biostudies-literature
| S-EPMC4122380 | biostudies-literature
| S-EPMC4006459 | biostudies-literature