Unknown

Dataset Information

0

RNA-seq mixology: designing realistic control experiments to compare protocols and analysis methods.


ABSTRACT: Carefully designed control experiments provide a gold standard for benchmarking different genomics research tools. A shortcoming of many gene expression control studies is that replication involves profiling the same reference RNA sample multiple times. This leads to low, pure technical noise that is atypical of regular studies. To achieve a more realistic noise structure, we generated a RNA-sequencing mixture experiment using two cell lines of the same cancer type. Variability was added by extracting RNA from independent cell cultures and degrading particular samples. The systematic gene expression changes induced by this design allowed benchmarking of different library preparation kits (standard poly-A versus total RNA with Ribozero depletion) and analysis pipelines. Data generated using the total RNA kit had more signal for introns and various RNA classes (ncRNA, snRNA, snoRNA) and less variability after degradation. For differential expression analysis, voom with quality weights marginally outperformed other popular methods, while for differential splicing, DEXSeq was simultaneously the most sensitive and the most inconsistent method. For sample deconvolution analysis, DeMix outperformed IsoPure convincingly. Our RNA-sequencing data set provides a valuable resource for benchmarking different protocols and data pre-processing workflows. The extra noise mimics routine lab experiments more closely, ensuring any conclusions are widely applicable.

SUBMITTER: Holik AZ 

PROVIDER: S-EPMC5389713 | biostudies-literature | 2017 Mar

REPOSITORIES: biostudies-literature

altmetric image

Publications

RNA-seq mixology: designing realistic control experiments to compare protocols and analysis methods.

Holik Aliaksei Z AZ   Law Charity W CW   Liu Ruijie R   Wang Zeya Z   Wang Wenyi W   Ahn Jaeil J   Asselin-Labat Marie-Liesse ML   Smyth Gordon K GK   Ritchie Matthew E ME  

Nucleic acids research 20170301 5


Carefully designed control experiments provide a gold standard for benchmarking different genomics research tools. A shortcoming of many gene expression control studies is that replication involves profiling the same reference RNA sample multiple times. This leads to low, pure technical noise that is atypical of regular studies. To achieve a more realistic noise structure, we generated a RNA-sequencing mixture experiment using two cell lines of the same cancer type. Variability was added by extr  ...[more]

Similar Datasets

2018-08-20 | GSE118767 | GEO
2019-12-19 | GSE142286 | GEO
2018-08-18 | GSE118706 | GEO
2018-07-25 | GSE117618 | GEO
2018-08-18 | GSE118704 | GEO
2018-08-19 | GSE117617 | GEO
2019-02-22 | GSE126908 | GEO
2018-08-19 | GSE117450 | GEO
2019-02-22 | GSE126906 | GEO
| PRJNA486769 | ENA