Unknown

Dataset Information

0

Complementarity of assembly-first and mapping-first approaches for alternative splicing annotation and differential analysis from RNAseq data.


ABSTRACT: Genome-wide analyses estimate that more than 90% of multi exonic human genes produce at least two transcripts through alternative splicing (AS). Various bioinformatics methods are available to analyze AS from RNAseq data. Most methods start by mapping the reads to an annotated reference genome, but some start by a de novo assembly of the reads. In this paper, we present a systematic comparison of a mapping-first approach (FARLINE) and an assembly-first approach (KISSPLICE). We applied these methods to two independent RNAseq datasets and found that the predictions of the two pipelines overlapped (70% of exon skipping events were common), but with noticeable differences. The assembly-first approach allowed to find more novel variants, including novel unannotated exons and splice sites. It also predicted AS in recently duplicated genes. The mapping-first approach allowed to find more lowly expressed splicing variants, and splice variants overlapping repeats. This work demonstrates that annotating AS with a single approach leads to missing out a large number of candidates, many of which are differentially regulated across conditions and can be validated experimentally. We therefore advocate for the combined use of both mapping-first and assembly-first approaches for the annotation and differential analysis of AS from RNAseq datasets.

SUBMITTER: Benoit-Pilven C 

PROVIDER: S-EPMC5844962 | biostudies-literature | 2018 Mar

REPOSITORIES: biostudies-literature

altmetric image

Publications

Complementarity of assembly-first and mapping-first approaches for alternative splicing annotation and differential analysis from RNAseq data.

Benoit-Pilven Clara C   Marchet Camille C   Chautard Emilie E   Lima Leandro L   Lambert Marie-Pierre MP   Sacomoto Gustavo G   Rey Amandine A   Cologne Audric A   Terrone Sophie S   Dulaurier Louis L   Claude Jean-Baptiste JB   Bourgeois Cyril F CF   Auboeuf Didier D   Lacroix Vincent V  

Scientific reports 20180309 1


Genome-wide analyses estimate that more than 90% of multi exonic human genes produce at least two transcripts through alternative splicing (AS). Various bioinformatics methods are available to analyze AS from RNAseq data. Most methods start by mapping the reads to an annotated reference genome, but some start by a de novo assembly of the reads. In this paper, we present a systematic comparison of a mapping-first approach (FARLINE) and an assembly-first approach (KISSPLICE). We applied these meth  ...[more]

Similar Datasets

| S-EPMC8144374 | biostudies-literature
| S-EPMC4578894 | biostudies-literature
| S-EPMC5028328 | biostudies-literature
| S-EPMC2745633 | biostudies-literature
| S-EPMC2798830 | biostudies-literature
| S-EPMC5860559 | biostudies-literature
| S-EPMC2838070 | biostudies-literature
| S-EPMC4575812 | biostudies-literature
| S-EPMC2000902 | biostudies-literature
| S-EPMC1853102 | biostudies-literature