Unknown

Dataset Information

0

InteMAP: Integrated metagenomic assembly pipeline for NGS short reads.


ABSTRACT: Next-generation sequencing (NGS) has greatly facilitated metagenomic analysis but also raised new challenges for metagenomic DNA sequence assembly, owing to its high-throughput nature and extremely short reads generated by sequencers such as Illumina. To date, how to generate a high-quality draft assembly for metagenomic sequencing projects has not been fully addressed.We conducted a comprehensive assessment on state-of-the-art de novo assemblers and revealed that the performance of each assembler depends critically on the sequencing depth. To address this problem, we developed a pipeline named InteMAP to integrate three assemblers, ABySS, IDBA-UD and CABOG, which were found to complement each other in assembling metagenomic sequences. Making a decision of which assembling approaches to use according to the sequencing coverage estimation algorithm for each short read, the pipeline presents an automatic platform suitable to assemble real metagenomic NGS data with uneven coverage distribution of sequencing depth. By comparing the performance of InteMAP with current assemblers on both synthetic and real NGS metagenomic data, we demonstrated that InteMAP achieves better performance with a longer total contig length and higher contiguity, and contains more genes than others.We developed a de novo pipeline, named InteMAP, that integrates existing tools for metagenomics assembly. The pipeline outperforms previous assembly methods on metagenomic assembly by providing a longer total contig length, a higher contiguity and covering more genes. InteMAP, therefore, could potentially be a useful tool for the research community of metagenomics.

SUBMITTER: Lai B 

PROVIDER: S-EPMC4545859 | biostudies-literature | 2015 Aug

REPOSITORIES: biostudies-literature

altmetric image

Publications

InteMAP: Integrated metagenomic assembly pipeline for NGS short reads.

Lai Binbin B   Wang Fumeng F   Wang Xiaoqi X   Duan Liping L   Zhu Huaiqiu H  

BMC bioinformatics 20150807


<h4>Background</h4>Next-generation sequencing (NGS) has greatly facilitated metagenomic analysis but also raised new challenges for metagenomic DNA sequence assembly, owing to its high-throughput nature and extremely short reads generated by sequencers such as Illumina. To date, how to generate a high-quality draft assembly for metagenomic sequencing projects has not been fully addressed.<h4>Results</h4>We conducted a comprehensive assessment on state-of-the-art de novo assemblers and revealed t  ...[more]

Similar Datasets

| S-EPMC3537596 | biostudies-literature
| S-EPMC3018814 | biostudies-other
| S-EPMC3092772 | biostudies-literature
| S-EPMC10926707 | biostudies-literature
| S-EPMC9508831 | biostudies-literature
| S-EPMC7214025 | biostudies-literature
| S-EPMC3424124 | biostudies-literature
2023-10-14 | GSE215357 | GEO
2023-10-14 | GSE215355 | GEO
2021-07-26 | E-MTAB-9189 | biostudies-arrayexpress