Unknown

Dataset Information

0

Lighter: fast and memory-efficient sequencing error correction without counting.


ABSTRACT: Lighter is a fast, memory-efficient tool for correcting sequencing errors. Lighter avoids counting k-mers. Instead, it uses a pair of Bloom filters, one holding a sample of the input k-mers and the other holding k-mers likely to be correct. As long as the sampling fraction is adjusted in inverse proportion to the depth of sequencing, Bloom filter size can be held constant while maintaining near-constant accuracy. Lighter is parallelized, uses no secondary storage, and is both faster and more memory-efficient than competing approaches while achieving comparable accuracy.

SUBMITTER: Song L 

PROVIDER: S-EPMC4248469 | biostudies-literature | 2014

REPOSITORIES: biostudies-literature

altmetric image

Publications

Lighter: fast and memory-efficient sequencing error correction without counting.

Song Li L   Florea Liliana L   Langmead Ben B  

Genome biology 20140101 11


Lighter is a fast, memory-efficient tool for correcting sequencing errors. Lighter avoids counting k-mers. Instead, it uses a pair of Bloom filters, one holding a sample of the input k-mers and the other holding k-mers likely to be correct. As long as the sampling fraction is adjusted in inverse proportion to the depth of sequencing, Bloom filter size can be held constant while maintaining near-constant accuracy. Lighter is parallelized, uses no secondary storage, and is both faster and more mem  ...[more]

Similar Datasets

| S-EPMC6280799 | biostudies-other
| S-EPMC3879328 | biostudies-literature
| S-EPMC3382444 | biostudies-literature
| S-EPMC3664804 | biostudies-literature
| S-EPMC3228814 | biostudies-other
| S-EPMC8261727 | biostudies-literature
| S-EPMC5704532 | biostudies-literature
| S-EPMC3169665 | biostudies-literature
| S-EPMC4253826 | biostudies-other
| S-EPMC4403973 | biostudies-literature