Lighter

fast and memory-efficient sequencing error correction without counting

Li Song, Liliana D Florea, Ben Langmead

Research output: Contribution to journalArticle

Abstract

Lighter is a fast, memory-efficient tool for correcting sequencing errors. Lighter avoids counting k-mers. Instead, it uses a pair of Bloom filters, one holding a sample of the input k-mers and the other holding k-mers likely to be correct. As long as the sampling fraction is adjusted in inverse proportion to the depth of sequencing, Bloom filter size can be held constant while maintaining near-constant accuracy. Lighter is parallelized, uses no secondary storage, and is both faster and more memory-efficient than competing approaches while achieving comparable accuracy.

Original languageEnglish (US)
Pages (from-to)509
Number of pages1
JournalGenome Biology
Volume15
Issue number11
StatePublished - 2014
Externally publishedYes

Fingerprint

error correction
algal bloom
filter
sampling

ASJC Scopus subject areas

  • Medicine(all)

Cite this

Lighter : fast and memory-efficient sequencing error correction without counting. / Song, Li; Florea, Liliana D; Langmead, Ben.

In: Genome Biology, Vol. 15, No. 11, 2014, p. 509.

Research output: Contribution to journalArticle

@article{2587a20327464b32b6a681a1fffaf449,
title = "Lighter: fast and memory-efficient sequencing error correction without counting",
abstract = "Lighter is a fast, memory-efficient tool for correcting sequencing errors. Lighter avoids counting k-mers. Instead, it uses a pair of Bloom filters, one holding a sample of the input k-mers and the other holding k-mers likely to be correct. As long as the sampling fraction is adjusted in inverse proportion to the depth of sequencing, Bloom filter size can be held constant while maintaining near-constant accuracy. Lighter is parallelized, uses no secondary storage, and is both faster and more memory-efficient than competing approaches while achieving comparable accuracy.",
author = "Li Song and Florea, {Liliana D} and Ben Langmead",
year = "2014",
language = "English (US)",
volume = "15",
pages = "509",
journal = "Genome Biology",
issn = "1474-7596",
publisher = "BioMed Central",
number = "11",

}

TY - JOUR

T1 - Lighter

T2 - fast and memory-efficient sequencing error correction without counting

AU - Song, Li

AU - Florea, Liliana D

AU - Langmead, Ben

PY - 2014

Y1 - 2014

N2 - Lighter is a fast, memory-efficient tool for correcting sequencing errors. Lighter avoids counting k-mers. Instead, it uses a pair of Bloom filters, one holding a sample of the input k-mers and the other holding k-mers likely to be correct. As long as the sampling fraction is adjusted in inverse proportion to the depth of sequencing, Bloom filter size can be held constant while maintaining near-constant accuracy. Lighter is parallelized, uses no secondary storage, and is both faster and more memory-efficient than competing approaches while achieving comparable accuracy.

AB - Lighter is a fast, memory-efficient tool for correcting sequencing errors. Lighter avoids counting k-mers. Instead, it uses a pair of Bloom filters, one holding a sample of the input k-mers and the other holding k-mers likely to be correct. As long as the sampling fraction is adjusted in inverse proportion to the depth of sequencing, Bloom filter size can be held constant while maintaining near-constant accuracy. Lighter is parallelized, uses no secondary storage, and is both faster and more memory-efficient than competing approaches while achieving comparable accuracy.

UR - http://www.scopus.com/inward/record.url?scp=84965186660&partnerID=8YFLogxK

UR - http://www.scopus.com/inward/citedby.url?scp=84965186660&partnerID=8YFLogxK

M3 - Article

VL - 15

SP - 509

JO - Genome Biology

JF - Genome Biology

SN - 1474-7596

IS - 11

ER -