Skip to main navigation Skip to search Skip to main content

Whisper: Read sorting allows robust mapping of DNA sequencing data

  • Silesian University of Technology
  • Lodz University of Technology

Research output: Contribution to journalArticlepeer-review

6 Citations (Scopus)

Abstract

Motivation: Mapping reads to a reference genome is often the first step in a sequencing data analysis pipeline. The reduction of sequencing costs implies a need for algorithms able to process increasing amounts of generated data in reasonable time. Results: We present Whisper, an accurate and high-performant mapping tool, based on the idea of sorting reads and then mapping them against suffix arrays for the reference genome and its reverse complement. Employing task and data parallelism as well as storing temporary data on disk result in superior time efficiency at reasonable memory requirements. Whisper excels at large NGS read collections, in particular Illumina reads with typical WGS coverage. The experiments with real data indicate that our solution works in about 15% of the time needed by the well-known BWA-MEM and Bowtie2 tools at a comparable accuracy, validated in a variant calling pipeline.

Original languageEnglish
Pages (from-to)2043-2050
Number of pages8
JournalBioinformatics
Volume35
Issue number12
DOIs
Publication statusPublished - 1 Jun 2019

ASJC Scopus subject areas

  • Statistics and Probability
  • Biochemistry
  • Molecular Biology
  • Computer Science Applications
  • Computational Theory and Mathematics
  • Computational Mathematics

Fingerprint

Dive into the research topics of 'Whisper: Read sorting allows robust mapping of DNA sequencing data'. Together they form a unique fingerprint.

Cite this