arrow
Return

CoLoRd: compressing long reads

delete2022-03-28
delete7
delete
OA
AI
M
Marek Kokot
A
Adam Gudyś
H
Heng Li *
D
Deorowicz, Sebastian *
DOI:10.1038/s41592-022-01432-3delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The cost of maintaining exabytes of data produced by sequencing experiments every year has become a major issue in today's genomic research. In spite of the increasing popularity of third-generation sequencing, the existing algorithms for compressing long reads exhibit a minor advantage over the general-purpose gzip. We present CoLoRd, an algorithm able to reduce the size of third-generation sequencing data by an order of magnitude without affecting the accuracy of downstream analyses. CoLoRd achieves high compression rates for long-read sequencing data without affecting downstream analyses.
Keywords:
GENOME

Journal

Nature Methods cover
Nature Methods
IF:
32.1
Papers:
7.2K
Citations:
12.7W

Organization

H
Harvard University
Scholars:
26.5W
Papers: 22.0W
Citations: 28.7W
S
Silesian University of Technology
Scholars:
6.2K
Papers: 6.2K
Citations: 5.9K