Return
CoLoRd: compressing long reads
DOI:10.1038/s41592-022-01432-3.png)
Abstract
En 中文
The cost of maintaining exabytes of data produced by sequencing experiments every year has become a major issue in today's genomic research. In spite of the increasing popularity of third-generation sequencing, the existing algorithms for compressing long reads exhibit a minor advantage over the general-purpose gzip. We present CoLoRd, an algorithm able to reduce the size of third-generation sequencing data by an order of magnitude without affecting the accuracy of downstream analyses. CoLoRd achieves high compression rates for long-read sequencing data without affecting downstream analyses.
Keywords:
GENOME
Journal
IF:
32.1
Papers:
7.2K
Citations:
12.7W

