CoLoRd: compressing long reads.
other · Level V
Where this comes from
- Record sourced from PubMed, PMID 35347321.
- Also identified by DOI 10.1038/s41592-022-01432-3 and PMC identifier 9337911.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
The cost of maintaining exabytes of data produced by sequencing experiments every year has become a major issue in today's genomic research. In spite of the increasing popularity of third-generation sequencing, the existing algorithms for compressing long reads exhibit a minor advantage over the general-purpose gzip. We present CoLoRd, an algorithm able to reduce the size of third-generation sequencing data by an order of magnitude without affecting the accuracy of downstream analyses.
Medical subject headings
- Genomics
- High-Throughput Nucleotide Sequencing