Swarm v3: towards tera-scale amplicon clustering.
Where this comes from
- Record sourced from PubMed, PMID 34244702.
- Also identified by DOI 10.1093/bioinformatics/btab493 and PMC identifier 8696092.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Previously we presented swarm, an open-source amplicon clustering programme that produces fine-scale molecular operational taxonomic units (OTUs) that are free of arbitrary global clustering thresholds. Here, we present swarm v3 to address issues of contemporary datasets that are growing towards tera-byte sizes. When compared with previous swarm versions, swarm v3 has modernized C++ source code, reduced memory footprint by up to 50%, optimized CPU-usage and multithreading (more than 7 times faster with default parameters), and it has been extensively tested for its robustness and logic. Source code and binaries are available at https://github.com/torognes/swarm. Supplementary data are available at Bioinformatics online.
Medical subject headings
- Software