READSCAN: a fast and scalable pathogen discovery program with accurate genome relative abundance estimation.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 23193222.
- Also identified by DOI 10.1093/bioinformatics/bts684 and PMC identifier 3562070.
- Licence recorded as CC BY.
- The licence permits redistribution, so the abstract is shown in full and the full text is available from the publisher.
Abstract
READSCAN is a highly scalable parallel program to identify non-host sequences (of potential pathogen origin) and estimate their genome relative abundance in high-throughput sequence datasets. READSCAN accurately classified human and viral sequences on a 20.1 million reads simulated dataset in <27 min using a small Beowulf compute cluster with 16 nodes (Supplementary Material). http://cbrc.kaust.edu.sa/readscan.
Medical subject headings
- Genome, Viral
- High-Throughput Nucleotide Sequencing
- Software