READSCAN: a fast and scalable pathogen discovery program with accurate genome relative abundance estimation.

Naeem, Raeece; Rashid, Mamoon; Pain, Arnab · Bioinformatics · 2013

basic_science · Level V

Where this comes from

Abstract

READSCAN is a highly scalable parallel program to identify non-host sequences (of potential pathogen origin) and estimate their genome relative abundance in high-throughput sequence datasets. READSCAN accurately classified human and viral sequences on a 20.1 million reads simulated dataset in <27 min using a small Beowulf compute cluster with 16 nodes (Supplementary Material). http://cbrc.kaust.edu.sa/readscan.

Medical subject headings