Two-pass alignment improves novel splice junction quantification.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 26519505.
- Also identified by DOI 10.1093/bioinformatics/btv642 and PMC identifier 5006238.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Discovery of novel splicing from RNA sequence data remains a critical and exciting focus of transcriptomics, but reduced alignment power impedes expression quantification of novel splice junctions. Here, we profile performance characteristics of two-pass alignment, which separates splice junction discovery from quantification. Per sample, across a variety of transcriptome sequencing datasets, two-pass alignment improved quantification of at least 94% of simulated novel splice junctions, and provided as much as 1.7-fold deeper median read depth over those splice junctions. We further demonstrate that two-pass alignment works by increasing alignment of reads to splice junctions by short lengths, and that potential alignment errors are readily identifiable by simple classification. Taken together, two-pass alignment promises to advance quantification and discovery of novel splicing events. arul@med.umich.edu, nesvi@med.umich.edu Two-pass alignment was implemented here as sequential alignment, genome indexing, and re-alignment steps with STAR. Full parameters are provided in Supplementary Table 2. Supplementary data are available at Bioinformatics online.
Medical subject headings
- RNA Splice Sites
- RNA Splicing
- Sequence Alignment