Combining probabilistic alignments with read pair information improves accuracy of split-alignments.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 29790902.
- Also identified by DOI 10.1093/bioinformatics/bty398 and PMC identifier 6198854.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Split-alignments provide base-pair-resolution evidence of genomic rearrangements. In practice, they are found by first computing high-scoring local alignments, parts of which are then combined into a split-alignment. This approach is challenging when aligning a short read to a large and repetitive reference, as it tends to produce many spurious local alignments leading to ambiguities in identifying the correct split-alignment. This problem is further exacerbated by the fact that rearrangements tend to occur in repeat-rich regions. We propose a split-alignment technique that combats the issue of ambiguous alignments by combining information from probabilistic alignment with positional information from paired-end reads. We demonstrate that our method finds accurate split-alignments, and that this translates into improved performance of variant-calling tools that rely on split-alignments. An open-source implementation is freely available at: https://bitbucket.org/splitpairedend/last-split-pe. Supplementary data are available at Bioinformatics online.
Medical subject headings
- Genomics
- Software