Detecting the impact of sequencing errors on SAGE data.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 11590101.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
SAGE data are obtained by sequencing short DNA tags. Due to the mistakes in DNA sequencing, SAGE data contain errors. We propose a new approach to identify tags whose abundance is biased by sequencing errors. This approach is based on a concept of neighbourhood: abundant tags can contaminate tags whose sequence is very close. The application of our approach reveals that moderately abundant tags can be generated by sequencing errors uniquely. It also allows for detecting correct rare tags. Software is available only to non-profit entities and for non-commercial purposes upon request.
Medical subject headings
- Gene Expression Profiling
- Gene Library
- Sequence Analysis, DNA