reGenotyper: Detecting mislabeled samples in genetic data.
Where this comes from
- Record sourced from PubMed, PMID 28192439.
- Also identified by DOI 10.1371/journal.pone.0171324 and PMC identifier 5305221.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
In high-throughput molecular profiling studies, genotype labels can be wrongly assigned at various experimental steps; the resulting mislabeled samples seriously reduce the power to detect the genetic basis of phenotypic variation. We have developed an approach to detect potential mislabeling, recover the "ideal" genotype and identify "best-matched" labels for mislabeled samples. On average, we identified 4% of samples as mislabeled in eight published datasets, highlighting the necessity of applying a "data cleaning" step before standard data analysis.
Medical subject headings
- Algorithms
- Computational Biology
- Gene Expression Profiling
- Polymorphism, Single Nucleotide
- Quantitative Trait Loci