Quantitative frame analysis and the annotation of GC-rich (and other) prokaryotic genomes. An application to Anaeromyxobacter dehalogenans.
Where this comes from
- Record sourced from PubMed, PMID 26048600.
- Also identified by DOI 10.1093/bioinformatics/btv339 and PMC identifier 4595893.
- Licence recorded as CC BY-NC.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Graphical representations of contrasts in GC usage among codon frame positions (frame analysis) provide evidence of genes missing from the annotations of prokaryotic genomes of high GC content but the qualitative approach of visual frame analysis prevents its applicability on a genomic scale. We developed two quantitative methods for the identification and statistical characterization in sequence regions of three-base periodicity (hits) associated with open reading frame structures. The methods were implemented in the N-Profile Analysis Computational Tool (NPACT), which highlights in graphical representations inconsistencies between newly identified ORFs and pre-existing annotations of coding-regions. We applied the NPACT procedures to two recently annotated strains of the deltaproteobacterium Anaeromyxobacter dehalogenans, identifying in both genomes numerous conserved ORFs not included in the published annotation of coding regions. NPACT is available as a web-based service and for download at http://genome.ufl.edu/npact. lucianob@ufl.edu Supplementary data are available at Bioinformatics online.
Medical subject headings
- Genome, Bacterial
- Genomics
- Molecular Sequence Annotation
- Myxococcales