Learning HMMs for nucleotide sequences from amino acid alignments.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 25638811.
- Also identified by DOI 10.1093/bioinformatics/btv054.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Profile hidden Markov models (profile HMMs) are known to efficiently predict whether an amino acid (AA) sequence belongs to a specific protein family. Profile HMMs can also be used to search for protein domains in genome sequences. In this case, HMMs are typically learned from AA sequences and then used to search on the six-frame translation of nucleotide (NT) sequences. However, this approach demands additional processing of the original data and search results. Here, we propose an alternative and more direct method which converts an AA alignment into an NT one, after which an NT-based HMM is trained to be applied directly on a genome.
Medical subject headings
- Genomics
- Sequence Alignment
- Sequence Analysis, Protein