An information-based network approach for protein classification.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 28350835.
- Also identified by DOI 10.1371/journal.pone.0174386 and PMC identifier 5370107.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Protein classification is one of the critical problems in bioinformatics. Early studies used geometric distances and polygenetic-tree to classify proteins. These methods use binary trees to present protein classification. In this paper, we propose a new protein classification method, whereby theories of information and networks are used to classify the multivariate relationships of proteins. In this study, protein universe is modeled as an undirected network, where proteins are classified according to their connections. Our method is unsupervised, multivariate, and alignment-free. It can be applied to the classification of both protein sequences and structures. Nine examples are used to demonstrate the efficiency of our new method.
Medical subject headings
- Algorithms
- Proteins
- Proteomics