An efficient graph kernel method for non-coding RNA functional prediction.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 28475710.
- Also identified by DOI 10.1093/bioinformatics/btx295.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
The importance of RNA protein-coding gene regulation is by now well appreciated. Non-coding RNAs (ncRNAs) are known to regulate gene expression at practically every stage, ranging from chromatin packaging to mRNA translation. However the functional characterization of specific instances remains a challenging task in genome scale settings. For this reason, automatic annotation approaches are of interest. Existing computational methods are either efficient but non-accurate or they offer increased precision, but present scalability problems. In this article, we present a predictive system based on kernel methods, a type of machine learning algorithm grounded in statistical learning theory. We employ a flexible graph encoding to preserve multiple structural hypotheses and exploit recent advances in representation and model induction to scale to large data volumes. Experimental results on tens of thousands of ncRNA sequences available from the Rfam database indicate that we can not only improve upon state-of-the-art predictors, but also achieve speedups of several orders of magnitude. The code is available from http://www.bioinf.uni-freiburg.de/~costa/EDeN.tgz . f.costa@exeter.ac.uk.
Medical subject headings
- Computational Biology
- Molecular Sequence Annotation
- RNA, Untranslated
- Supervised Machine Learning