FusionLearn: a biomarker selection algorithm on cross-platform data.
other · Level V
Where this comes from
- Record sourced from PubMed, PMID 30918944.
- Also identified by DOI 10.1093/bioinformatics/btz223.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
In high dimensional genetic data analysis, the objective is to select important biomarkers which are involved in some biological processes, such as disease progression, immune response, etc. The experimental data are often collected from different platforms including microarray experiments and proteomic experiments. The conventional single-platform approach lacks the capability to learn from multiple platforms, and the resulted lists of biomarkers vary across different platforms. There is a great need to develop an algorithm which can aggregate information across platforms and provide a consolidated list of biomarkers across different platforms. In this paper, we introduce an R package FusionLearn, which implements a fusion learning algorithm to analyze cross-platform data. The consolidated list of biomarkers is selected by the technique of group penalization. We first apply the algorithm on a collection of breast cancer microarray experiments from the NCBI (National Centre for Biotechnology Information) microarray database and the resulted list of selected genes have higher classification accuracy rate across different datasets than the lists generated from each single dataset. Secondly, we use the software to analyze a combined microarray and proteomic dataset for the study of the growth phase versus the stationary phase in Streptomyces coelicolor. The selected biomarkers demonstrate consistent differential behavior across different platforms. R package: https://cran.r-project.org/package=FusionLearn.
Medical subject headings
- Gene Expression Profiling