BioVAE: a pre-trained latent variable language model for biomedical text mining.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 34636886.
- Also identified by DOI 10.1093/bioinformatics/btab702 and PMC identifier 8756089.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Large-scale pre-trained language models (PLMs) have advanced state-of-the-art (SOTA) performance on various biomedical text mining tasks. The power of such PLMs can be combined with the advantages of deep generative models. These are examples of these combinations. However, they are trained only on general domain text, and biomedical models are still missing. In this work, we describe BioVAE, the first large-scale pre-trained latent variable language model for the biomedical domain, which uses the OPTIMUS framework to train on large volumes of biomedical text. The model shows SOTA performance on several biomedical text mining tasks when compared to existing publicly available biomedical PLMs. In addition, our model can generate more accurate biomedical sentences than the original OPTIMUS output. Our source code and pre-trained models are freely available: https://github.com/aistairc/BioVAE. Supplementary data are available at Bioinformatics online.
Medical subject headings
- Data Mining
- Language