Approximate subgraph matching-based literature mining for biomedical events and relations.
Where this comes from
- Record sourced from PubMed, PMID 23613763.
- Also identified by DOI 10.1371/journal.pone.0060954 and PMC identifier 3629260.
- Licence recorded as CC0.
- The licence permits redistribution, so the abstract is shown in full and the full text is available from the publisher.
Abstract
The biomedical text mining community has focused on developing techniques to automatically extract important relations between biological components and semantic events involving genes or proteins from literature. In this paper, we propose a novel approach for mining relations and events in the biomedical literature using approximate subgraph matching. Extraction of such knowledge is performed by searching for an approximate subgraph isomorphism between key contextual dependencies and input sentence graphs. Our approach significantly increases the chance of retrieving relations or events encoded within complex dependency contexts by introducing error tolerance into the graph matching process, while maintaining the extraction precision at a high level. When evaluated on practical tasks, it achieves a 51.12% F-score in extracting nine types of biological events on the GE task of the BioNLP-ST 2011 and an 84.22% F-score in detecting protein-residue associations. The performance is comparable to the reported systems across these tasks, and thus demonstrates the generalizability of our proposed approach.
Medical subject headings
- Algorithms
- Biomedical Technology
- Data Mining
- Publications