PPI-IRO: a Two-Stage Method for Protein-Protein Interaction Extraction Based on Interaction Relation Ontology.

Chuan-Xi Li,Peng Chen,Ru-Jing Wang,Xiu-Jie Wang,Ya-Ru Su,Jinyan Li
DOI: https://doi.org/10.1504/ijdmb.2014.062890
2014-01-01
International Journal of Data Mining and Bioinformatics
Abstract:Mining Protein-Protein Interactions (PPIs) from the fast-growing biomedical literature resources has been proven as an effective approach for the identification of biological regulatory networks. This paper presents a novel method based on the idea of Interaction Relation Ontology (IRO), which specifies and organises words of various proteins interaction relationships. Our method is a two-stage PPI extraction method. At first, IRO is applied in a binary classifier to determine whether sentences contain a relation or not. Then, IRO is taken to guide PPI extraction by building sentence dependency parse tree. Comprehensive and quantitative evaluations and detailed analyses are used to demonstrate the significant performance of IRO on relation sentences classification and PPI extraction. Our PPI extraction method yielded a recall of around 80% and 90% and an F1 of around 54% and 66% on corpora of AIMed and BioInfer, respectively, which are superior to most existing extraction methods.
What problem does this paper attempt to address?