A sequence labeling framework for extracting drug-protein relations from biomedical literature

Ling Luo; Po Ting Lai; Chih Hsuan Wei; Zhiyong Lu

Journal ArticleOPEN ACCESS

A sequence labeling framework for extracting drug-protein relations from biomedical literature

Database : the journal of biological databases and curation (2022) 2022

DOI: 10.1093/database/baac058

4Citations

17Readers

Abstract

Automatic extracting interactions between chemical compound/drug and gene/protein are significantly beneficial to drug discovery, drug repurposing, drug design and biomedical knowledge graph construction. To promote the development of the relation extraction between drug and protein, the BioCreative VII challenge organized the DrugProt track. This paper describes the approach we developed for this task. In addition to the conventional text classification framework that has been widely used in relation extraction tasks, we propose a sequence labeling framework to drug-protein relation extraction. We first comprehensively compared the cutting-edge biomedical pre-trained language models for both frameworks. Then, we explored several ensemble methods to further improve the final performance. In the evaluation of the challenge, our best submission (i.e. the ensemble of models in two frameworks via major voting) achieved the F1-score of 0.795 on the official test set. Further, we realized the sequence labeling framework is more efficient and achieves better performance than the text classification framework. Finally, our ensemble of the sequence labeling models with majority voting achieves the best F1-score of 0.800 on the test set. DATABASE URL: https://github.com/lingluodlut/BioCreativeVII_DrugProt.

Cite

CITATION STYLE

APA

Luo, L., Lai, P. T., Wei, C. H., & Lu, Z. (2022). A sequence labeling framework for extracting drug-protein relations from biomedical literature. Database : The Journal of Biological Databases and Curation, 2022. https://doi.org/10.1093/database/baac058

A sequence labeling framework for extracting drug-protein relations from biomedical literature

Abstract

Cite

Register to see more suggestions