Domain-Adaptive Fine-Tuning of BioMedBERT for Medical Text Classification

1Citations
Citations of this article
9Readers
Mendeley users who have this article in their library.

Abstract

Accurate classification of medical notes and texts is a critical task for improving biomedical information retrieval and decision support systems. In this study, we propose a hybrid deep learning model combining BioMedBERT with Cross-Attention and BiLSTM, aimed at enhancing the classification performance of disease-related abstracts across five categories. The proposed model was evaluated using a dataset comprising 14k annotated samples derived from scientific medical literature. The proposed architecture achieves a macro F1-score of 63.82, outperforming traditional methods such as sentence embedding models (SimCSE, SBERT), zero-shot entailment approaches, and BioBERT variants integrated with MLP classifiers. Findings show that while the model effectively distinguishes between categories such as neoplasms and cardiovascular diseases, challenges persist in classifying abstracts with overlapping semantics, particularly general pathological conditions. This research demonstrates the efficacy of combining domainspecific language models with sequence and attention mechanisms, proposing a viable method for scalable and interpretable biomedical text classification.

Cite

CITATION STYLE

APA

Buntoro, G. A., Putra, O. V., & Purnomo, M. H. (2025). Domain-Adaptive Fine-Tuning of BioMedBERT for Medical Text Classification. Engineering, Technology and Applied Science Research, 15(6), 28523–28529. https://doi.org/10.48084/etasr.13026

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free