sBERT: Parameter-Efficient Transformer-Based Deep Learning Model for Scientific Literature Classification

  • Ahanger M
  • Wani M
  • Palade V
N/ACitations
Citations of this article
24Readers
Mendeley users who have this article in their library.

Abstract

This paper introduces a parameter-efficient transformer-based model designed for scientific literature classification. By optimizing the transformer architecture, the proposed model significantly reduces memory usage, training time, inference time, and the carbon footprint associated with large language models. The proposed approach is evaluated against various deep learning models and demonstrates superior performance in classifying scientific literature. Comprehensive experiments conducted on datasets from Web of Science, ArXiv, Nature, Springer, and Wiley reveal that the proposed model’s multi-headed attention mechanism and enhanced embeddings contribute to its high accuracy and efficiency, making it a robust solution for text classification tasks.

Cite

CITATION STYLE

APA

Ahanger, M. M., Wani, M. A., & Palade, V. (2024). sBERT: Parameter-Efficient Transformer-Based Deep Learning Model for Scientific Literature Classification. Knowledge, 4(3), 397–421. https://doi.org/10.3390/knowledge4030022

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free