The Development of NLP Models

  • Zhang K
N/ACitations
Citations of this article
5Readers
Mendeley users who have this article in their library.

Abstract

This article reviews the development of Natural Language Processing (NLP) models from early statistical approaches to modern large language models (LLMs). Beginning with probability-based n-gram models, it outlines their limitations in data sparsity and long-term dependency. It then introduces neural network-based models, including word embeddings, logistic regression, multi-layer perceptrons, and Word2Vec, followed by sequential models such as RNNs and LSTMs that capture temporal dependencies. The shift to pre-trained models, marked by Word2Vec and the Transformer architecture, enabled scalable transfer learning and laid the foundation for state-of-the-art models like BERT, GPT, and their derivatives. Applications in text classification are illustrated through experiments with RoBERTa, hybrid BERT-LightGBM models, and fine-tuning techniques such as LoRA. Finally, the article discusses the broader ecosystem of large models, including chat models, multimodal models, and agent frameworks that integrate planning, memory, and tool use. The review emphasizes both theoretical principles and practical workflows, highlighting the necessity of iterative learning, coding practice, and ecosystem familiarity for effectively leveraging NLP technologies.

Cite

CITATION STYLE

APA

Zhang, K. (2025). The Development of NLP Models. Frontiers in Science and Engineering, 5(9), 127–141. https://doi.org/10.54691/enn7qh67

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free