CRNN model for text detection and classification from natural scenes

N/ACitations
Citations of this article
26Readers
Mendeley users who have this article in their library.

Abstract

In the emerging field of computer vision, text recognition in natural settings remains a significant challenge due to variables like font, text size, and background complexity. This study introduces a method focusing on the automatic detection and classification of cursive text in multiple languages: English, Hindi, Tamil, and Kannada using a deep convolutional recurrent neural network (CRNN). The architecture combines convolutional neural networks (CNN) and long short-term memory (LSTM) networks for effective spatial and temporal learning. We employed pre-trained CNN models like VGG-16 and ResNet-18 for feature extraction and evaluated their performance. The method outperformed existing techniques, achieving an accuracy of 95.0%, 96.3%, and 96.2% on ICDAR 2015, ICDAR 2017, and a custom dataset (PDT2023), respectively. The findings not only push the boundaries of text detection technology but also offer promising prospects for practical applications.

Cite

CITATION STYLE

APA

Prakash, P., Yeliyur Hanumanthaiah, S. K., & Mayigowda, S. B. (2024). CRNN model for text detection and classification from natural scenes. IAES International Journal of Artificial Intelligence, 13(1), 839–849. https://doi.org/10.11591/ijai.v13.i1.pp839-849

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free