Recognition of inscribed cursive Pashtu numeral through optimized deep learning

Sibtain Syed; Khalil Khan; Maqbool Khan; Rehan Ullah Khan; Abdulrahman Aloraini

doi:10.7717/peerj-cs.2124

PeerJ Computer Science (Jul 2024)

Recognition of inscribed cursive Pashtu numeral through optimized deep learning

Sibtain Syed,
Khalil Khan,
Maqbool Khan,
Rehan Ullah Khan,
Abdulrahman Aloraini

Affiliations

Sibtain Syed: Department of IT & CS, Pak-Austria Fachhochschule Institute of Applied Sciences and Technology, Haripur, KP, Pakistan
Khalil Khan: Department of Computer Science, School of Engineering and Digital Sciences, Nazarbayev University, Astana, Kazakhstan
Maqbool Khan: Department of IT & CS, Pak-Austria Fachhochschule Institute of Applied Sciences and Technology, Haripur, KP, Pakistan
Rehan Ullah Khan: Department of Information Technology, College of Computer, Qassim University, Buraydah, Saudi Arabia
Abdulrahman Aloraini: Department of Information Technology, College of Computer, Qassim University, Buraydah, Saudi Arabia

DOI: https://doi.org/10.7717/peerj-cs.2124
Journal volume & issue: Vol. 10
p. e2124

Abstract

Read online Read online

Pashtu is one of the most widely spoken languages in south-east Asia. Pashtu Numerics recognition poses challenges due to its cursive nature. Despite this, employing a machine learning-based optical character recognition (OCR) model can be an effective way to tackle this issue. The main aim of the study is to propose an optimized machine learning model which can efficiently identify Pashtu numerics from 0–9. The methodology includes data organizing into different directories each representing labels. After that, the data is preprocessed i.e., images are resized to 32 × 32 images, then they are normalized by dividing their pixel value by 255, and the data is reshaped for model input. The dataset was split in the ratio of 80:20. After this, optimized hyperparameters were selected for LSTM and CNN models with the help of trial-and-error technique. Models were evaluated by accuracy and loss graphs, classification report, and confusion matrix. The results indicate that the proposed LSTM model slightly outperforms the proposed CNN model with a macro-average of precision: 0.9877, recall: 0.9876, F1 score: 0.9876. Both models demonstrate remarkable performance in accurately recognizing Pashtu numerics, achieving an accuracy level of nearly 98%. Notably, the LSTM model exhibits a marginal advantage over the CNN model in this regard.

Published in PeerJ Computer Science

ISSN: 2376-5992 (Online)
Publisher: PeerJ Inc.
Country of publisher: United States
LCC subjects: Science: Mathematics: Instruments and machines: Electronic computers. Computer science
Website: https://peerj.com/computer-science/

About the journal

Abstract

Keywords