Recognition of inscribed cursive Pashtu numeral through optimized deep learning.

PeerJ Comput Sci

Department of Information Technology, College of Computer, Qassim University, Buraydah, Saudi Arabia.

Published: July 2024

Pashtu is one of the most widely spoken languages in south-east Asia. Pashtu Numerics recognition poses challenges due to its cursive nature. Despite this, employing a machine learning-based optical character recognition (OCR) model can be an effective way to tackle this issue. The main aim of the study is to propose an optimized machine learning model which can efficiently identify Pashtu numerics from 0-9. The methodology includes data organizing into different directories each representing labels. After that, the data is preprocessed , images are resized to 32 × 32 images, then they are normalized by dividing their pixel value by 255, and the data is reshaped for model input. The dataset was split in the ratio of 80:20. After this, optimized hyperparameters were selected for LSTM and CNN models with the help of trial-and-error technique. Models were evaluated by accuracy and loss graphs, classification report, and confusion matrix. The results indicate that the proposed LSTM model slightly outperforms the proposed CNN model with a macro-average of precision: 0.9877, recall: 0.9876, F1 score: 0.9876. Both models demonstrate remarkable performance in accurately recognizing Pashtu numerics, achieving an accuracy level of nearly 98%. Notably, the LSTM model exhibits a marginal advantage over the CNN model in this regard.

Download full-text PDF

Source
http://www.ncbi.nlm.nih.gov/pmc/articles/PMC11323096PMC
http://dx.doi.org/10.7717/peerj-cs.2124DOI Listing

Publication Analysis

Top Keywords

pashtu numerics
12
lstm model
8
cnn model
8
model
7
pashtu
5
recognition inscribed
4
inscribed cursive
4
cursive pashtu
4
pashtu numeral
4
numeral optimized
4

Similar Publications

Recognition of inscribed cursive Pashtu numeral through optimized deep learning.

PeerJ Comput Sci

July 2024

Department of Information Technology, College of Computer, Qassim University, Buraydah, Saudi Arabia.

Pashtu is one of the most widely spoken languages in south-east Asia. Pashtu Numerics recognition poses challenges due to its cursive nature. Despite this, employing a machine learning-based optical character recognition (OCR) model can be an effective way to tackle this issue.

View Article and Find Full Text PDF

Background: Providing medical care to newly arrived migrants presents multiple challenges. A major challenge is a lack of a common language in the absence of language interpretation services. We examine the multilingualism of German physicians and clinical psychotherapists providing ambulatory care.

View Article and Find Full Text PDF

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!