Recognition of inscribed cursive Pashtu numeral through optimized deep learning

Autor:	Sibtain Syed, Khalil Khan, Maqbool Khan, Rehan Ullah Khan, Abdulrahman Aloraini
Jazyk:	angličtina
Rok vydání:	2024
Předmět:	Convolution neural networks Long short term memory Pashtu script Optical character recognition Pattern recognition Electronic computers. Computer science QA75.5-76.95
Zdroj:	PeerJ Computer Science, Vol 10, p e2124 (2024)
Druh dokumentu:	article
ISSN:	2376-5992
DOI:	10.7717/peerj-cs.2124
Popis:	Pashtu is one of the most widely spoken languages in south-east Asia. Pashtu Numerics recognition poses challenges due to its cursive nature. Despite this, employing a machine learning-based optical character recognition (OCR) model can be an effective way to tackle this issue. The main aim of the study is to propose an optimized machine learning model which can efficiently identify Pashtu numerics from 0–9. The methodology includes data organizing into different directories each representing labels. After that, the data is preprocessed i.e., images are resized to 32 × 32 images, then they are normalized by dividing their pixel value by 255, and the data is reshaped for model input. The dataset was split in the ratio of 80:20. After this, optimized hyperparameters were selected for LSTM and CNN models with the help of trial-and-error technique. Models were evaluated by accuracy and loss graphs, classification report, and confusion matrix. The results indicate that the proposed LSTM model slightly outperforms the proposed CNN model with a macro-average of precision: 0.9877, recall: 0.9876, F1 score: 0.9876. Both models demonstrate remarkable performance in accurately recognizing Pashtu numerics, achieving an accuracy level of nearly 98%. Notably, the LSTM model exhibits a marginal advantage over the CNN model in this regard.
Databáze:	Directory of Open Access Journals
Externí odkaz:	https://doaj.org/article/25630a7d86b9418880aa1c49a224ad73 Zobrazit plný text záznamu View record in DOAJ