Рақамли технологияларнинг назарий ва амалий масалалари 8-jild 1-son (2025) · 85-94-betlar

Detection and correction of spelling errors in Uzbek texts based on machine learning algorithms

Ochilov, M.M., Narzullayev, O.O., Xolmatov, O.A.

Manbada oʻqish PDF

Annotatsiya

This study addresses the problem of detecting and correcting spelling errors in Uzbek texts. Due to the complex morphological structure and agglutinative nature of the Uzbek language, traditional spell-checking methods do not provide sufficient accuracy. Therefore, this research employs the Levenshtein distance algorithm to measure word similarity and utilizes neural network-based language models for contextual correction. KenLM (a statistical language model), LSTM (Long Short-Term Memory), and BiLSTM (Bidirectional LSTM) approaches were used as language models. A text corpus of 80 million words was collected and analyzed for model training. The test results indicate that the BiLSTM model achieved the highest accuracy (90.09%) in correcting spelling errors, while the LSTM model recorded 84.62% accuracy. The KenLM model demonstrated an accuracy of 62.21% as well. These findings highlight that deep learning models capable of contextual analysis can significantly improve the automatic detection and correction of spelling errors in the Uzbek language. Based on the study results, future research plans include the application of transformer models, the expansion of annotated corpora, and the development of models that consider various morphological characteristics of the Uzbek language.

O‘zbek tiliimlo xatolarini tuzatishtabiiy tilni qayta ishlash (NLP)Levenshteyn masofasitil modeliKenLMLSTMBiLSTMneyron tarmoqlarkontekstual tahlil

Metadata manbasi: jurnal OAI-PMH arxivi · Sindex toʻliq matnni saqlamaydi, manbaga havola beradi.

Iqtibos olish

APA 7
Ochilov, M.M., Narzullayev, O.O. & Xolmatov, O.A. (2025). Detection and correction of spelling errors in Uzbek texts based on machine learning algorithms. Рақамли технологияларнинг назарий ва амалий масалалари, 8(1), 85-94.
GOST R 7.0.5
Ochilov, M.M., Narzullayev, O.O., Xolmatov, O.A. Detection and correction of spelling errors in Uzbek texts based on machine learning algorithms // Рақамли технологияларнинг назарий ва амалий масалалари. 2025. Т. 8. № 1. С. 85-94.
BibTeX
@article{m.m.2025,
  author  = {Ochilov, M.M. and Narzullayev, O.O. and Xolmatov, O.A.},
  title   = {Detection and correction of spelling errors in Uzbek texts based on machine learning algorithms},
  journal = {Рақамли технологияларнинг назарий ва амалий масалалари},
  year    = {2025},
  volume  = {8},
  number  = {1},
  pages   = {85-94}
}
RIS
TY  - JOUR
AU  - Ochilov, M.M.
AU  - Narzullayev, O.O.
AU  - Xolmatov, O.A.
TI  - Detection and correction of spelling errors in Uzbek texts based on machine learning algorithms
JO  - Рақамли технологияларнинг назарий ва амалий масалалари
PY  - 2025
VL  - 8
IS  - 1
SP  - 85
EP  - 94
ER  -