CREATION OF ARTIFICIAL INTELLIGENCE-BASED CONTINUOUS SPEECH RECOGNITION SYSTEM FOR UZBEKISTAN: DESIGN OF CORPUS, ACOUSTIC MODEL AND LANGUAGE MODEL

Arzikulov, Husan

Techscience.uz - техника фанлари долзарб масалалри · 2026-yil

Annotatsiya

This article analyzes the scientific basis for creating a continuous speech recognition system based on artificial intelligence for the Uzbek language. The study systematically summarizes the stages of corpus formation, data preprocessing, acoustic modeling, language model construction, decoding and results evaluation based on the primary sources under investigation. Comparative analysis, structural synthesis and interpretation of published results were used as a methodological basis. The studied literature shows that the quality of speech recognition in the Uzbek language is determined, first of all, by the quality and size of the corpus, the preprocessing discipline, hybrid deep learning architectures and language models adapted to the agglutinative nature of the Uzbek language. Based on these results, a six-stage architecture is proposed, consisting of corpus formation, preprocessing, character separation, acoustic modeling, language model construction and decoding with error evaluation.

Maqola ma’lumotlari
MualliflarArzikulov, Husan
JurnalTechscience.uz - техника фанлари долзарб масалалри
Nashr sanasi2026-03-25
Jild4
Son3
Betlar25-31
TilO‘zbek
DOI10.47390/ts-v4i3y2026n04

Kalit so‘zlar

artificial intelligence; Uzbek language; automatic speech recognition; speech corpus; acoustic model; language model; deep learning; CTC-attention; WER; CER., sun’iy intellekt; o‘zbek tili; avtomatik nutqni tanish; nutq korpusi; akustik model; til modeli; chuqur o‘rganish; CTC - diqqat; WER; CER.

Ilmiy soha

Techscience.uz - техника фанлари долзарб масалалри jurnalidan boshqa maqolalar

Techscience.uz - техника фанлари долзарб масалалри — barcha maqolalar