Technology for Preparing a Dataset Based on Mediapipe for Translating Text to Uzbek Sign Language

Джураев, Д.Б.

Рақамли технологияларнинг назарий ва амалий масалалари · 2025-yil

Annotatsiya

This paper discusses the problem of creating a multimodal dataset based on MediaPipe technology for automatic translation of Uzbek sign language into text and speech. Due to the multifaceted nature of sign language  hand movements, facial expressions, body posture, and gaze direction  the technical, linguistic, and ethical aspects of dataset preparation are analyzed. Using MediaPipe, key points of the hands, face, and body are extracted, and experiments are conducted with GRU, LSTM, and BiLSTM models. Comparative analysis demonstrates the efficiency and resource-saving capability of the GRU model, while LSTM and BiLSTM show advantages in processing complex sequences. Thus, the multimodal dataset developed in this study provides a reliable foundation for the development of real-time sign language recognition systems.

Maqola ma’lumotlari
MualliflarДжураев, Д.Б.
JurnalРақамли технологияларнинг назарий ва амалий масалалари
Nashr sanasi2025-09-15
Jild8
Son3
Betlar82-93
TilRus
DOI10.62132/ijdt.v8i3.290

Kalit so‘zlar

жестовый язык, MediaPipe, мультимодальный датасет, RNN (GRU, LSTM, BiLSTM), компьютерное зрение, обработка естественного языка (NLP), искусственный интеллект, sign language, MediaPipe, multimodal dataset, RNN (GRU, LSTM, BiLSTM), computer vision, natural language processing (NLP), artificial intelligence

Ilmiy soha

Рақамли технологияларнинг назарий ва амалий масалалари jurnalidan boshqa maqolalar

Рақамли технологияларнинг назарий ва амалий масалалари — barcha maqolalar