Publication Details
Issue: Vol 3, No 3 (2026)
Pages: 154-159
ISSN: 2997-3902

Abstract

This thesis analyzes the issues of creating an automatic classification model for machine-building terms in the Uzbek language based on a corpus. The study examines methods for the linguistic, semantic, and structural separation of terms. A corpus of texts on mechanical engineering has been formed, and statistical and rule-based approaches to the automatic identification of terms and their division into semantic groups have been described. As a result of the study, the possibilities of systematizing technical terminology in the Uzbek language and its application in the field of digital linguistics have been substantiated.

Keywords
corpus linguistics mechanical engineering terms terminology automatic classification NLP technical vocabulary semantic model Uzbek language