Publication Details
Issue: Vol 3, No 3 (2026)
Pages: 154-159
ISSN: 2997-3902
Abstract
This thesis analyzes the issues of creating an automatic classification model for machine-building terms in the Uzbek language based on a corpus. The study examines methods for the linguistic, semantic, and structural separation of terms. A corpus of texts on mechanical engineering has been formed, and statistical and rule-based approaches to the automatic identification of terms and their division into semantic groups have been described. As a result of the study, the possibilities of systematizing technical terminology in the Uzbek language and its application in the field of digital linguistics have been substantiated.
Keywords
corpus linguistics
mechanical engineering terms
terminology
automatic classification
NLP
technical vocabulary
semantic model
Uzbek language