Publication Details
Issue: Vol 6, No 4 (2025)
Pages: 880-886
ISSN: 2660-6828

Abstract

This study investigates the vocabulary of the work Shajara-i Tarākima through a statistical lens, focusing on parts of speech. The analysis is anchored in the importance of language as a communication tool that reflects the social, economic, and cultural status of its time. With the advancement of computational linguistics, this research employs statistical methods to analyze and compare the frequency of lexical categories, aiding in the understanding of language structure and style. Despite the relevance of historical works in shaping modern language, there is a lack of comprehensive statistical analyses of classic texts like Shajara-i Tarākima, particularly in terms of parts of speech and their semantic classifications. The lexicon of Shajara-i Tarākima was digitized and analyzed using Excel, categorizing 14,093 lexemes into various parts of speech, including nouns, verbs, adjectives, and other lexical categories. The frequency of each category was calculated, and a semantic classification was conducted to provide deeper insight into the vocabulary. The analysis revealed that nouns, particularly proper nouns such as anthroponyms, toponyms, and ethnonyms, formed the largest portion of the lexicon. Verbs and adjectives followed, with a smaller proportion of auxiliary words and conjunctions. Notably, the work exhibits historical and archaic lexical forms, which distinguish it from modern Uzbek. The study concludes that the vocabulary of Shajara-i Tarākima is rich in Turkic onomastic units, showing minimal grammatical deviation from the modern Uzbek language but containing archaic lexemes. This research contributes to understanding the evolution of the Uzbek language, offering a statistical framework for further linguistic studies and comparisons across historical texts. It also highlights the relevance of linguistic tools in studying classical works for cultural, historical, and literary analysis.

Keywords
Lexeme parts of speech words denoting a person and an object words denoting an action words denoting a quality words denoting number and quantity lexemes denoting reference