QARAQALPAQ TILINIŃ LINGVISTIKALÍQ KORPUS PLATFORMASÍN JARATÍW (Tólepbergen Qayıpbergenovtıń shıǵarmaları mısalında)

Authors

  • G.R. Abdalieva

    Tashkent informaciyalıq texnologiya universiteti Nókis filialı

  • N.U. Uteuliev

    Tashkent informaciyalıq texnologiya universiteti Nókis filialı

  • B.K. Kalmuratov

    Tashkent informaciyalıq texnologiya universiteti Nókis filialı

  • A.S. Utesinova

    Tashkent informaciyalıq texnologiya universiteti Nókis filialı

Keywords: karakalpak language, linguistic corpus, cultural heritage, morphological analysis, syntactic analysis, semantic analysis, lemmatization, tokenization, artificial intelligence, phraseology

Abstract

This article focuses on the process of creating a linguistic corpus based on the Karakalpak language. Using T.Kaipbergenov's works as an example, it highlights methods for systematic language analysis and the potential of information systems. The importance of natural language processing technologies and linguistic corpora is analyzed as an innovative approach in linguistics. The article outlines the stages of corpus creation, its technical aspects, and its scientific and practical significance.

References

1. McEnery T., & Hardie A. 2011. Corpus linguistics: Method, theory and practice. Cambridge University Press.

2. Sinclair J. 2004. Trust the text: Language, corpus and discourse. Routledge.

3. Dash N.S. 2008. Corpus Linguistics: An Introduction. Pearson Education India.