АЛГОРИТМ СБОРА ТЕКСТОВЫХ ДАННЫХ НА КАЗАХСКОМ ЯЗЫКЕ

D.R. Rakhimova; A.R. Satybaldiev

doi:10.51889/2020-2.1728-7901.45

Vol. 70 No. 2 (2020)

ALGORITHM COLLECTION OF TEXT DATA IN THE KAZAKH LANGUAGE

Published June 2020

243

269

D.R. Rakhimova⁺⁻

Al-Farabi Kazakh National University, Almaty

A.R. Satybaldiev ⁺⁻

Al-Farabi Kazakh National University, Almaty

DOI: 10.51889/2020-2.1728-7901.45

Abstract

This work is devoted to the creation of a system for the automatic collection and processing of open data in Kazakh from Internet resources, and bears practical significance in the tasks of collecting and analyzing text. The introduction substantiates the relevance of the chosen topic, a review of existing approaches, formulates the objectives of the study. We consider such a problem as the collection and primary processing of text data with subsequent analysis. Data collection is a priority, since open data from Internet resources is not structured and needs to be processed. The authors provide a system for processing web pages of Kazakh-language portals, and also gives practical application of this approach to real data of open resources using the created system. The approach of indexing documents using features is presented. The system will help structure open data from Internet resources, as well as analyze collected data. Practical results are presented

.pdf (Русский)

Keywords

algorithm data collection text analysis Kazakh language

Language

Русский

How to Cite

[1]

Rakhimova Д. and Satybaldiev .А. 2020. ALGORITHM COLLECTION OF TEXT DATA IN THE KAZAKH LANGUAGE . Bulletin of Abai KazNPU. Series of Physical and Mathematical sciences. 70, 2 (Jun. 2020), 283–289. DOI:https://doi.org/10.51889/2020-2.1728-7901.45.

ALGORITHM COLLECTION OF TEXT DATA IN THE KAZAKH LANGUAGE

Download Citation