Skip to main content Skip to main navigation menu Skip to site footer
Bulletin of Abai KazNPU. Series of Physical and mathematical sciences

ALGORITHM COLLECTION OF TEXT DATA IN THE KAZAKH LANGUAGE

Published June 2020

161

192

Al-Farabi Kazakh National University, Almaty
Al-Farabi Kazakh National University, Almaty
Abstract

This work is devoted to the creation of a system for the automatic collection and processing of open data in Kazakh from Internet resources, and bears practical significance in the tasks of collecting and analyzing text. The introduction substantiates the relevance of the chosen topic, a review of existing approaches, formulates the objectives of the study. We consider such a problem as the collection and primary processing of text data with subsequent analysis. Data collection is a priority, since open data from Internet resources is not structured and needs to be processed. The authors provide a system for processing web pages of Kazakh-language portals, and also gives practical application of this approach to real data of open resources using the created system. The approach of indexing documents using features is presented. The system will help structure open data from Internet resources, as well as analyze collected data. Practical results are presented

.pdf (Русский)
Language

Русский

How to Cite

[1]
Рахимова, Д. and Сатыбалдиев, .А. 2020. ALGORITHM COLLECTION OF TEXT DATA IN THE KAZAKH LANGUAGE . Bulletin of Abai KazNPU. Series of Physical and mathematical sciences. 70, 2 (Jun. 2020), 283–289. DOI:https://doi.org/10.51889/2020-2.1728-7901.45.