The New Version of the Corpus Data Management System “Tugan Tel” Using Graph Knowledge Base
Damir Mukhamedshin, Ayrat Rafizovich Gatiatullin, Гильмуллин Ринат Абрекович · 2024
This article describes the experience of developing a new version of the corpus data management system “Tugan Tel” using a graph knowledge base. The authors describe the model of linguistic knowledge graph of Turkic languages TurkLang. Also, this article describes the process of selecting a graph DBMS based on an analysis of existing solutions for selection criteria. Further, the development of a graph knowledge base is described in detail, with query examples. The model of the linguistic knowledge graph TurkLang was first implemented programmatically and the article describes in detail the process of this implementation. The article also touches upon the development of the search functionality of the new version of the system. The new conceptual model of the linguistic knowledge graph significantly expands the search and research capabilities of the corpus data management system. In addition, the developed graph knowledge base can be extended by other pluggable knowledge graphs and vice versa can be connected to other graph knowledge bases.