Research on the Method of Constructing Distributed Data Lake Driven by Virtualization Model

Chuan Liu · 2020

In this paper, the method of constructing distributed data lake driven by virtualization model is analyzed and studied. Through the selection of rich information resources, and the definition of the data model, combined with a set of data specifications of the virtualization model, the edge computing technology has achieved the self-governance of integrated economic internal data. Comparison and virtualization model-driven method of building distributed Data Lake are with logical and physical dispersion characteristics, and practice the purpose” of data handling. It not only solves the problem that some comprehensive economic entities do not want to upload the original data federation of industry and analysis the business demand for large data, but also is very good to meet the need of real-time processing business for live data, reduce the amount of data handling costs at the same time, and finally improve the economy.

Read the paper · More papers on PaperTik