Guest Editorial: Big Data Infrastructure I
Jinjun Chen, Honggang Wang · IEEE Transactions on Big Data · 2018
The papers in this special section focuses on Big Data infrastructure. These papers address Big Data Infrastructure with emerging computing platforms such as heterogeneous clouds, hybrid architectures. Data is becoming an increasingly decisive resource in modern societies, economies, and governmental organizations. Big Data is an emerging paradigm encompassing various kinds of complex and large scale information beyond the processing capability of conventional software and databases. Various technologies are being discussed to support the handling of big data such as massively parallel processing databases, scalable storage systems, cloud computing platforms, Hadoop and Spark. Due to the multisource, massive, heterogeneous, and dynamic characteristics of application data involved in a distributed environment, one of the most important characteristics of Big Data is to carry out computing on the petabyte (PB), even the exabyte (EB)-level data with a complex computing process. Therefore, large-scale scalable Big Data Infrastructure with corresponding programming language support and software models for efficient processing in distributed environments such as cloud is on demand.