Using model Driven Engineering to transform Big Data query languages to MapReduce jobs
Allae Erraissi · International Journal of Computing and Digital Systems · 2021
Big Data processing is done by using MapReduce which is a clustered data processing framework.As it is Composed of Map and Reduce functions, it distributes data processing tasks between different computers, and hence reduces the results in a single summary.Most data analysts prefer to use query languages like Pig and Hive to process Big Data, given the complexity of the MapReduce paradigm.In this paper, we shall propose an approach based on Model Engineering to transform requests written by Pig or Hive to MapReduce jobs thanks to the use of the ATL transformation language.Our proposal will allow us to easily obtain MapReduce programs from requests written in Pig or Hive.