Performance Evaluation of Read and Write Operations in Hadoop Distributed File System
Talluri Lakshmi Siva Rama Krishna, Thirumalaisamy Ragunathan, Sudheer Kumar Battula · 2014
Hadoop Distributed File System (HDFS) is the core component of Apache Hadoop project. In HDFS, the computation is carried out in the nodes where relevant data is stored. Hadoop also implemented a parallel computational paradigm named as Map-Reduce. In this paper, we have measured the performance of read and write operations in HDFS by considering small and large files. For performance evaluation, we have used a Hadoop cluster with five nodes. The results indicate that HDFS performs well for the files with the size greater than the default block size and performs poorly for the files with the size less than the default block size.