Hadoop has become a key component of big data, and gained more and more support. Since users have recognized the enormous potential of Hadoop, some of them are working to develop and optimize the existing technologies to supplement Hadoop when using it. This paper gives the basic framework of Hadoop system and describes the optimization work on the parallel computing framework MapReduce, the performance of HDFS. We study and analyze the advantages and disadvantages of these technologies. Finally some future research directions are given.
Discussion(0)
No comments yet. Be the first to comment.