2018Unpublished venueRequires access

Performance Analysis of Hadoop Distributed File System Writing File Process

Yunyue Xie, Abobaker Mohammed Qasem Farhan, Meihua Zhou

Open publisher page 1 citations

Abstract

With the Internet development, the data in the world have had a drastic increase in the past decade. Traditional IT architecture cannot meet the needs of saving and processing data, the emergence of Cloud Computing solved the problem. Hadoop is a distributed processing software architecture run on Cloud Computing platform, it can store and process big data. Hadoop Distributed File System (HDFS) and MapReduce are its two main core components, which implements distributed file storage and parallel task processing respectively. In this paper, we will model, analysis, and evaluate HDFS based on Performance Evaluation Process Algebra (PEPA).

About this research paper

What this paper is about

With the Internet development, the data in the world have had a drastic increase in the past decade. Traditional IT architecture cannot meet the needs of saving and processing data, the emergence of Cloud Computing solved the problem. Hadoop is a distributed processing software architecture run on Cloud Computing platform, it can store and process big data. Hadoop Distributed File System (HDFS) and MapReduce are its two main core components, which implements distributed file storage and parallel task processing respectively. In this paper, we will model, analysis, and evaluate HDFS based on Performance Evaluation Process Algebra (PEPA).

Why it matters

OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

With the Internet development, the data in the world have had a drastic increase in the past decade. Traditional IT architecture cannot meet the needs of saving and processing data, the emergence of Cloud Computing solved the problem. Hadoop is a distributed processing software architecture run on Cloud Computing platform, it can store and process big data. Hadoop Distributed File System (HDFS) and MapReduce are its two main core components, which implements distributed file storage and parallel task processing respectively. In this paper, we will model, analysis, and evaluate HDFS based on Performance Evaluation Process Algebra (PEPA).

Key concepts: Computer science, Distributed File System, Cloud computing, Operating system, Data-intensive computing, Big data, Task (project management), Process (computing)

Related papers

Back to paper searchBrowse research topicsOriginal source
Performance Analysis of Hadoop Distributed File System Writing File Process — Research Paper | ScholarLens