article Open access

Data Transfers in Hadoop: A Comparative Study

  • RonPub -- Research Online Publishing (RonPub UG)
Research footprint

At a glance

Citations
4
References
27
Comments
0
Paper overview

Abstract

Hadoop is an open source framework for processing large amounts of data in distributed computing environment. It plays an important role in processing and analyzing the Big Data. This framework is used for storing data on large clusters of commodity hardware. Data input and output to and from Hadoop is an indispensable action for any data processing job. At present, many tools have been evolved for importing and exporting Data in Hadoop. In this article, some commonly used tools for importing and exporting data have been emphasized. Moreover, a state-of-the-art comparative study among the various tools has been made. With this study, it has been decided that where to use one tool over the other with emphasis on the data transfer to and from Hadoop system. This article also discusses about how Hadoop handles backup and disaster recovery along with some open research questions in terms of Big Data transfer when dealing with cloud-based services.

Record transparency

Publication details

OpenAlex
W2277296357
Document type
article
Language
EN
Source
RonPub -- Research Online Publishing (RonPub UG)
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.