article Open access

Study on Big Data Frameworks

  • International Journal of Scientific Research in Science and Technology
  • Technoscience Academy
Research footprint

At a glance

Citations
2
References
8
Comments
0
Paper overview

Abstract

Big data analytics is becoming more and more popular every day as a tool for evaluating large volumes of data on demand. Apache Hadoop, Spark, Storm, and Flink are four of the most widely used big data processing frameworks. Although all four architectures support big data analysis, they vary in how they are used and the infrastructure that supports it. This paper defines a general collection of main performance metrics, which include Processing Time, CPU Use, Latency, Execution Time, Performance, Scalability, and Fault-tolerance, and contrasting the four big data architectures against these KPIs in a literature review. When compared to Apache Hadoop and Apache Storm frameworks for non-real-time results, Spark was found to be the winner over multiple KPIs, including processing time, CPU usage, Latency, Execution time, and Scalability. In terms of processing time, CPU consumption, latency, execution time, and performance, Flink surpassed Apache Spark and Apache Storm architectures.

Record transparency

Publication details

DOI
10.32628/ijsrst218475
OpenAlex
W3198287910
Document type
article
Language
EN
Source
International Journal of Scientific Research in Science and Technology
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.