• https://doi.org/10.37591/ecft.v7i3.2606Copy DOI Icon

Big Data Tools: A Survey

  • Jan 14, 2021
  • Shwetha M Banakar +2 more
Show More
  • Abstract
  • Literature Map
  • Similar Papers
Abstract

Abstract Nowadays, a large volume of data is generated in the form of text, voice, video, images, and sound. It is a very challenging job to handle and to get processed these different types of data. It is a very laborious process to analyze big data by using traditional data processing applications. Due to huge scattered file systems, a big data analysis is a difficult task. So, to analyze big data, a number of tools and techniques are required, Hadoop, Apache Spark, MongoDB, Cassandra are the tools used to handle big data . Keywords: Apache Spark, Cassandra, Hadoop, MongoDB

Similar Papers
  • Research Article

A Hybrid Machine Learning Model for Predictive Analytics in Big Data Frameworks

  • Mar 05, 2025
  • AVE Trends in Intelligent Computing Systems
  • Anjan Kumar Reddy Ayyadapu
  • Book Chapter
  • Citations58

SPARQLGX: Efficient Distributed Evaluation of SPARQL with Apache Spark

  • Jan 01, 2016
  • Damien Graux +3
  • Research Article
  • Citations3

Big Data and Java are integrated with machine learning

  • Mar 29, 2024
  • International Journal of Multidisciplinary Sciences and Arts
  • Anis Ahmed Qazi +1
  • Research Article
  • Citations14

Performance evaluation of Map-reduce jar pig hive and spark with machine learning using big data

  • Aug 01, 2020
  • International Journal of Electrical and Computer Engineering (IJECE)
  • Santosh Jankatti +3
  • Research Article
  • Citations1

Distributed Systems and Big Data Analytics in Predictive Healthcare: Transforming Modern Medicine

  • Mar 25, 2025
  • International Journal of Scientific Research in Computer Science, Engineering and Information Technology
  • Shridhar Bhalekar
  • Conference Article
  • Citations20

Fuzzy Based Clustering Algorithms to Handle Big Data with Implementation on Apache Spark

  • Mar 01, 2016
  • Neha Bharill +2
  • Supplementary Content
  • Citations3

Performance evaluation of GPU- and cluster-computing for parallelization of compute-intensive tasks

  • Aug 06, 2021
  • International Journal of Web Information Systems
  • Alexander Döschl +2
  • Research Article
  • Citations1

DNA barcoding using particle swarm optimization on apache spark SQL case study: DNA of covid-19

  • Jan 01, 2021
  • International Journal of Nonlinear Analysis and Applications
  • Lala Septem Riza +3
  • Book Chapter
  • Citations2

Chapter 21 - Big Data Integration

  • Jan 01, 2013
  • Managing Data in Motion
  • April Reeve
  • Research Article

Big Data Analytics for Business Growth: Leveraging Large-Scale Data for Competitive Advantage

  • Dec 09, 2025
  • European Journal of Applied Science, Engineering and Technology
  • Syed Mohammed Walid Karim +1
  • PDF
  • Research Article
  • Citations3

Legal Governance of Brain Data Derived from Artificial Intelligence

  • Jun 02, 2021
  • Voices in Bioethics
  • Mahika Ahluwalia
  • Conference Article
  • Citations13

Review of Apriori based Frequent Itemset Mining Solutions on Big Data

  • Apr 01, 2020
  • Mohammad Javad Shayegan Fard +1
  • Conference Article
  • Citations24

SWAT

  • May 31, 2016
  • Max Grossman +1
  • Conference Article
  • Citations28

Apache Hadoop Yarn Parameter configuration Challenges and Optimization

  • Feb 01, 2015
  • Bhavin J Mathiya +1
  • Conference Article
  • Citations3

Perldoop2: A Big Data-Oriented Source-to-Source Perl-Java Compiler

  • Nov 01, 2017
  • Cesar Pineiro +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.