• Home
  • Search
  • Multi-Q: Multiple Queries Optimization Based on MapReduce in Cloud
  • Cite Icon3
  • https://doi.org/10.1109/.19Copy DOI Icon

Multi-Q: Multiple Queries Optimization Based on MapReduce in Cloud

  • Nov 20, 2014
  • Ding Ding +2 more
Show More
  • Abstract
  • Literature Map
  • Citations
  • Similar Papers
Abstract

With the explosion of data in the past decade, big data is becoming a research hotspot in the information field. Many cloud-based distributed data processing platforms have been proposed to provide efficient and cost effective solutions for big data query processing, such as Hadoop, Hive, Pig, etc. However, most of the current research works are focus on improving the performance of query processing based on the view of systematics while without considering the characteristics of queries themselves, such as the query similarity, which will cause large numbers of redundant computation, effect query execution efficiency, thus having an adverse impact on promotion of the multi-queries processing performance. To solve this problem, in this paper, we propose a Multi-queries optimization framework based on MapReduce-oriented cloud environment (Multi-Q), which utilizes the dependence between multiple queries to realize query results reuse. Firstly, a cluster-based partition algorithm called CPA has been exploited to conduct the logic partition of the search range of query workload. Secondly, a multi-queries reuse dependence graph (MRDG) construction method on the basis of the cluster-based partition results has been presented to depict the dependence between the multiple queries. Finally, a Multi-Q processing algorithm based on Multi-Q Reuse Dependence Graph has been put forward to achieve the query results reuse and improve the overall query processing performance. We evaluate our approach by deploying Multi-Q based on Hadoop in a real cloud environment, called SEU-Cloud, and conducting extensive experiments based on the standard TPC-H. The result verifies that compared with Hive, the performance of improvement is approximately 39.3% by using our Multi-Q.

Similar Papers
  • Conference Article
  • Citations2

Map Reduce Programming for Electronic Medical Records Data Analysis on Cloud Using Apache Hadoop, Hive and Sqoop

  • Aug 01, 2015
  • Sreekanth Rallapalli +1
  • Conference Article

Survey on Scientific Data Processing Using Hadoop MapReduce in Cloud Environments

  • May 11, 2012
  • Xiangming Kong
  • Research Article

Achieving Accountable MapReduce in cloud computing

  • Jan 01, 2014
  • Future Generation Computer Systems
  • Xiaozhifeng +1
  • Research Article

A Survey on Big Data Analytics Using HADOOP

  • Jun 05, 2019
  • Asian Journal of Computer Science and Technology
  • S Mamatha +1
  • Conference Article
  • Citations47

Efficiently supporting multiple similarity queries for mining in metric databases

  • Feb 01, 2000
  • B Braunmuller +3
  • Book Chapter
  • Citations5

Efficient Group Processing for Multiple Reverse Top-k Geo-Social Keyword Queries

  • Jan 01, 2020
  • Pengfei Jin +3
  • Research Article
  • Citations3

MapReduce in the Cloud: Data-Location-Aware VM Scheduling

  • Dec 25, 2013
  • ZTE communications
  • Tung Nguyen And Weisong Shi
  • Book Chapter

Parallel and Distributed Query Processing in Attributed Networks

  • Jan 01, 2023
  • A Sandhya Rani +1
  • Research Article
  • Citations851

The state of the art in distributed query processing

  • Dec 01, 2000
  • ACM Computing Surveys
  • Donald Kossmann
  • Conference Article
  • Citations6

An efficient framework of data mining and its analytics on massive streams of big data repositories

  • Aug 01, 2016
  • D N Disha +3
  • Conference Article
  • Citations5

Optimal k-Nearest-Neighbor Query Processing via Multiple Lower Bound Approximations

  • Dec 01, 2018
  • Christian Beecks +1
  • Research Article
  • Citations6

E-Commerce Security Research in Big Data Environment

  • Jan 01, 2018
  • International Journal of Enterprise Information Systems
  • Mei Zhang +2
  • Research Article
  • Citations21

Materialized view selection using evolutionary algorithm for speeding up big data query processing

  • Mar 16, 2017
  • Journal of Intelligent Information Systems
  • Rajib Goswami +2
  • Conference Article
  • Citations4

Performance Enhancement in Big Data handling

  • Feb 01, 2020
  • Himadri Sekhar Ray +2
  • Book Chapter
  • Citations3

A Peer-to-Peer Architecture for Cloud Based Data Cubes Allocation

  • Nov 10, 2015
  • Mohammed Ezzat +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.