On Spatial Joins in MapReduce

Ibrahim Sabek, Mohamed Mokbel

Research output: Chapter in Book/Report/Conference proceedingConference contribution

2 Citations (Scopus)

Abstract

This paper provides the first attempt for a full-fledged query optimizer for MapReduce-based spatial join algorithms. The optimizer develops its own taxonomy that covers almost all possible ways of doing a spatial join for any two input datasets. The optimizer comes in two flavors; cost-based and rule-based. Given two input data sets, the cost-based query optimizer evaluates the costs of all possible options in the developed taxonomy, and selects the one with the lowest cost. The rule-based query optimizer abstracts the developed cost models of the cost-based optimizer into a set of simple easy-to-check heuristic rules. Then, it applies its rules to select the lowest cost option. Both query optimizers are deployed and experimentally evaluated inside a widely used open-source MapReduce-based big spatial data system. Exhaustive experiments show that both query optimizers are always successful in taking the right decision for spatially joining any two datasets of up to 500GB each.

Original languageEnglish
Title of host publicationGIS
Subtitle of host publicationProceedings of the ACM International Symposium on Advances in Geographic Information Systems
PublisherAssociation for Computing Machinery
Volume2017-November
ISBN (Print)9781450354905
DOIs
Publication statusPublished - 7 Nov 2017
Event25th ACM SIGSPATIAL International Conference on Advances in Geographic Information Systems, ACM SIGSPATIAL GIS 2017 - Redondo Beach, United States
Duration: 7 Nov 201710 Nov 2017

Other

Other25th ACM SIGSPATIAL International Conference on Advances in Geographic Information Systems, ACM SIGSPATIAL GIS 2017
CountryUnited States
CityRedondo Beach
Period7/11/1710/11/17

    Fingerprint

Keywords

  • Hadoop
  • MapReduce
  • Query Optimization
  • Spatial Join

ASJC Scopus subject areas

  • Earth-Surface Processes
  • Computer Science Applications
  • Modelling and Simulation
  • Computer Graphics and Computer-Aided Design
  • Information Systems

Cite this

Sabek, I., & Mokbel, M. (2017). On Spatial Joins in MapReduce. In GIS: Proceedings of the ACM International Symposium on Advances in Geographic Information Systems (Vol. 2017-November). [21] Association for Computing Machinery. https://doi.org/10.1145/3139958.3139967