LocationSpark: A Distributed In-Memory Data Management System for Big Spatial Data
Summary: LocationSpark is a distributed in-memory spatial DB on Spark, offering range, kNN, spatio-textual queries, and spatial-join. Key ideas: skew-aware scheduling, plan-aware exec, spatial Bloom filters to cut cross-node traffic, and hot data flushing. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Mingjie Tang (Purdue University)
- 2. Yongyang Yu (Purdue University)
- 3. Qutaibah M. Malluhi (Qatar University)
- 4. Mourad Ouzzani (Hamad Bin Khalifa University; Qatar Computing Research Institute)
- 5. Walid G. Aref (Purdue University)
BibTeX Citation
@article{tang_vldb16,
title = {{LocationSpark: A Distributed In-Memory Data Management System for Big Spatial Data}},
author = {Tang, Mingjie and Yu, Yongyang and Malluhi, Qutaibah M. and Ouzzani, Mourad and Aref, Walid G.},
journal = {PVLDB},
series = {{VLDB} '16},
volume = {9},
number = {13},
pages = {1565--1576},
doi = {10.14778/3007263.3007276},
url = {https://doi.org/10.14778/3007263.3007276},
year = {2016}
}
Incoming Citations (Sorted by Pagerank)
Showing 8 of 8 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,732 | How Good Are Modern Spatial Analytics Systems? | 2018 | VLDB | 8.1934602e-05 |
| 5,714 | Hu-Fu: Efficient and Secure Spatial Queries over Data Federation | 2022 | VLDB | 6.1117296e-05 |
| 7,155 | VRE: A Versatile, Robust, and Economical Trajectory Data System | 2022 | VLDB | 5.6869288e-05 |
| 8,057 | Architecting a Query Compiler for Spatial Workloads | 2020 | SIGMOD | 5.4982831e-05 |
| 9,544 | SwiftSpatial: Spatial Joins on Modern Hardware | 2025 | SIGMOD | 5.2528121e-05 |
| 11,392 | ST4ML: Machine Learning Oriented Spatio-Temporal Data Processing at Scale | 2023 | SIGMOD | 5.093636e-05 |
| 11,797 | SSTD: A Distributed System on Streaming Spatio-Textual Data | 2020 | VLDB | 5.093636e-05 |
| 12,008 | CarStream: An Industrial System of Big Data Processing for Internet-of-Vehicles | 2017 | VLDB | 5.093636e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 6 of 6 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 24 | Spark SQL: Relational Data Processing in Spark | 2015 | SIGMOD | 0.00054865648 |
| 1,067 | Hadoop-GIS: A High Performance Spatial Data Warehousing System over MapReduce | 2013 | VLDB | 0.00012327784 |
| 1,319 | SkewTune: Mitigating Skew in MapReduce Applications | 2012 | SIGMOD | 0.00011175005 |
| 2,137 | Efficient Processing of k Nearest Neighbor Joins using MapReduce | 2012 | VLDB | 9.110238e-05 |
| 4,740 | An Experimental Analysis of Iterated Spatial Joins in Main Memory | 2013 | VLDB | 6.5283833e-05 |
| 5,395 | AQWA: Adaptive Query-Workload-Aware Partitioning of Big Spatial Data | 2015 | VLDB | 6.2331619e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 2,046 | Distributed Trajectory Similarity Search | 2017 | VLDB |
| 2 | 5,395 | AQWA: Adaptive Query-Workload-Aware Partitioning of Big Spatial Data | 2015 | VLDB |
| 3 | 8,057 | Architecting a Query Compiler for Spatial Workloads | 2020 | SIGMOD |
| 4 | 8,644 | Incremental Partitioning for Efficient Spatial Data Analytics | 2022 | VLDB |
| 5 | 1,067 | Hadoop-GIS: A High Performance Spatial Data Warehousing System over MapReduce | 2013 | VLDB |
| 6 | 2,794 | A Demonstration of SpatialHadoop: An Efficient MapReduce Framework for Spatial Data | 2013 | VLDB |
| 7 | 8,990 | STAR: A Distributed Stream Warehouse System for Spatial Data | 2020 | SIGMOD |
| 8 | 1,441 | Efficient Processing of Top-k Spatial Preference Queries | 2011 | VLDB |
| 9 | 12,141 | A Demonstration of AQWA: Adaptive Query-Workload-Aware Partitioning of Big Spatial Data | 2015 | VLDB |
| 10 | 1,175 | Simba: Efficient In-Memory Spatial Analytics | 2016 | SIGMOD |