[Demo] Low-latency Spark Queries on Updatable Data
Summary: Demo: Indexed DataFrame—cached Spark DataFrame with an integrated index for fast lookups and joins on updatable data. Supports multi-version concurrency for updates; evaluated on growing social-network graphs with microbenchmarks and real-world queries. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Alexandru Uta
- 2. Bogdan Ghit
- 3. Ankur Dave
- 4. Peter Boncz
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 8,754 | Hyperspace: The Indexing Subsystem of Azure Synapse | 2021 | VLDB | 4.4520434e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 3 of 3 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 66 | Spark SQL: Relational Data Processing in Spark | 2015 | SIGMOD | 0.00061707583 |
| 530 | The LDBC Social Network Benchmark: Interactive Workload | 2015 | SIGMOD | 0.00020823189 |
| 1,546 | Structured Streaming: A Declarative API for Real-Time Applications in Apache Spark | 2018 | SIGMOD | 0.00011418993 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 9,122 | Dynamic Speculative Optimizations for SQL Compilation in Apache Spark | 2020 | VLDB | 4.3877539e-05 |
| 4,056 | Are Updatable Learned Indexes Ready? | 2022 | VLDB | 6.4905689e-05 |
| 11,408 | SparkCAD: Caching Anomalies Detector for Spark Applications | 2022 | VLDB | 4.1905499e-05 |
| 1,546 | Structured Streaming: A Declarative API for Real-Time Applications in Apache Spark | 2018 | SIGMOD | 0.00011418993 |
| 3,207 | Big Data Analytics with Datalog Queries on Spark | 2016 | SIGMOD | 7.3847098e-05 |
| 4,751 | Indexing for Interactive Exploration of Big Data Series | 2014 | SIGMOD | 5.9411478e-05 |
| 11,199 | QaaD (Query-as-a-Data): Scalable Execution of Massive Number of Small Queries in Spark | 2023 | SIGMOD | 4.1905499e-05 |
| 3,536 | Scaling Spark in the Real World: Performance and Usability | 2015 | VLDB | 6.9938207e-05 |
| 4,646 | LocationSpark: A Distributed In-Memory Data Management System for Big Spatial Data | 2016 | VLDB | 6.0176549e-05 |
| 9,505 | Supporting Scalable Analytics with Latency Constraints | 2015 | VLDB | 4.3300131e-05 |