Selecting Subexpressions to Materialize at Datacenter Scale
Summary: Selects cross-job subexpressions to materialize at datacenter scale, formulating the problem as ILP/bipartite labeling. BIG SUBS, a distributed vertex-centric algorithm, handles tens of thousands of jobs and cuts machine-hours by up to 40% on SCOPE workloads. (summarized by gpt-5.6-luna on Jul 24 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Alekh Jindal (Microsoft)
- 2. Konstantinos Karanasos (Microsoft)
- 3. Sriram Rao (Microsoft)
- 4. Hiren Patel (Microsoft)
BibTeX Citation
@article{jindal_vldb18,
title = {{Selecting Subexpressions to Materialize at Datacenter Scale}},
author = {Jindal, Alekh and Karanasos, Konstantinos and Rao, Sriram and Patel, Hiren},
journal = {PVLDB},
series = {{VLDB} '18},
volume = {11},
number = {7},
pages = {800--812},
doi = {10.14778/3192965.3192971},
url = {https://doi.org/10.14778/3192965.3192971},
year = {2018}
}
Incoming Citations (Sorted by Pagerank)
Showing 36 of 36 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 27 of 27 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 1,036 | Parallel Subgraph Listing in a Large-Scale Graph | 2014 | SIGMOD |
| 2 | 10,558 | Characterizing Parallel Subgraph Matching Performance: A Systematic Study of Interactions, Scalability, and Enumeration | 2026 | VLDB |
| 3 | 10,050 | Towards the Scheduling of Vertex-constrained Multi Subgraph Matching Query | 2020 | SIGMOD |
| 4 | 3,821 | Distributed Subgraph Matching on Timely Dataflow | 2019 | VLDB |
| 5 | 6,194 | Incorporating Super-Operators in Big-Data Query Optimizers | 2020 | VLDB |
| 6 | 442 | Efficient Subgraph Matching on Billion Node Graphs | 2012 | VLDB |
| 7 | 10,204 | Beyond Maximum Common Subgraph: A Framework Maximizing Shared Computation for Multi-Query Subgraph Matching | 2026 | SIGMOD |
| 8 | 4,158 | HUGE: An Efficient and Scalable Subgraph Enumeration System | 2021 | SIGMOD |
| 9 | 1,246 | Distributed Evaluation of Subgraph Queries Using Worst-case Optimal Low-Memory Dataflows | 2018 | VLDB |
| 10 | 1,154 | Efficient Exploitation of Similar Subexpressions for Query Processing | 2007 | SIGMOD |