Back to papers
Templating Shuffles
Summary: TeShu: a unified, extensible shuffle service that represents optimization choices as parameterized shuffle templates capturing application, workload, and data‑center variability. Templates are instantiated via efficient sampling to quickly find near‑optimal shuffles.
(summarized by gpt-5-mini on Feb 09 2026)
Paper ID
heb09b2135a608d03
Venue
CIDR
Year
2023
Pagerank
4.9793485e-05
Overall Rank
11,676 | 21.50%
DOI
-
PDF
Download
(CC BY 4.0)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
BibTeX Citation
Copy BibTeX
@inproceedings{zhang_cidr23,
address = {Amsterdam, Netherlands},
series = {{CIDR} '23},
title = {{Templating Shuffles}},
booktitle = {Proceedings of the {Conference} on {Innovative} {Data} {Systems} {Research}},
author = {Zhang, Qizhen and Wu, Jiacheng and Chen, Ang and Liu, Vincent and Loo, Boon Thau},
year = {2023}
}
Incoming Citations (Sorted by Pagerank)
Showing 0 of 0 citing papers.
Rank
Citing Paper
Year
Venue
Pagerank
Outgoing Citations (Sorted by Pagerank)
Showing 18 of 18 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Rank
Cited Paper
Year
Venue
Pagerank
3
Pregel: A System for Large-Scale Graph Processing
2010
SIGMOD
0.0012092602
23
Spark SQL: Relational Data Processing in Spark
2015
SIGMOD
0.00055406774
231
Storm @Twitter
2014
SIGMOD
0.00023841089
394
One Trillion Edges: Graph Processing at Facebook-Scale
2015
VLDB
0.00019191286
415
SystemML: Declarative Machine Learning on Spark
2016
VLDB
0.0001865959
932
Starling: A Scalable Query Engine on Cloud Functions
2020
SIGMOD
0.00013008551
1,206
NUMA-aware algorithms: the case of data shuffling
2013
CIDR
0.00011541214
1,763
Lambada: Interactive Data Analytics on Cold Data Using Serverless Cloud Infrastructure
2020
SIGMOD
9.7003359e-05
2,350
Understanding the Effect of Data Center Resource Disaggregation on Production DBMSs
2020
VLDB
8.5984201e-05
2,635
Big Data Analytics with Datalog Queries on Spark
2016
SIGMOD
8.1965216e-05
3,135
Rethinking Data Management Systems for Disaggregated Data Centers
2020
CIDR
7.6092468e-05
3,608
DFI: The Data Flow Interface for High-Speed Networks
2021
SIGMOD
7.1650501e-05
4,403
Boxer: Data Analytics on Network-enabled Serverless Platforms
2021
CIDR
6.6138772e-05
4,533
AdaptDB: Adaptive Partitioning for Distributed Joins
2017
VLDB
6.5565658e-05
5,958
Cheetah: Accelerating Database Queries with Switch Pruning
2020
SIGMOD
5.9338314e-05
8,485
Optimizing Declarative Graph Queries at Large Scale
2019
SIGMOD
5.3326349e-05
8,550
Topology-aware Parallel Data Processing: Models, Algorithms and Systems at Scale
2020
CIDR
5.3168829e-05
9,816
Supporting Scalable Analytics with Latency Constraints
2015
VLDB
5.1257999e-05
Semantically Similar Papers
#
Overall Rank
Paper
Year
Venue
1
5,306
A Software-Defined Networking based Approach for Performance Management of Analytical Queries on Distributed Data Stores
2014
SIGMOD
2
4,008
Automating Distributed Tiered Storage Management in Cluster Computing
2020
VLDB
3
5,726
Privacy Amplification via Shuffling: Unified, Simplified, and Tightened
2024
VLDB
4
5,293
Optimizing Data-intensive Systems in Disaggregated Data Centers with TELEPORT
2022
SIGMOD
5
5,504
Magnet: Push-based Shuffle Service for Large-scale Data Processing
2020
VLDB
6
8,657
Network Shuffling: Privacy Amplification via Random Walks
2022
SIGMOD
7
8,752
Zero-sided RDMA: Network-driven Data Shuffling for Disaggregated Heterogeneous Cloud DBMSs
2024
SIGMOD
8
8,601
Towards Resource Efficiency: Practical Insights into Large-Scale Spark Workloads at ByteDance
2024
VLDB
9
1,206
NUMA-aware algorithms: the case of data shuffling
2013
CIDR
10
3,895
Hyper Dimension Shuffle: Efficient Data Repartition at Petabyte Scale in SCOPE
2019
VLDB