DBScholar

Back to papers

Hyperspace: The Indexing Subsystem of Azure Synapse

Summary: Hyperspace is Synapse’s transparent secondary-indexing subsystem for lake files and warehouse tables, supporting multiple index types, concurrent maintenance, and automatic query rewrites. It delivers up to 10× benchmark and 100× real-workload acceleration without application changes. (summarized by gpt-5.6-luna on Jul 21 2026)

Paper ID
h76f0a7db331651a1
Venue
VLDB
Year
2021
Pagerank
5.2283159e-05
Overall Rank
9,088 | 38.90%
DOI
10.14778/3476311.3476382

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{potharaju_vldb21,
        title = {{Hyperspace: The Indexing Subsystem of Azure Synapse}},
        author = {Potharaju, Rahul and Kim, Terry and Song, Eunjin and Wu, Wentao and Novik, Lev and Dave, Apoorve and Fogarty, Andrew and Pirzadeh, Pouria and Acharya, Vidip and Dhody, Gurleen and Li, Jiying and Ramanujam, Sinduja and Bruno, Nicolas and Galindo-Legaria, César A. and Narasayya, Vivek and Chaudhuri, Surajit and Nori, Anil K. and Talius, Tomas and Ramakrishnan, Raghu},
        journal = {PVLDB},
        series = {{VLDB} '21},
        volume = {14},
        number = {12},
        pages = {3043--3055},
        doi = {10.14778/3476311.3476382},
        url = {https://doi.org/10.14778/3476311.3476382},
        year = {2021}
}

Incoming Citations (Sorted by Pagerank)

Showing 1 of 1 citing papers.

Rank Citing Paper Year Venue Pagerank
11,269 LogCloud: Fast Search of Compressed Logs on Object Storage 2025 VLDB 4.9793485e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 23 of 23 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.001992968
12 C-Store: A Column-oriented DBMS 2005 VLDB 0.00068998927
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00055406774
31 Hive - A Warehousing Solution Over a Map-Reduce Framework 2009 VLDB 0.00049839909
107 Improved Query Performance with Variant Indexes 1997 SIGMOD 0.00033460288
151 An Efficient, Cost-Driven Index Selection Tool for Microsoft SQL Server 1997 VLDB 0.00028672526
216 Small Materialized Aggregates: A Light Weight Index Structure for Data Warehousing 1998 VLDB 0.00024485024
378 AutoAdmin "What-if" Index Analysis Utility 1998 SIGMOD 0.00019549382
459 Delta Lake: High-Performance ACID Table Storage over Cloud Object Stores 2020 VLDB 0.00017856221
753 Profiling, What-if Analysis, and Cost-based Optimization of MapReduce Programs 2011 VLDB 0.00014237583
1,397 Magic mirror in my hand, which is the best in the land? An Experimental Evaluation of Index Selection Algorithms 2020 VLDB 0.00010789242
1,745 Selecting Subexpressions to Materialize at Datacenter Scale 2018 VLDB 9.7343818e-05
1,790 POLARIS: The Distributed SQL Engine in Azure Synapse 2020 VLDB 9.6272006e-05
2,446 Azure Data Lake Store: A Hyperscale Distributed File Service for Big Data Analytics 2017 SIGMOD 8.4547121e-05
3,033 Chi: A Scalable and Programmable Control Plane for Distributed Stream Processing Systems 2018 VLDB 7.7337998e-05
3,073 Pushing Data-Induced Predicates Through Joins in Big-Data Clusters 2020 VLDB 7.6777283e-05
3,402 RHEEM: Enabling Cross-Platform Data Processing - May The Big Data Be With You! - 2018 VLDB 7.3304477e-05
3,545 Computation Reuse in Analytics Job Service at Microsoft 2018 SIGMOD 7.2134803e-05
3,895 Hyper Dimension Shuffle: Efficient Data Repartition at Petabyte Scale in SCOPE 2019 VLDB 6.9393157e-05
4,533 AdaptDB: Adaptive Partitioning for Distributed Joins 2017 VLDB 6.5565658e-05
6,174 Helios: Hyperscale Indexing for the Cloud & Edge 2020 VLDB 5.8612829e-05
7,485 Building the Enterprise Fabric for Big Data with Vertica and Spark Integration 2016 SIGMOD 5.5126823e-05
9,832 [Demo] Low-latency Spark Queries on Updatable Data 2019 SIGMOD 5.1245795e-05
Previous Page 1 / 1 Next

Semantically Similar Papers