DBScholar

Back to papers

Hyperspace: The Indexing Subsystem of Azure Synapse

Summary: Hyperspace is Synapse’s transparent secondary-indexing subsystem for lake files and warehouse tables, supporting multiple index types, concurrent maintenance, and automatic query rewrites. It delivers up to 10× benchmark and 100× real-workload acceleration without application changes. (summarized by gpt-5.6-luna on Jul 21 2026)

Paper ID
12698
Venue
VLDB
Year
2021
Pagerank
5.3483178e-05
Overall Rank
8,927 | 38.76%
DOI
10.14778/3476311.3476382

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{potharaju_vldb21,
        title = {{Hyperspace: The Indexing Subsystem of Azure Synapse}},
        author = {Potharaju, Rahul and Kim, Terry and Song, Eunjin and Wu, Wentao and Novik, Lev and Dave, Apoorve and Fogarty, Andrew and Pirzadeh, Pouria and Acharya, Vidip and Dhody, Gurleen and Li, Jiying and Ramanujam, Sinduja and Bruno, Nicolas and Galindo-Legaria, César A. and Narasayya, Vivek and Chaudhuri, Surajit and Nori, Anil K. and Talius, Tomas and Ramakrishnan, Raghu},
        journal = {PVLDB},
        series = {{VLDB} '21},
        volume = {14},
        number = {12},
        pages = {3043--3055},
        doi = {10.14778/3476311.3476382},
        url = {https://doi.org/10.14778/3476311.3476382},
        year = {2021}
}

Incoming Citations (Sorted by Pagerank)

Showing 1 of 1 citing papers.

Rank Citing Paper Year Venue Pagerank
10,866 LogCloud: Fast Search of Compressed Logs on Object Storage 2025 VLDB 5.093636e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 23 of 23 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.0020210012
12 C-Store: A Column-oriented DBMS 2005 VLDB 0.00069513174
24 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00054865648
32 Hive - A Warehousing Solution Over a Map-Reduce Framework 2009 VLDB 0.00050111008
104 Improved Query Performance with Variant Indexes 1997 SIGMOD 0.00033932213
156 An Efficient, Cost-Driven Index Selection Tool for Microsoft SQL Server 1997 VLDB 0.00028636811
227 Small Materialized Aggregates: A Light Weight Index Structure for Data Warehousing 1998 VLDB 0.00023958508
387 AutoAdmin "What-if" Index Analysis Utility 1998 SIGMOD 0.00019442332
520 Delta Lake: High-Performance ACID Table Storage over Cloud Object Stores 2020 VLDB 0.00017136828
735 Profiling, What-if Analysis, and Cost-based Optimization of MapReduce Programs 2011 VLDB 0.00014522606
1,481 Magic mirror in my hand, which is the best in the land? An Experimental Evaluation of Index Selection Algorithms 2020 VLDB 0.00010644613
1,765 Selecting Subexpressions to Materialize at Datacenter Scale 2018 VLDB 9.8079546e-05
1,925 POLARIS: The Distributed SQL Engine in Azure Synapse 2020 VLDB 9.4764398e-05
2,477 Azure Data Lake Store: A Hyperscale Distributed File Service for Big Data Analytics 2017 SIGMOD 8.5239378e-05
3,117 Chi: A Scalable and Programmable Control Plane for Distributed Stream Processing Systems 2018 VLDB 7.7382559e-05
3,137 Pushing Data-Induced Predicates Through Joins in Big-Data Clusters 2020 VLDB 7.7204167e-05
3,351 RHEEM: Enabling Cross-Platform Data Processing - May The Big Data Be With You! - 2018 VLDB 7.4937347e-05
3,605 Computation Reuse in Analytics Job Service at Microsoft 2018 SIGMOD 7.2640711e-05
3,964 Hyper Dimension Shuffle: Efficient Data Repartition at Petabyte Scale in SCOPE 2019 VLDB 6.9855158e-05
4,456 AdaptDB: Adaptive Partitioning for Distributed Joins 2017 VLDB 6.692321e-05
6,045 Helios: Hyperscale Indexing for the Cloud & Edge 2020 VLDB 5.9957544e-05
7,344 Building the Enterprise Fabric for Big Data with Vertica and Spark Integration 2016 SIGMOD 5.638371e-05
9,655 [Demo] Low-latency Spark Queries on Updatable Data 2019 SIGMOD 5.2422003e-05
Previous Page 1 / 1 Next

Semantically Similar Papers