DBScholar

Back to papers

Helios: Hyperscale Indexing for the Cloud & Edge

Summary: Helios is a hyperscale cloud/edge indexing system that feeds real-time streams into relational engines. It ingests quadrillions of events and indexes trillions of keys daily across dozens of data centers, using a simple data model and asynchronous indexing - a scalable blueprint for Microsoft-scale analytics. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h3366fa01e8189050
Venue
VLDB
Year
2020
Pagerank
5.8612829e-05
Overall Rank
6,174 | 58.50%
DOI
10.14778/3415478.3415547

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{potharaju_vldb20,
        title = {{Helios: Hyperscale Indexing for the Cloud \& Edge}},
        author = {Potharaju, Rahul and Kim, Terry and Wu, Wentao and Acharya, Vidip and Suh, Steve and Fogarty, Andrew and Dave, Apoorve and Ramanujam, Sinduja and Talius, Tomas and Novik, Lev and Ramakrishnan, Raghu},
        journal = {PVLDB},
        series = {{VLDB} '20},
        volume = {13},
        number = {12},
        pages = {3231--3244},
        doi = {10.14778/3415478.3415547},
        url = {https://doi.org/10.14778/3415478.3415547},
        year = {2020}
}

Incoming Citations (Sorted by Pagerank)

Showing 5 of 5 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 26 of 26 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.001992968
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00055406774
30 SCOPE: Easy and Efficient Parallel Processing of Massive Data Sets 2008 VLDB 0.00050495102
35 Hekaton: SQL Server’s Memory-Optimized OLTP Engine 2013 SIGMOD 0.00048001919
40 The Case for Learned Index Structures 2018 SIGMOD 0.00046284649
47 PNUTS: Yahoo!'s Hosted Data Serving Platform 2008 VLDB 0.00044066403
63 Amazon Aurora: Design Considerations for High Throughput Cloud-Native Relational Databases 2017 SIGMOD 0.00038531147
88 Automated Selection of Materialized Views and Indexes for SQL Databases 2000 VLDB 0.00035351639
139 Megastore: Providing Scalable, Highly Available Storage for Interactive Services 2011 CIDR 0.00029480065
151 An Efficient, Cost-Driven Index Selection Tool for Microsoft SQL Server 1997 VLDB 0.00028672526
195 Integrating Vertical and Horizontal Partitioning into Automated Physical Database Design 2004 SIGMOD 0.00025628849
206 Generalized Search Trees for Database Systems (Extended Abstract) 1995 VLDB 0.00024986675
218 MillWheel: Fault-Tolerant Stream Processing at Internet Scale 2013 VLDB 0.00024390324
224 Self-Driving Database Management Systems 2017 CIDR 0.00024013745
231 Storm @Twitter 2014 SIGMOD 0.00023841089
325 The Dataflow Model: A Practical Approach to Balancing Correctness, Latency, and Cost in Massive-Scale, Unbounded, Out-of-Order Data Processing 2015 VLDB 0.00020964941
365 Hyder - A Transactional Record Manager for Shared Flash 2011 CIDR 0.00019941855
685 Trill: A High-Performance Incremental Query Processor for Diverse Analytics 2015 VLDB 0.00014782777
1,123 Structured Streaming: A Declarative API for Real-Time Applications in Apache Spark 2018 SIGMOD 0.0001193233
1,751 Split Query Processing in Polybase 2013 SIGMOD 9.7291888e-05
2,263 Concurrency and Recovery in Generalized Search Trees 1997 SIGMOD 8.7304715e-05
2,446 Azure Data Lake Store: A Hyperscale Distributed File Service for Big Data Analytics 2017 SIGMOD 8.4547121e-05
2,909 Deuteronomy: Transaction Support for Cloud Data 2011 CIDR 7.8721738e-05
3,033 Chi: A Scalable and Programmable Control Plane for Distributed Stream Processing Systems 2018 VLDB 7.7337998e-05
7,836 Quill: Efficient, Transferable, and Rich Analytics at Scale 2016 VLDB 5.4435399e-05
8,810 Demonstration: MacroBase, A Fast Data Analysis Engine 2017 SIGMOD 5.2723423e-05
Previous Page 1 / 1 Next

Semantically Similar Papers