DBScholar

Back to papers

Helios: Hyperscale Indexing for the Cloud & Edge

Summary: Helios is a hyperscale cloud/edge indexing system that feeds real-time streams into relational engines. It ingests quadrillions of events and indexes trillions of keys daily across dozens of data centers, using a simple data model and asynchronous indexing - a scalable blueprint for Microsoft-scale analytics. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h3366fa01e8189050
Venue
VLDB
Year
2020
Pagerank
5.8585082e-05
Overall Rank
6,176 | 58.50%
DOI
10.14778/3415478.3415547
PDF
Download (CC BY-NC-ND 4.0)

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{potharaju_vldb20,
        title = {{Helios: Hyperscale Indexing for the Cloud \& Edge}},
        author = {Potharaju, Rahul and Kim, Terry and Wu, Wentao and Acharya, Vidip and Suh, Steve and Fogarty, Andrew and Dave, Apoorve and Ramanujam, Sinduja and Talius, Tomas and Novik, Lev and Ramakrishnan, Raghu},
        journal = {PVLDB},
        series = {{VLDB} '20},
        volume = {13},
        number = {12},
        pages = {3231--3244},
        doi = {10.14778/3415478.3415547},
        url = {https://doi.org/10.14778/3415478.3415547},
        year = {2020}
}

Incoming Citations (Sorted by Pagerank)

Showing 5 of 5 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 26 of 26 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.0019923528
23 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00055384955
30 SCOPE: Easy and Efficient Parallel Processing of Massive Data Sets 2008 VLDB 0.00050475202
35 Hekaton: SQL Server’s Memory-Optimized OLTP Engine 2013 SIGMOD 0.00047996489
40 The Case for Learned Index Structures 2018 SIGMOD 0.00046363107
47 PNUTS: Yahoo!'s Hosted Data Serving Platform 2008 VLDB 0.00044047047
63 Amazon Aurora: Design Considerations for High Throughput Cloud-Native Relational Databases 2017 SIGMOD 0.00038521221
88 Automated Selection of Materialized Views and Indexes for SQL Databases 2000 VLDB 0.00035340164
139 Megastore: Providing Scalable, Highly Available Storage for Interactive Services 2011 CIDR 0.00029466567
151 An Efficient, Cost-Driven Index Selection Tool for Microsoft SQL Server 1997 VLDB 0.00028664776
195 Integrating Vertical and Horizontal Partitioning into Automated Physical Database Design 2004 SIGMOD 0.00025619089
207 Generalized Search Trees for Database Systems (Extended Abstract) 1995 VLDB 0.00024976482
218 MillWheel: Fault-Tolerant Stream Processing at Internet Scale 2013 VLDB 0.00024379041
224 Self-Driving Database Management Systems 2017 CIDR 0.00024011047
231 Storm @Twitter 2014 SIGMOD 0.00023830094
325 The Dataflow Model: A Practical Approach to Balancing Correctness, Latency, and Cost in Massive-Scale, Unbounded, Out-of-Order Data Processing 2015 VLDB 0.0002095522
365 Hyder - A Transactional Record Manager for Shared Flash 2011 CIDR 0.00019935521
685 Trill: A High-Performance Incremental Query Processor for Diverse Analytics 2015 VLDB 0.00014778299
1,123 Structured Streaming: A Declarative API for Real-Time Applications in Apache Spark 2018 SIGMOD 0.00011926683
1,752 Split Query Processing in Polybase 2013 SIGMOD 9.7246277e-05
2,265 Concurrency and Recovery in Generalized Search Trees 1997 SIGMOD 8.7267108e-05
2,448 Azure Data Lake Store: A Hyperscale Distributed File Service for Big Data Analytics 2017 SIGMOD 8.4508839e-05
2,911 Deuteronomy: Transaction Support for Cloud Data 2011 CIDR 7.8685228e-05
3,034 Chi: A Scalable and Programmable Control Plane for Distributed Stream Processing Systems 2018 VLDB 7.7301387e-05
7,840 Quill: Efficient, Transferable, and Rich Analytics at Scale 2016 VLDB 5.4409637e-05
8,818 Demonstration: MacroBase, A Fast Data Analysis Engine 2017 SIGMOD 5.2698464e-05
Previous Page 1 / 1 Next

Semantically Similar Papers