DBScholar

Back to papers

Helios: Hyperscale Indexing for the Cloud & Edge

Summary: Helios is a hyperscale cloud/edge indexing system that feeds real-time streams into relational engines. It ingests quadrillions of events and indexes trillions of keys daily across dozens of data centers, using a simple data model and asynchronous indexing - a scalable blueprint for Microsoft-scale analytics. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
12393
Venue
VLDB
Year
2020
Pagerank
5.9957544e-05
Overall Rank
6,045 | 58.53%
DOI
10.14778/3415478.3415547

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{potharaju_vldb20,
        title = {{Helios: Hyperscale Indexing for the Cloud \& Edge}},
        author = {Potharaju, Rahul and Kim, Terry and Wu, Wentao and Acharya, Vidip and Suh, Steve and Fogarty, Andrew and Dave, Apoorve and Ramanujam, Sinduja and Talius, Tomas and Novik, Lev and Ramakrishnan, Raghu},
        journal = {PVLDB},
        series = {{VLDB} '20},
        volume = {13},
        number = {12},
        pages = {3231--3244},
        doi = {10.14778/3415478.3415547},
        url = {https://doi.org/10.14778/3415478.3415547},
        year = {2020}
}

Incoming Citations (Sorted by Pagerank)

Showing 5 of 5 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 26 of 26 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
2 R-Trees: A Dynamic Index Structure For Spatial Searching 1984 SIGMOD 0.0020210012
24 Spark SQL: Relational Data Processing in Spark 2015 SIGMOD 0.00054865648
30 SCOPE: Easy and Efficient Parallel Processing of Massive Data Sets 2008 VLDB 0.00051174276
38 Hekaton: SQL Server’s Memory-Optimized OLTP Engine 2013 SIGMOD 0.00047648573
43 The Case for Learned Index Structures 2018 SIGMOD 0.00046060254
47 PNUTS: Yahoo!'s Hosted Data Serving Platform 2008 VLDB 0.00044503718
73 Amazon Aurora: Design Considerations for High Throughput Cloud-Native Relational Databases 2017 SIGMOD 0.00037333356
87 Automated Selection of Materialized Views and Indexes for SQL Databases 2000 VLDB 0.00035281619
144 Megastore: Providing Scalable, Highly Available Storage for Interactive Services 2011 CIDR 0.00029554682
156 An Efficient, Cost-Driven Index Selection Tool for Microsoft SQL Server 1997 VLDB 0.00028636811
199 Integrating Vertical and Horizontal Partitioning into Automated Physical Database Design 2004 SIGMOD 0.00025612088
202 Generalized Search Trees for Database Systems (Extended Abstract) 1995 VLDB 0.00025454884
220 Storm @Twitter 2014 SIGMOD 0.00024244587
224 MillWheel: Fault-Tolerant Stream Processing at Internet Scale 2013 VLDB 0.00024130894
234 Self-Driving Database Management Systems 2017 CIDR 0.00023810722
361 The Dataflow Model: A Practical Approach to Balancing Correctness, Latency, and Cost in Massive-Scale, Unbounded, Out-of-Order Data Processing 2015 VLDB 0.00020138717
379 Hyder - A Transactional Record Manager for Shared Flash 2011 CIDR 0.00019611067
710 Trill: A High-Performance Incremental Query Processor for Diverse Analytics 2015 VLDB 0.00014715033
1,190 Structured Streaming: A Declarative API for Real-Time Applications in Apache Spark 2018 SIGMOD 0.00011743246
1,739 Split Query Processing in Polybase 2013 SIGMOD 9.8814935e-05
2,227 Concurrency and Recovery in Generalized Search Trees 1997 SIGMOD 8.9111106e-05
2,477 Azure Data Lake Store: A Hyperscale Distributed File Service for Big Data Analytics 2017 SIGMOD 8.5239378e-05
2,903 Deuteronomy: Transaction Support for Cloud Data 2011 CIDR 7.9761235e-05
3,117 Chi: A Scalable and Programmable Control Plane for Distributed Stream Processing Systems 2018 VLDB 7.7382559e-05
7,709 Quill: Efficient, Transferable, and Rich Analytics at Scale 2016 VLDB 5.5628268e-05
8,660 Demonstration: MacroBase, A Fast Data Analysis Engine 2017 SIGMOD 5.3894205e-05
Previous Page 1 / 1 Next

Semantically Similar Papers