DBScholar

Back to papers

Indexing HDFS Data in PDW: Splitting the data from the index

Summary: Proposes using B+-tree indices in an RDBMS to access data stored in HDFS, effectively splitting storage from indexing. Demonstrates that the PDW-driven approach yields efficient, highly selective query processing by exploiting RDBMS indexing on Hadoop data. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
11000
Venue
VLDB
Year
2014
Pagerank
5.5073329e-05
Overall Rank
8,014 | 45.02%
DOI
10.14778/2733004.2733024

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{gankidi_vldb14,
        title = {{Indexing HDFS Data in PDW: Splitting the data from the index}},
        author = {Gankidi, Vinitha Reddy and Teletia, Nikhil and Patel, Jignesh M. and Halverson, Alan and DeWitt, David J.},
        journal = {PVLDB},
        series = {{VLDB} '14},
        volume = {7},
        number = {13},
        pages = {1520--1531},
        doi = {10.14778/2733004.2733024},
        url = {https://doi.org/10.14778/2733004.2733024},
        year = {2014}
}

Incoming Citations (Sorted by Pagerank)

Showing 2 of 2 citing papers.

Rank Citing Paper Year Venue Pagerank
3,413 Slalom: Coasting Through Raw Data via Adaptive Partitioning and Indexing 2017 VLDB 7.4326381e-05
8,490 FusionInsight LibrA: Huawei’s Enterprise Cloud Data Analytics Platform 2018 VLDB 5.4152314e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 4 of 4 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers