DBScholar

Back to papers

Apache Hive: From MapReduce to Enterprise-grade Big Data Warehousing

Summary: Apache Hive evolves from MapReduce to an enterprise-grade data warehouse via a hybrid MPP/big-data architecture that integrates SQL, storage formats, and cloud concepts. Innovations cover Transactions, optimizer, runtime, and federation, with experiments on typical workloads and a forward-looking community roadmap. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h3bd0ee3c00c57f96
Venue
SIGMOD
Year
2019
Pagerank
7.2942885e-05
Overall Rank
3,446 | 76.84%
DOI
10.1145/3299869.3314045

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{camachorodriguez_sigmod19,
        title = {{Apache Hive: From MapReduce to Enterprise-grade Big Data Warehousing}},
        author = {Camacho-Rodríguez, Jesús and Chauhan, Ashutosh and Gates, Alan and Koifman, Eugene and O’Malley, Owen and Garg, Vineet and Haindrich, Zoltan and Shelukhin, Sergey and Jayachandran, Prasanth and Seth, Siddharth and Jaiswal, Deepak and Bouguerra, Slim and Bangarwa, Nishant and Hariappan, Sankar and Agarwal, Anishek and Dere, Jason and Dai, Daniel and Nair, Thejas and Dembla, Nita and Vijayaraghavan, Gopal and Hagleitner, Günther},
        series = {{SIGMOD} '19},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3299869.3314045},
        url = {https://dl.acm.org/doi/10.1145/3299869.3314045},
        year = {2019}
}

Incoming Citations (Sorted by Pagerank)

Showing 13 of 13 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 20 of 20 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
11 Implementing Data Cubes Efficiently 1996 SIGMOD 0.00071084324
49 Dremel: Interactive Analysis of Web-Scale Datasets 2010 VLDB 0.00043160717
52 The Snowflake Elastic Data Warehouse 2016 SIGMOD 0.00041219077
88 Automated Selection of Materialized Views and Indexes for SQL Databases 2000 VLDB 0.00035351639
98 LEO - DB2's LEarning Optimizer 2001 VLDB 0.00034106982
187 DB2 Design Advisor: Integrated Automatic Physical Database Design 2004 VLDB 0.0002592488
233 Amazon Redshift and the Case for Simpler Data Warehouses 2015 SIGMOD 0.00023783585
237 Serializable Isolation for Snapshot Databases 2008 SIGMOD 0.00023660039
327 Impala: A Modern, Open-Source SQL Engine for Hadoop 2015 CIDR 0.0002095191
379 Apache Calcite: A Foundational Framework for Optimized Query Processing Over Heterogeneous Data Sources 2018 SIGMOD 0.00019514689
388 Incremental Maintenance of Views with Duplicates 1995 SIGMOD 0.00019350381
553 Optimizing Queries Using Materialized Views: A Practical, Scalable Solution 2001 SIGMOD 0.0001652591
944 Enhancements to SQL Server Column Stores 2013 SIGMOD 0.00012939697
1,236 Druid: A Real-time Analytical Data Store 2014 SIGMOD 0.00011402848
1,276 Orca: A Modular Query Optimizer Architecture for Big Data 2014 SIGMOD 0.00011239266
2,689 Major Technical Advancements in Apache Hive 2014 SIGMOD 8.1264718e-05
3,545 Computation Reuse in Analytics Job Service at Microsoft 2018 SIGMOD 7.2134803e-05
3,891 Apache Tez: A Unifying Framework for Modeling and Building Data Processing Applications 2015 SIGMOD 6.9432955e-05
4,234 Adaptive Statistics in Oracle 12c 2017 VLDB 6.7134191e-05
7,528 Invisible Glue: Scalable Self-Tuning Multi-Stores 2015 CIDR 5.5011847e-05
Previous Page 1 / 1 Next

Semantically Similar Papers