DBScholar

Back to papers

Apache Hive: From MapReduce to Enterprise-grade Big Data Warehousing

Summary: Apache Hive evolves from MapReduce to an enterprise-grade data warehouse via a hybrid MPP/big-data architecture that integrates SQL, storage formats, and cloud concepts. Innovations cover Transactions, optimizer, runtime, and federation, with experiments on typical workloads and a forward-looking community roadmap. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
5719
Venue
SIGMOD
Year
2019
Pagerank
7.3115321e-05
Overall Rank
3,555 | 75.62%
DOI
10.1145/3299869.3314045

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{camachorodriguez_sigmod19,
        title = {{Apache Hive: From MapReduce to Enterprise-grade Big Data Warehousing}},
        author = {Camacho-Rodríguez, Jesús and Chauhan, Ashutosh and Gates, Alan and Koifman, Eugene and O’Malley, Owen and Garg, Vineet and Haindrich, Zoltan and Shelukhin, Sergey and Jayachandran, Prasanth and Seth, Siddharth and Jaiswal, Deepak and Bouguerra, Slim and Bangarwa, Nishant and Hariappan, Sankar and Agarwal, Anishek and Dere, Jason and Dai, Daniel and Nair, Thejas and Dembla, Nita and Vijayaraghavan, Gopal and Hagleitner, Günther},
        series = {{SIGMOD} '19},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/3299869.3314045},
        url = {https://dl.acm.org/doi/10.1145/3299869.3314045},
        year = {2019}
}

Incoming Citations (Sorted by Pagerank)

Showing 12 of 12 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 20 of 20 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
11 Implementing Data Cubes Efficiently 1996 SIGMOD 0.00071822821
51 Dremel: Interactive Analysis of Web-Scale Datasets 2010 VLDB 0.0004291425
66 The Snowflake Elastic Data Warehouse 2016 SIGMOD 0.00038561587
87 Automated Selection of Materialized Views and Indexes for SQL Databases 2000 VLDB 0.00035281619
100 LEO - DB2's LEarning Optimizer 2001 VLDB 0.00034385207
184 DB2 Design Advisor: Integrated Automatic Physical Database Design 2004 VLDB 0.00026256101
237 Amazon Redshift and the Case for Simpler Data Warehouses 2015 SIGMOD 0.0002369895
245 Serializable Isolation for Snapshot Databases 2008 SIGMOD 0.00023458287
330 Impala: A Modern, Open-Source SQL Engine for Hadoop 2015 CIDR 0.0002104801
375 Incremental Maintenance of Views with Duplicates 1995 SIGMOD 0.00019681204
445 Apache Calcite: A Foundational Framework for Optimized Query Processing Over Heterogeneous Data Sources 2018 SIGMOD 0.00018336751
559 Optimizing Queries Using Materialized Views: A Practical, Scalable Solution 2001 SIGMOD 0.00016528822
940 Enhancements to SQL Server Column Stores 2013 SIGMOD 0.00013081205
1,242 Druid: A Real-time Analytical Data Store 2014 SIGMOD 0.00011516162
1,621 Orca: A Modular Query Optimizer Architecture for Big Data 2014 SIGMOD 0.00010203114
2,706 Major Technical Advancements in Apache Hive 2014 SIGMOD 8.2287564e-05
3,605 Computation Reuse in Analytics Job Service at Microsoft 2018 SIGMOD 7.2640711e-05
3,885 Apache Tez: A Unifying Framework for Modeling and Building Data Processing Applications 2015 SIGMOD 7.0475239e-05
4,277 Adaptive Statistics in Oracle 12c 2017 VLDB 6.7873816e-05
7,428 Invisible Glue: Scalable Self-Tuning Multi-Stores 2015 CIDR 5.6205394e-05
Previous Page 1 / 1 Next

Semantically Similar Papers