DBScholar

Back to papers

HAWQ: A Massively Parallel Processing SQL Engine in Hadoop

Summary: MPP SQL on HDFS combining DBMS-style parallelism with Hadoop; standard SQL with full ACID transactions. UDP interconnect, fault tolerance, read-optimized storage, and extensible data-store support for Hadoop formats; ~40x Stinger, ~35–45x Hive. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h35076ec5428c11a9
Venue
SIGMOD
Year
2014
Pagerank
7.8816694e-05
Overall Rank
2,898 | 80.53%
DOI
10.1145/2588555.2595636

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{chang_sigmod14,
        title = {{HAWQ: A Massively Parallel Processing SQL Engine in Hadoop}},
        author = {Chang, Lei and Wang, Zhanwei and Ma, Tao and Jian, Lirong and Ma, Lili and Goldshuv, Alon and Lonergan, Luke and Cohen, Jeffrey and Welton, Caleb and Sherry, Gavin and Bhandarkar, Milind},
        series = {{SIGMOD} '14},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/2588555.2595636},
        url = {https://dl.acm.org/doi/10.1145/2588555.2595636},
        year = {2014}
}

Incoming Citations (Sorted by Pagerank)

Showing 10 of 10 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 20 of 20 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
6 Pig Latin: A Not-So-Foreign Language for Data Processing 2008 SIGMOD 0.0010515896
19 A Critique of ANSI SQL Isolation Levels 1995 SIGMOD 0.00058759613
31 Hive - A Warehousing Solution Over a Map-Reduce Framework 2009 VLDB 0.00049821554
43 A Comparison of Approaches to Large-Scale Data Analysis 2009 SIGMOD 0.0004552807
48 Weaving Relations for Cache Performance 2001 VLDB 0.00043795812
49 Dremel: Interactive Analysis of Web-Scale Datasets 2010 VLDB 0.0004314366
93 Encapsulation of Parallelism in the Volcano Query Processing System 1990 SIGMOD 0.00034607573
105 The MADlib Analytics Library or MAD Skills, the SQL 2012 VLDB 0.00033633007
120 HadoopDB: An Architectural Hybrid of MapReduce and DBMS Technologies for Analytical Workloads 2009 VLDB 0.00031099083
154 MAD Skills: New Analysis Practices for Big Data 2009 VLDB 0.00028568843
432 Shark: SQL and Rich Analytics at Scale 2013 SIGMOD 0.00018331051
617 F1: A Distributed SQL Database That Scales 2013 VLDB 0.00015548752
727 Serializable Snapshot Isolation in PostgreSQL 2012 VLDB 0.00014456009
876 Tenzing: A SQL Implementation On The MapReduce Framework 2011 VLDB 0.00013304112
1,328 Column-oriented Database Systems 2009 VLDB 0.00010998855
1,466 The Performance of MapReduce: An In-depth Study 2010 VLDB 0.00010565007
1,752 Split Query Processing in Polybase 2013 SIGMOD 9.7246277e-05
2,176 Efficient Processing of Data Warehousing Queries in a Split Execution Environment 2011 SIGMOD 8.9150466e-05
7,847 Oracle In-Database Hadoop: When MapReduce Meets RDBMS 2012 SIGMOD 5.4398885e-05
8,074 Emerging Trends in the Enterprise Data Analytics: Connecting Hadoop and DB2 Warehouse 2011 SIGMOD 5.3921606e-05
Previous Page 1 / 1 Next

Semantically Similar Papers