DBScholar

Back to papers

HAWQ: A Massively Parallel Processing SQL Engine in Hadoop

Summary: MPP SQL on HDFS combining DBMS-style parallelism with Hadoop; standard SQL with full ACID transactions. UDP interconnect, fault tolerance, read-optimized storage, and extensible data-store support for Hadoop formats; ~40x Stinger, ~35–45x Hive. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
h35076ec5428c11a9
Venue
SIGMOD
Year
2014
Pagerank
7.8853204e-05
Overall Rank
2,898 | 80.52%
DOI
10.1145/2588555.2595636

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{chang_sigmod14,
        title = {{HAWQ: A Massively Parallel Processing SQL Engine in Hadoop}},
        author = {Chang, Lei and Wang, Zhanwei and Ma, Tao and Jian, Lirong and Ma, Lili and Goldshuv, Alon and Lonergan, Luke and Cohen, Jeffrey and Welton, Caleb and Sherry, Gavin and Bhandarkar, Milind},
        series = {{SIGMOD} '14},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/2588555.2595636},
        url = {https://dl.acm.org/doi/10.1145/2588555.2595636},
        year = {2014}
}

Incoming Citations (Sorted by Pagerank)

Showing 10 of 10 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 20 of 20 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
6 Pig Latin: A Not-So-Foreign Language for Data Processing 2008 SIGMOD 0.001052036
19 A Critique of ANSI SQL Isolation Levels 1995 SIGMOD 0.00058781151
31 Hive - A Warehousing Solution Over a Map-Reduce Framework 2009 VLDB 0.00049839909
43 A Comparison of Approaches to Large-Scale Data Analysis 2009 SIGMOD 0.00045546775
48 Weaving Relations for Cache Performance 2001 VLDB 0.00043805923
49 Dremel: Interactive Analysis of Web-Scale Datasets 2010 VLDB 0.00043160717
93 Encapsulation of Parallelism in the Volcano Query Processing System 1990 SIGMOD 0.00034622929
105 The MADlib Analytics Library or MAD Skills, the SQL 2012 VLDB 0.00033638251
120 HadoopDB: An Architectural Hybrid of MapReduce and DBMS Technologies for Analytical Workloads 2009 VLDB 0.000311132
154 MAD Skills: New Analysis Practices for Big Data 2009 VLDB 0.00028579704
432 Shark: SQL and Rich Analytics at Scale 2013 SIGMOD 0.00018339357
617 F1: A Distributed SQL Database That Scales 2013 VLDB 0.00015555815
726 Serializable Snapshot Isolation in PostgreSQL 2012 VLDB 0.00014459037
876 Tenzing: A SQL Implementation On The MapReduce Framework 2011 VLDB 0.00013309176
1,327 Column-oriented Database Systems 2009 VLDB 0.00011003776
1,466 The Performance of MapReduce: An In-depth Study 2010 VLDB 0.00010569837
1,751 Split Query Processing in Polybase 2013 SIGMOD 9.7291888e-05
2,173 Efficient Processing of Data Warehousing Queries in a Split Execution Environment 2011 SIGMOD 8.91924e-05
7,843 Oracle In-Database Hadoop: When MapReduce Meets RDBMS 2012 SIGMOD 5.4424628e-05
8,068 Emerging Trends in the Enterprise Data Analytics: Connecting Hadoop and DB2 Warehouse 2011 SIGMOD 5.394659e-05
Previous Page 1 / 1 Next

Semantically Similar Papers