Mesa: Geo-Replicated, Near Real-Time, Scalable Data Warehousing
Summary: Geo-replicated, near real-time analytics warehouse for petabyte-scale measurement data, handling millions of row updates per second and billions of queries daily. Guarantees consistent, repeatable query answers across datacenters and during datacenter outages, highlighting end-to-end scalability, availability, and fault tolerance. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Ashish Gupta (Google)
- 2. Fan Yang (Google)
- 3. Jason Govig (Google)
- 4. Adam Kirsch (Google)
- 5. Kelvin Chan (Google)
- 6. Kevin Lai (Google)
- 7. Shuo Wu (Google)
- 8. Sandeep Govind Dhoot (Google)
- 9. Abhilash Rajesh Kumar (Google)
- 10. Ankur Agiwal (Google)
- 11. Sanjay Bhansali (Google)
- 12. Mingsheng Hong (Google)
- 13. Jamie Cameron (Google)
- 14. Masood Siddiqi (Google)
- 15. David Jones (Google)
- 16. Jeff Shute (Google)
- 17. Andrey Gubarev (Google)
- 18. Shivakumar Venkataraman (Google)
- 19. Divyakant Agrawal (Google)
BibTeX Citation
@article{gupta_vldb14,
title = {{Mesa: Geo-Replicated, Near Real-Time, Scalable Data Warehousing}},
author = {Gupta, Ashish and Yang, Fan and Govig, Jason and Kirsch, Adam and Chan, Kelvin and Lai, Kevin and Wu, Shuo and Dhoot, Sandeep Govind and Kumar, Abhilash Rajesh and Agiwal, Ankur and Bhansali, Sanjay and Hong, Mingsheng and Cameron, Jamie and Siddiqi, Masood and Jones, David and Shute, Jeff and Gubarev, Andrey and Venkataraman, Shivakumar and Agrawal, Divyakant},
journal = {PVLDB},
series = {{VLDB} '14},
volume = {7},
number = {12},
pages = {1259--1270},
doi = {10.14778/2732977.2732999},
url = {https://doi.org/10.14778/2732977.2732999},
year = {2014}
}
Incoming Citations (Sorted by Pagerank)
Showing 23 of 23 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 31 of 31 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 11,823 | A Drop-in Middleware for Serializable DB Clustering across Geo-distributed Sites | 2020 | VLDB |
| 2 | 8,001 | Pangea: Monolithic Distributed Storage for Data Analytics | 2019 | VLDB |
| 3 | 1,593 | F1 – The Fault-Tolerant Distributed RDBMS Supporting Google's Ad Business | 2012 | SIGMOD |
| 4 | 10,059 | Progressive Partitioning for Parallelized Query Execution in Google’s Napa | 2023 | VLDB |
| 5 | 9,274 | DataGarage: Warehousing Massive Performance Data on Commodity Servers | 2010 | VLDB |
| 6 | 6,281 | Peta-Scale Data Warehousing at Yahoo! | 2009 | SIGMOD |
| 7 | 3,627 | Data Management Projects at Google | 2006 | SIGMOD |
| 8 | 8,272 | Shasta: Interactive Reporting At Scale | 2016 | SIGMOD |
| 9 | 144 | Megastore: Providing Scalable, Highly Available Storage for Interactive Services | 2011 | CIDR |
| 10 | 4,017 | Napa: Powering Scalable Data Warehousing with Robust Query Performance at Google | 2021 | VLDB |