On Brewing Fresh Espresso: LinkedIn’s Distributed Data Serving Platform
Summary: Espresso: LinkedIn's scalable, document-oriented store with cross-document transactions, real-time indexing, on-the-fly schema evolution, and timeline-consistent change capture. Innovations include a generic distributed cluster manager, partition-aware change capture, and a high-performance inverted index, with empirical results. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Lin Qiao (LinkedIn)
- 2. Kapil Surlaker (LinkedIn)
- 3. Shirshanka Das (LinkedIn)
- 4. Tom Quiggle (LinkedIn)
- 5. Bob Schulman (LinkedIn)
- 6. Bhaskar Ghosh (LinkedIn)
- 7. Antony Curtis (LinkedIn)
- 8. Oliver Seeliger (LinkedIn)
- 9. Zhen Zhang (LinkedIn)
- 10. Aditya Auradkar (LinkedIn)
- 11. Chris Beavers (LinkedIn)
- 12. Gregory Brandt (LinkedIn)
- 13. Mihir Gandhi (LinkedIn)
- 14. Kishore Gopalakrishna (LinkedIn)
- 15. Wai Ip (LinkedIn)
- 16. Swaroop Jagadish (LinkedIn)
- 17. Shi Lu (LinkedIn)
- 18. Alexander Pachev (LinkedIn)
- 19. Aditya Ramesh (LinkedIn)
- 20. Abraham Sebastian (LinkedIn)
- 21. Rupa Shanbhag (LinkedIn)
- 22. Subbu Subramaniam (LinkedIn)
- 23. Yun Sun (LinkedIn)
- 24. Sajid Topiwala (LinkedIn)
- 25. Cuong Tran (LinkedIn)
- 26. Jemiah Westerman (LinkedIn)
- 27. David Zhang (LinkedIn)
BibTeX Citation
@inproceedings{qiao_sigmod13,
title = {{On Brewing Fresh Espresso: LinkedIn’s Distributed Data Serving Platform}},
author = {Qiao, Lin and Surlaker, Kapil and Das, Shirshanka and Quiggle, Tom and Schulman, Bob and Ghosh, Bhaskar and Curtis, Antony and Seeliger, Oliver and Zhang, Zhen and Auradkar, Aditya and Beavers, Chris and Brandt, Gregory and Gandhi, Mihir and Gopalakrishna, Kishore and Ip, Wai and Jagadish, Swaroop and Lu, Shi and Pachev, Alexander and Ramesh, Aditya and Sebastian, Abraham and Shanbhag, Rupa and Subramaniam, Subbu and Sun, Yun and Topiwala, Sajid and Tran, Cuong and Westerman, Jemiah and Zhang, David},
series = {{SIGMOD} '13},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/2463676.2465298},
url = {https://dl.acm.org/doi/10.1145/2463676.2465298},
year = {2013}
}
Incoming Citations (Sorted by Pagerank)
Showing 14 of 14 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 3 of 3 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 47 | PNUTS: Yahoo!'s Hosted Data Serving Platform | 2008 | VLDB | 0.00044503718 |
| 144 | Megastore: Providing Scalable, Highly Available Storage for Interactive Services | 2011 | CIDR | 0.00029554682 |
| 1,593 | F1 – The Fault-Tolerant Distributed RDBMS Supporting Google's Ad Business | 2012 | SIGMOD | 0.00010252961 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 1,895 | Samza: Stateful Scalable Stream Processing at LinkedIn | 2017 | VLDB |
| 2 | 6,004 | Data Ingestion for the Connected World | 2017 | CIDR |
| 3 | 7,492 | Scalable Distributed Inverted List Indexes in Disaggregated Memory | 2024 | SIGMOD |
| 4 | 11,689 | A Client-centric Approach to Transactional Datastores | 2021 | SIGMOD |
| 5 | 5,790 | Supporting a Semantic Data Model in a Distributed Database System | 1983 | VLDB |
| 6 | 3,109 | Managing Large Dynamic Graphs Efficiently | 2012 | SIGMOD |
| 7 | 5,419 | Building a Replicated Logging System with Apache Kafka | 2015 | VLDB |
| 8 | 1,879 | Predictable Performance for Unpredictable Workloads | 2009 | VLDB |
| 9 | 4,595 | The "Big Data" Ecosystem at LinkedIn | 2013 | SIGMOD |
| 10 | 6,734 | Liquid: Unifying Nearline and Offline Big Data Integration | 2015 | CIDR |