Back to authors
Reynold Xin
- Author ID
- o0009-0002-5173-1578
- ORCID
-
0009-0002-5173-1578
- Links
-
(found by gpt-5.6-luna on jul 24 2026)
- Most Frequent Institution
- Databricks
- Pagerank
- 0.24645181
- Overall Rank
- 217 | 99.00%
- Paper Count
- 27
Affiliation Timeline
Incoming Non-self Citations Over Time
Total yearly non-self incoming citations across all papers by this author.
Publications by Paper Pagerank
Showing 27 of 27 publications.
| Rank |
Title |
Year |
Venue |
Pagerank |
| 23 |
Spark SQL: Relational Data Processing in Spark |
2015 |
SIGMOD |
0.00055406774 |
| 92 |
CrowdDB: Answering Queries with Crowdsourcing |
2011 |
SIGMOD |
0.00034672523 |
| 432 |
Shark: SQL and Rich Analytics at Scale |
2013 |
SIGMOD |
0.00018339357 |
| 459 |
Delta Lake: High-Performance ACID Table Storage over Cloud Object Stores |
2020 |
VLDB |
0.00017856221 |
| 717 |
Finding Related Tables |
2012 |
SIGMOD |
0.00014532116 |
| 950 |
Lakehouse: A New Generation of Open Platforms that Unify Data Warehousing and Advanced Analytics |
2021 |
CIDR |
0.00012895553 |
| 1,036 |
Fine-grained Partitioning for Aggressive Data Skipping |
2014 |
SIGMOD |
0.00012377471 |
| 1,123 |
Structured Streaming: A Declarative API for Real-Time Applications in Apache Spark |
2018 |
SIGMOD |
0.0001193233 |
| 1,487 |
Photon: A Fast Query Engine for Lakehouse Systems |
2022 |
SIGMOD |
0.00010521722 |
| 1,851 |
CrowdDB: Query Processing with the VLDB Crowd |
2011 |
VLDB |
9.5037355e-05 |
| 2,232 |
Shark: Fast Data Analysis Using Coarse-grained Distributed Memory |
2012 |
SIGMOD |
8.7960223e-05 |
| 3,455 |
Scaling Spark in the Real World: Performance and Usability |
2015 |
VLDB |
7.2884813e-05 |
| 5,348 |
Adaptive and Robust Query Execution for Lakehouses at Scale |
2024 |
VLDB |
6.1690434e-05 |
| 6,987 |
SparkR: Scaling R Programs with Spark |
2016 |
SIGMOD |
5.6283496e-05 |
| 7,582 |
Unity Catalog: Open and Universal Governance for the Lakehouse and Beyond |
2025 |
SIGMOD |
5.4910839e-05 |
| 7,996 |
Databricks Lakeguard: Supporting Fine-grained Access Control and Multi-user Capabilities for Apache Spark Workloads |
2025 |
SIGMOD |
5.4108483e-05 |
| 9,282 |
Making Data Engineering Declarative |
2023 |
CIDR |
5.2034134e-05 |
| 9,988 |
Introduction to Spark 2.0 for Database Researchers |
2016 |
SIGMOD |
5.0998353e-05 |
| 10,079 |
Delta Sharing: An Open Protocol for Cross-Platform Data Sharing |
2025 |
VLDB |
5.0830849e-05 |
| 10,190 |
MEET DB2: Automated Database Migration Evaluation |
2010 |
VLDB |
5.0644123e-05 |
| 10,916 |
AutoLiquid: Autonomic Data Layout Optimization for the Databricks Lakehouse |
2026 |
VLDB |
4.9793485e-05 |
| 10,941 |
Ultron: History-Based Query Optimization at Databricks |
2026 |
VLDB |
4.9793485e-05 |
| 10,943 |
Lakebase: Serverless Postgres over Open Lake Storage |
2026 |
VLDB |
4.9793485e-05 |
| 12,482 |
A Partitioning Framework for Aggressive Data Skipping |
2014 |
VLDB |
4.9793485e-05 |
| 12,805 |
Linkage Query Writer |
2009 |
VLDB |
4.9793485e-05 |
| 13,614 |
The Three Golden Ages of Database Engineering: From SIGMOD’85 to the Agentic Era |
2026 |
VLDB |
- |
| 13,623 |
Blink Twice - Automatic Workload Pinning and Regression Detection for Versionless Apache Spark using Retries |
2025 |
SIGMOD |
- |
Frequent Co-authors
Co-authored at least 5 papers.