DBScholar

Back to authors

Christopher Ré

Author ID
u314
ORCID
-
Links
(found by gpt-5.6-luna on jul 24 2026)
Most Frequent Institution
Stanford University
Pagerank
0.50355869
Overall Rank
63 | 99.71%
Paper Count
55

Affiliation Timeline

Incoming Non-self Citations Over Time

Total yearly non-self incoming citations across all papers by this author.

Publications by Paper Pagerank

Showing all 55 publications. Total citations include self and non-self citations.

Rank Title Year Venue Total Citations Pagerank
104 HoloClean: Holistic Data Repairs with Probabilistic Inference 2017 VLDB 143 0.00033676943
105 The MADlib Analytics Library or MAD Skills, the SQL 2012 VLDB 108 0.00033633007
205 Snorkel: Rapid Training Data Creation with Weak Supervision 2018 VLDB 72 0.00025171314
208 EmptyHeaded: A Relational Engine for Graph Processing 2016 SIGMOD 96 0.00024899872
329 Can Foundation Models Wrangle Your Data? 2023 VLDB 64 0.00020867521
402 Worst-case Optimal Join Algorithms 2012 PODS 63 0.00019095982
496 Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes 2024 VLDB 40 0.00017318538
503 Towards a Unified Architecture for in-RDBMS Analytics 2012 SIGMOD 49 0.00017195428
579 Incremental Knowledge Base Construction Using DeepDive 2015 VLDB 44 0.00016083582
654 Materialization Optimizations for Feature Selection Workloads 2014 SIGMOD 46 0.00015096817
696 MYSTIQ: A system for finding more answers by using probabilities 2005 SIGMOD 33 0.00014691377
1,063 Tuffy: Scaling up Statistical Inference in Markov Logic Networks using an RDBMS 2011 VLDB 23 0.00012208
1,120 Snuba: Automating Weak Supervision to Label Training Data 2019 VLDB 26 0.000119406
1,167 DimmWitted: A Study of Main-Memory Statistical Analytics 2014 VLDB 25 0.0001172597
1,282 Automatic Optimization for MapReduce Programs 2011 VLDB 19 0.00011208192
1,570 AJAR: Aggregations and Joins over Annotated Relations 2016 PODS 25 0.0001020376
1,726 Beyond Worst-case Analysis for Joins with Minesweeper 2014 PODS 18 9.7857214e-05
1,759 Structured Querying of Web Text: A Technical Challenge 2007 CIDR 11 9.7090189e-05
1,785 Approximate Lineage for Probabilistic Databases 2008 VLDB 20 9.6438009e-05
1,813 Joins via Geometric Resolutions: Worst-case and Beyond 2015 PODS 23 9.5759542e-05
2,587 Brainwash: A Data System for Feature Engineering 2013 CIDR 23 8.2523942e-05
2,985 Materialized Views in Probabilistic Databases: For Information Exchange and Query Optimization 2007 VLDB 12 7.7794901e-05
3,063 Fonduer: Knowledge Base Construction from Richly Formatted Data 2018 SIGMOD 12 7.689108e-05
3,112 Event Queries on Correlated Probabilistic Streams 2008 SIGMOD 18 7.6351481e-05
3,521 Ember: No-Code Context Enrichment via Similarity-Based Keyless Joins 2022 VLDB 10 7.2320843e-05
3,627 The Role of Massively Multi-Task and Weak Supervision in Software 2.0 2019 CIDR 5 7.1492314e-05
3,718 Machine Learning and Databases: The Sound of Things to Come or a Cacophony of Hype? 2015 SIGMOD 7 7.0703796e-05
3,737 Snorkel: Fast Training Set Generation for Information Extraction 2017 SIGMOD 8 7.0610219e-05
3,793 SLiMFast: Guaranteed Results for Data Fusion and Source Reliability 2017 SIGMOD 15 7.0156889e-05
3,909 DunceCap: Query Plans Using Generalized Hypertree Decompositions 2015 SIGMOD 10 6.9287689e-05
3,988 Exploiting Correlations for Expensive Predicate Evaluation 2015 SIGMOD 6 6.8694751e-05
3,998 Overton: A Data System for Monitoring and Improving Machine-Learned Products 2020 CIDR 10 6.862274e-05
4,029 Towards High-Throughput Gibbs Sampling at Scale: A Study across Storage Managers 2013 SIGMOD 9 6.840359e-05
4,395 Extracting Databases from Dark Data with DeepDive 2016 SIGMOD 8 6.6175689e-05
5,182 Snorkel DryBell: A Case Study in Deploying Weak Supervision at Industrial Scale 2019 SIGMOD 9 6.2377015e-05
5,809 Incrementally Maintaining Classification using an RDBMS 2011 VLDB 6 5.9848544e-05
5,922 A Relational Framework for Classifier Engineering 2017 PODS 5 5.9442275e-05
5,974 Understanding Cardinality Estimation using Entropy Maximization 2010 PODS 5 5.9279222e-05
7,026 GeoDeepDive: Statistical Inference using Familiar Data-Processing Languages 2013 SIGMOD 2 5.6157971e-05
7,542 Feature Selection in Enterprise Analytics: A Demonstration using an R-based Data Analytics System 2013 VLDB 9 5.4976824e-05
7,872 Mind the Gap: Bridging Multi-Domain Query Workloads with EmptyHeaded 2017 VLDB 3 5.4339723e-05
8,314 DunceCap: Compiling Worst-Case Optimal Query Plans 2015 SIGMOD 4 5.3557545e-05
8,755 Probabilistic Management of OCR Data using an RDBMS 2012 VLDB 2 5.2841044e-05
8,756 A Demonstration of Cascadia Through a Digital Diary Application 2008 SIGMOD 1 5.2830007e-05
9,771 Bootleg: Chasing the Tail with Self-Supervised Named Entity Disambiguation 2021 CIDR 4 5.1320012e-05
9,838 Automating the Enterprise with Foundation Models 2024 VLDB 1 5.1221535e-05
10,179 Is Data Management the Beating Heart of AI Systems? 2022 SIGMOD 1 5.06553e-05
11,831 Data Management Opportunities for Foundation Models 2022 CIDR 0 4.9769913e-05
12,131 Leveraging Organizational Resources to Adapt Models to New Data Modalities 2020 VLDB 2 4.9769913e-05
12,434 Mindtagger: A Demonstration of Data Labeling in Knowledge Base Construction 2015 VLDB 1 4.9769913e-05
12,705 Transducing Markov Sequences 2010 PODS 3 4.9769913e-05
12,896 Systems Aspects of Probabilistic Data Management 2008 VLDB 0 4.9769913e-05
13,738 The DB Community vis-à-vis Environmental, Health, and Societal Grand Challenges: Innovation Engine, Plumber, or Bystander? 2022 SIGMOD 0 -
13,967 Ringtail: A Generalized Nowcasting System 2013 VLDB 1 -
14,062 Lahar Demonstration: Warehousing Markovian Streams 2009 VLDB 1 -

Frequent Co-authors

Co-authored at least 5 papers.

Co-author Shared Papers Rank Pagerank
Dan Suciu 8 5 1.1591003
Michael John Cafarella 7 80 0.44642508
Chun Zhang 7 97 0.39504494
Magdalena Balazinska 6 50 0.56350834
Arun Kumar 6 125 0.35553538