| 104 |
HoloClean: Holistic Data Repairs with Probabilistic Inference |
2017 |
VLDB |
143 |
0.00033676943 |
| 105 |
The MADlib Analytics Library or MAD Skills, the SQL |
2012 |
VLDB |
108 |
0.00033633007 |
| 205 |
Snorkel: Rapid Training Data Creation with Weak Supervision |
2018 |
VLDB |
72 |
0.00025171314 |
| 208 |
EmptyHeaded: A Relational Engine for Graph Processing |
2016 |
SIGMOD |
96 |
0.00024899872 |
| 329 |
Can Foundation Models Wrangle Your Data? |
2023 |
VLDB |
64 |
0.00020867521 |
| 402 |
Worst-case Optimal Join Algorithms |
2012 |
PODS |
63 |
0.00019095982 |
| 496 |
Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes |
2024 |
VLDB |
40 |
0.00017318538 |
| 503 |
Towards a Unified Architecture for in-RDBMS Analytics |
2012 |
SIGMOD |
49 |
0.00017195428 |
| 579 |
Incremental Knowledge Base Construction Using DeepDive |
2015 |
VLDB |
44 |
0.00016083582 |
| 654 |
Materialization Optimizations for Feature Selection Workloads |
2014 |
SIGMOD |
46 |
0.00015096817 |
| 696 |
MYSTIQ: A system for finding more answers by using probabilities |
2005 |
SIGMOD |
33 |
0.00014691377 |
| 1,063 |
Tuffy: Scaling up Statistical Inference in Markov Logic Networks using an RDBMS |
2011 |
VLDB |
23 |
0.00012208 |
| 1,120 |
Snuba: Automating Weak Supervision to Label Training Data |
2019 |
VLDB |
26 |
0.000119406 |
| 1,167 |
DimmWitted: A Study of Main-Memory Statistical Analytics |
2014 |
VLDB |
25 |
0.0001172597 |
| 1,282 |
Automatic Optimization for MapReduce Programs |
2011 |
VLDB |
19 |
0.00011208192 |
| 1,570 |
AJAR: Aggregations and Joins over Annotated Relations |
2016 |
PODS |
25 |
0.0001020376 |
| 1,726 |
Beyond Worst-case Analysis for Joins with Minesweeper |
2014 |
PODS |
18 |
9.7857214e-05 |
| 1,759 |
Structured Querying of Web Text: A Technical Challenge |
2007 |
CIDR |
11 |
9.7090189e-05 |
| 1,785 |
Approximate Lineage for Probabilistic Databases |
2008 |
VLDB |
20 |
9.6438009e-05 |
| 1,813 |
Joins via Geometric Resolutions: Worst-case and Beyond |
2015 |
PODS |
23 |
9.5759542e-05 |
| 2,587 |
Brainwash: A Data System for Feature Engineering |
2013 |
CIDR |
23 |
8.2523942e-05 |
| 2,985 |
Materialized Views in Probabilistic Databases: For Information Exchange and Query Optimization |
2007 |
VLDB |
12 |
7.7794901e-05 |
| 3,063 |
Fonduer: Knowledge Base Construction from Richly Formatted Data |
2018 |
SIGMOD |
12 |
7.689108e-05 |
| 3,112 |
Event Queries on Correlated Probabilistic Streams |
2008 |
SIGMOD |
18 |
7.6351481e-05 |
| 3,521 |
Ember: No-Code Context Enrichment via Similarity-Based Keyless Joins |
2022 |
VLDB |
10 |
7.2320843e-05 |
| 3,627 |
The Role of Massively Multi-Task and Weak Supervision in Software 2.0 |
2019 |
CIDR |
5 |
7.1492314e-05 |
| 3,718 |
Machine Learning and Databases: The Sound of Things to Come or a Cacophony of Hype? |
2015 |
SIGMOD |
7 |
7.0703796e-05 |
| 3,737 |
Snorkel: Fast Training Set Generation for Information Extraction |
2017 |
SIGMOD |
8 |
7.0610219e-05 |
| 3,793 |
SLiMFast: Guaranteed Results for Data Fusion and Source Reliability |
2017 |
SIGMOD |
15 |
7.0156889e-05 |
| 3,909 |
DunceCap: Query Plans Using Generalized Hypertree Decompositions |
2015 |
SIGMOD |
10 |
6.9287689e-05 |
| 3,988 |
Exploiting Correlations for Expensive Predicate Evaluation |
2015 |
SIGMOD |
6 |
6.8694751e-05 |
| 3,998 |
Overton: A Data System for Monitoring and Improving Machine-Learned Products |
2020 |
CIDR |
10 |
6.862274e-05 |
| 4,029 |
Towards High-Throughput Gibbs Sampling at Scale: A Study across Storage Managers |
2013 |
SIGMOD |
9 |
6.840359e-05 |
| 4,395 |
Extracting Databases from Dark Data with DeepDive |
2016 |
SIGMOD |
8 |
6.6175689e-05 |
| 5,182 |
Snorkel DryBell: A Case Study in Deploying Weak Supervision at Industrial Scale |
2019 |
SIGMOD |
9 |
6.2377015e-05 |
| 5,809 |
Incrementally Maintaining Classification using an RDBMS |
2011 |
VLDB |
6 |
5.9848544e-05 |
| 5,922 |
A Relational Framework for Classifier Engineering |
2017 |
PODS |
5 |
5.9442275e-05 |
| 5,974 |
Understanding Cardinality Estimation using Entropy Maximization |
2010 |
PODS |
5 |
5.9279222e-05 |
| 7,026 |
GeoDeepDive: Statistical Inference using Familiar Data-Processing Languages |
2013 |
SIGMOD |
2 |
5.6157971e-05 |
| 7,542 |
Feature Selection in Enterprise Analytics: A Demonstration using an R-based Data Analytics System |
2013 |
VLDB |
9 |
5.4976824e-05 |
| 7,872 |
Mind the Gap: Bridging Multi-Domain Query Workloads with EmptyHeaded |
2017 |
VLDB |
3 |
5.4339723e-05 |
| 8,314 |
DunceCap: Compiling Worst-Case Optimal Query Plans |
2015 |
SIGMOD |
4 |
5.3557545e-05 |
| 8,755 |
Probabilistic Management of OCR Data using an RDBMS |
2012 |
VLDB |
2 |
5.2841044e-05 |
| 8,756 |
A Demonstration of Cascadia Through a Digital Diary Application |
2008 |
SIGMOD |
1 |
5.2830007e-05 |
| 9,771 |
Bootleg: Chasing the Tail with Self-Supervised Named Entity Disambiguation |
2021 |
CIDR |
4 |
5.1320012e-05 |
| 9,838 |
Automating the Enterprise with Foundation Models |
2024 |
VLDB |
1 |
5.1221535e-05 |
| 10,179 |
Is Data Management the Beating Heart of AI Systems? |
2022 |
SIGMOD |
1 |
5.06553e-05 |
| 11,831 |
Data Management Opportunities for Foundation Models |
2022 |
CIDR |
0 |
4.9769913e-05 |
| 12,131 |
Leveraging Organizational Resources to Adapt Models to New Data Modalities |
2020 |
VLDB |
2 |
4.9769913e-05 |
| 12,434 |
Mindtagger: A Demonstration of Data Labeling in Knowledge Base Construction |
2015 |
VLDB |
1 |
4.9769913e-05 |
| 12,705 |
Transducing Markov Sequences |
2010 |
PODS |
3 |
4.9769913e-05 |
| 12,896 |
Systems Aspects of Probabilistic Data Management |
2008 |
VLDB |
0 |
4.9769913e-05 |
| 13,738 |
The DB Community vis-à-vis Environmental, Health, and Societal Grand Challenges: Innovation Engine, Plumber, or Bystander? |
2022 |
SIGMOD |
0 |
- |
| 13,967 |
Ringtail: A Generalized Nowcasting System |
2013 |
VLDB |
1 |
- |
| 14,062 |
Lahar Demonstration: Warehousing Markovian Streams |
2009 |
VLDB |
1 |
- |