| 104 |
HoloClean: Holistic Data Repairs with Probabilistic Inference |
2017 |
VLDB |
0.00033690989 |
| 105 |
The MADlib Analytics Library or MAD Skills, the SQL |
2012 |
VLDB |
0.00033638251 |
| 205 |
Snorkel: Rapid Training Data Creation with Weak Supervision |
2018 |
VLDB |
0.00025181304 |
| 208 |
EmptyHeaded: A Relational Engine for Graph Processing |
2016 |
SIGMOD |
0.00024884544 |
| 329 |
Can Foundation Models Wrangle Your Data? |
2023 |
VLDB |
0.00020858443 |
| 402 |
Worst-case Optimal Join Algorithms |
2012 |
PODS |
0.00019104625 |
| 501 |
Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes |
2024 |
VLDB |
0.00017267905 |
| 503 |
Towards a Unified Architecture for in-RDBMS Analytics |
2012 |
SIGMOD |
0.00017202276 |
| 579 |
Incremental Knowledge Base Construction Using DeepDive |
2015 |
VLDB |
0.00016086569 |
| 654 |
Materialization Optimizations for Feature Selection Workloads |
2014 |
SIGMOD |
0.0001510357 |
| 696 |
MYSTIQ: A system for finding more answers by using probabilities |
2005 |
SIGMOD |
0.00014697176 |
| 1,062 |
Tuffy: Scaling up Statistical Inference in Markov Logic Networks using an RDBMS |
2011 |
VLDB |
0.0001221201 |
| 1,120 |
Snuba: Automating Weak Supervision to Label Training Data |
2019 |
VLDB |
0.00011946047 |
| 1,167 |
DimmWitted: A Study of Main-Memory Statistical Analytics |
2014 |
VLDB |
0.00011729888 |
| 1,281 |
Automatic Optimization for MapReduce Programs |
2011 |
VLDB |
0.00011213384 |
| 1,570 |
AJAR: Aggregations and Joins over Annotated Relations |
2016 |
PODS |
0.0001020855 |
| 1,725 |
Beyond Worst-case Analysis for Joins with Minesweeper |
2014 |
PODS |
9.7902443e-05 |
| 1,760 |
Structured Querying of Web Text: A Technical Challenge |
2007 |
CIDR |
9.7106555e-05 |
| 1,785 |
Approximate Lineage for Probabilistic Databases |
2008 |
VLDB |
9.6482655e-05 |
| 1,812 |
Joins via Geometric Resolutions: Worst-case and Beyond |
2015 |
PODS |
9.5803973e-05 |
| 2,586 |
Brainwash: A Data System for Feature Engineering |
2013 |
CIDR |
8.2560072e-05 |
| 2,984 |
Materialized Views in Probabilistic Databases: For Information Exchange and Query Optimization |
2007 |
VLDB |
7.7830089e-05 |
| 3,061 |
Fonduer: Knowledge Base Construction from Richly Formatted Data |
2018 |
SIGMOD |
7.6927483e-05 |
| 3,109 |
Event Queries on Correlated Probabilistic Streams |
2008 |
SIGMOD |
7.6387635e-05 |
| 3,521 |
Ember: No-Code Context Enrichment via Similarity-Based Keyless Joins |
2022 |
VLDB |
7.2351481e-05 |
| 3,627 |
The Role of Massively Multi-Task and Weak Supervision in Software 2.0 |
2019 |
CIDR |
7.1514557e-05 |
| 3,716 |
Machine Learning and Databases: The Sound of Things to Come or a Cacophony of Hype? |
2015 |
SIGMOD |
7.0726435e-05 |
| 3,736 |
Snorkel: Fast Training Set Generation for Information Extraction |
2017 |
SIGMOD |
7.0642592e-05 |
| 3,791 |
SLiMFast: Guaranteed Results for Data Fusion and Source Reliability |
2017 |
SIGMOD |
7.0189865e-05 |
| 3,908 |
DunceCap: Query Plans Using Generalized Hypertree Decompositions |
2015 |
SIGMOD |
6.9320489e-05 |
| 3,988 |
Exploiting Correlations for Expensive Predicate Evaluation |
2015 |
SIGMOD |
6.871854e-05 |
| 3,997 |
Overton: A Data System for Monitoring and Improving Machine-Learned Products |
2020 |
CIDR |
6.8655222e-05 |
| 4,029 |
Towards High-Throughput Gibbs Sampling at Scale: A Study across Storage Managers |
2013 |
SIGMOD |
6.8419186e-05 |
| 4,393 |
Extracting Databases from Dark Data with DeepDive |
2016 |
SIGMOD |
6.620281e-05 |
| 5,181 |
Snorkel DryBell: A Case Study in Deploying Weak Supervision at Industrial Scale |
2019 |
SIGMOD |
6.2406557e-05 |
| 5,813 |
Incrementally Maintaining Classification using an RDBMS |
2011 |
VLDB |
5.9861029e-05 |
| 5,921 |
A Relational Framework for Classifier Engineering |
2017 |
PODS |
5.9470387e-05 |
| 5,973 |
Understanding Cardinality Estimation using Entropy Maximization |
2010 |
PODS |
5.9306716e-05 |
| 7,025 |
GeoDeepDive: Statistical Inference using Familiar Data-Processing Languages |
2013 |
SIGMOD |
5.6184467e-05 |
| 7,536 |
Feature Selection in Enterprise Analytics: A Demonstration using an R-based Data Analytics System |
2013 |
VLDB |
5.5002435e-05 |
| 7,867 |
Mind the Gap: Bridging Multi-Domain Query Workloads with EmptyHeaded |
2017 |
VLDB |
5.4365458e-05 |
| 8,308 |
DunceCap: Compiling Worst-Case Optimal Query Plans |
2015 |
SIGMOD |
5.3582899e-05 |
| 8,747 |
Probabilistic Management of OCR Data using an RDBMS |
2012 |
VLDB |
5.2865969e-05 |
| 8,748 |
A Demonstration of Cascadia Through a Digital Diary Application |
2008 |
SIGMOD |
5.2855025e-05 |
| 9,766 |
Bootleg: Chasing the Tail with Self-Supervised Named Entity Disambiguation |
2021 |
CIDR |
5.1344318e-05 |
| 9,831 |
Automating the Enterprise with Foundation Models |
2024 |
VLDB |
5.1245795e-05 |
| 10,175 |
Is Data Management the Beating Heart of AI Systems? |
2022 |
SIGMOD |
5.0679291e-05 |
| 11,825 |
Data Management Opportunities for Foundation Models |
2022 |
CIDR |
4.9793485e-05 |
| 12,125 |
Leveraging Organizational Resources to Adapt Models to New Data Modalities |
2020 |
VLDB |
4.9793485e-05 |
| 12,428 |
Mindtagger: A Demonstration of Data Labeling in Knowledge Base Construction |
2015 |
VLDB |
4.9793485e-05 |