| 106 |
The MADlib Analytics Library or MAD Skills, the SQL |
2012 |
VLDB |
0.00033539462 |
| 112 |
HoloClean: Holistic Data Repairs with Probabilistic Inference |
2017 |
VLDB |
0.00032801121 |
| 205 |
Snorkel: Rapid Training Data Creation with Weak Supervision |
2018 |
VLDB |
0.00025235185 |
| 211 |
EmptyHeaded: A Relational Engine for Graph Processing |
2016 |
SIGMOD |
0.00024797217 |
| 411 |
Worst-case Optimal Join Algorithms |
2012 |
PODS |
0.00018902089 |
| 420 |
Can Foundation Models Wrangle Your Data? |
2023 |
VLDB |
0.00018789852 |
| 518 |
Towards a Unified Architecture for in-RDBMS Analytics |
2012 |
SIGMOD |
0.00017167492 |
| 579 |
Incremental Knowledge Base Construction Using DeepDive |
2015 |
VLDB |
0.00016217563 |
| 640 |
Materialization Optimizations for Feature Selection Workloads |
2014 |
SIGMOD |
0.00015409494 |
| 683 |
MYSTIQ: A system for finding more answers by using probabilities |
2005 |
SIGMOD |
0.00015005338 |
| 713 |
Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes |
2024 |
VLDB |
0.00014672521 |
| 1,043 |
Tuffy: Scaling up Statistical Inference in Markov Logic Networks using an RDBMS |
2011 |
VLDB |
0.0001244565 |
| 1,094 |
Snuba: Automating Weak Supervision to Label Training Data |
2019 |
VLDB |
0.00012214617 |
| 1,150 |
DimmWitted: A Study of Main-Memory Statistical Analytics |
2014 |
VLDB |
0.00011943462 |
| 1,257 |
Automatic Optimization for MapReduce Programs |
2011 |
VLDB |
0.000114432 |
| 1,549 |
AJAR: Aggregations and Joins over Annotated Relations |
2016 |
PODS |
0.00010390168 |
| 1,699 |
Beyond Worst-case Analysis for Joins with Minesweeper |
2014 |
PODS |
9.975915e-05 |
| 1,728 |
Structured Querying of Web Text: A Technical Challenge |
2007 |
CIDR |
9.9098087e-05 |
| 1,752 |
Approximate Lineage for Probabilistic Databases |
2008 |
VLDB |
9.8358116e-05 |
| 1,857 |
Joins via Geometric Resolutions: Worst-case and Beyond |
2015 |
PODS |
9.6047945e-05 |
| 2,604 |
Brainwash: A Data System for Feature Engineering |
2013 |
CIDR |
8.3524514e-05 |
| 2,936 |
Materialized Views in Probabilistic Databases: For Information Exchange and Query Optimization |
2007 |
VLDB |
7.9437689e-05 |
| 3,050 |
Event Queries on Correlated Probabilistic Streams |
2008 |
SIGMOD |
7.8140428e-05 |
| 3,192 |
Fonduer: Knowledge Base Construction from Richly Formatted Data |
2018 |
SIGMOD |
7.65035e-05 |
| 3,589 |
Ember: No-Code Context Enrichment via Similarity-Based Keyless Joins |
2022 |
VLDB |
7.2812353e-05 |
| 3,600 |
The Role of Massively Multi-Task and Weak Supervision in Software 2.0 |
2019 |
CIDR |
7.2709969e-05 |
| 3,678 |
Machine Learning and Databases: The Sound of Things to Come or a Cacophony of Hype? |
2015 |
SIGMOD |
7.207554e-05 |
| 3,701 |
Snorkel: Fast Training Set Generation for Information Extraction |
2017 |
SIGMOD |
7.185321e-05 |
| 3,715 |
SLiMFast: Guaranteed Results for Data Fusion and Source Reliability |
2017 |
SIGMOD |
7.1763559e-05 |
| 3,841 |
DunceCap: Query Plans Using Generalized Hypertree Decompositions |
2015 |
SIGMOD |
7.0808098e-05 |
| 3,945 |
Exploiting Correlations for Expensive Predicate Evaluation |
2015 |
SIGMOD |
7.0055154e-05 |
| 3,947 |
Overton: A Data System for Monitoring and Improving Machine-Learned Products |
2020 |
CIDR |
7.0040437e-05 |
| 3,952 |
Towards High-Throughput Gibbs Sampling at Scale: A Study across Storage Managers |
2013 |
SIGMOD |
6.9967272e-05 |
| 4,410 |
Extracting Databases from Dark Data with DeepDive |
2016 |
SIGMOD |
6.717496e-05 |
| 5,058 |
Snorkel DryBell: A Case Study in Deploying Weak Supervision at Industrial Scale |
2019 |
SIGMOD |
6.3815523e-05 |
| 5,695 |
Incrementally Maintaining Classification using an RDBMS |
2011 |
VLDB |
6.1183313e-05 |
| 5,804 |
A Relational Framework for Classifier Engineering |
2017 |
PODS |
6.0829486e-05 |
| 5,860 |
Understanding Cardinality Estimation using Entropy Maximization |
2010 |
PODS |
6.0636893e-05 |
| 6,881 |
GeoDeepDive: Statistical Inference using Familiar Data-Processing Languages |
2013 |
SIGMOD |
5.7469874e-05 |
| 7,403 |
Feature Selection in Enterprise Analytics: A Demonstration using an R-based Data Analytics System |
2013 |
VLDB |
5.6249895e-05 |
| 7,733 |
Mind the Gap: Bridging Multi-Domain Query Workloads with EmptyHeaded |
2017 |
VLDB |
5.5564458e-05 |
| 8,168 |
DunceCap: Compiling Worst-Case Optimal Query Plans |
2015 |
SIGMOD |
5.4743814e-05 |
| 8,585 |
Probabilistic Management of OCR Data using an RDBMS |
2012 |
VLDB |
5.4075213e-05 |
| 8,588 |
A Demonstration of Cascadia Through a Digital Diary Application |
2008 |
SIGMOD |
5.4067872e-05 |
| 9,614 |
Bootleg: Chasing the Tail with Self-Supervised Named Entity Disambiguation |
2021 |
CIDR |
5.2444447e-05 |
| 9,654 |
Automating the Enterprise with Foundation Models |
2024 |
VLDB |
5.2422003e-05 |
| 9,984 |
Is Data Management the Beating Heart of AI Systems? |
2022 |
SIGMOD |
5.1842045e-05 |
| 11,516 |
Data Management Opportunities for Foundation Models |
2022 |
CIDR |
5.093636e-05 |
| 11,824 |
Leveraging Organizational Resources to Adapt Models to New Data Modalities |
2020 |
VLDB |
5.093636e-05 |
| 12,135 |
Mindtagger: A Demonstration of Data Labeling in Knowledge Base Construction |
2015 |
VLDB |
5.093636e-05 |