| 265 |
CrowdER: Crowdsourcing Entity Resolution |
2012 |
VLDB |
0.00029904018 |
| 788 |
ActiveClean: Interactive Data Cleaning For Statistical Modeling |
2016 |
VLDB |
0.00016618698 |
| 863 |
Leveraging Transitive Relations for Crowdsourced Joins |
2013 |
SIGMOD |
0.00015793243 |
| 1,338 |
Entity Matching: How Similar Is Similar |
2011 |
VLDB |
0.00012501156 |
| 1,396 |
Can We Beat the Prefix Filtering? An Adaptive Framework for Similarity Join and Search |
2012 |
SIGMOD |
0.00012215253 |
| 1,629 |
Data Cleaning: Overview and Emerging Challenges |
2016 |
SIGMOD |
0.00011073148 |
| 1,699 |
Are We Ready For Learned Cardinality Estimation? |
2021 |
VLDB |
0.00010848882 |
| 2,177 |
A Sample-and-Clean Framework for Fast and Accurate Query Processing on Dirty Data |
2014 |
SIGMOD |
9.371335e-05 |
| 2,588 |
Pass-Join: A Partition-based Method for Similarity Joins |
2012 |
VLDB |
8.4872437e-05 |
| 2,759 |
Complaint-driven Training Data Debugging for Query 2.0 |
2020 |
SIGMOD |
8.1646193e-05 |
| 3,198 |
Towards Dependable Data Repairing with Fixing Rules |
2014 |
SIGMOD |
7.4029546e-05 |
| 3,268 |
QASCA: A Quality-Aware Task Assignment System for Crowdsourcing Applications |
2015 |
SIGMOD |
7.3027561e-05 |
| 3,303 |
SCODED: Statistical Constraint Oriented Data Error Detection |
2020 |
SIGMOD |
7.2438671e-05 |
| 3,731 |
Skipping-oriented Partitioning for Columnar Layouts |
2017 |
VLDB |
6.8074069e-05 |
| 3,767 |
Cleaning Crowdsourced Labels Using Oracles for Statistical Classification |
2019 |
VLDB |
6.7748725e-05 |
| 3,944 |
AQP++: Connecting Approximate Query Processing With Aggregate Precomputation for Interactive Analytics |
2018 |
SIGMOD |
6.6056349e-05 |
| 4,215 |
Trie-Join: Efficient Trie-based String Similarity Joins with Edit-Distance Constraints |
2010 |
VLDB |
6.3464157e-05 |
| 4,453 |
CLAMShell: Speeding up Crowds for Low-latency Data Labeling |
2016 |
VLDB |
6.1690121e-05 |
| 4,664 |
PrivateClean: Data Cleaning and Differential Privacy |
2016 |
SIGMOD |
6.0058132e-05 |
| 5,227 |
Enabling SQL-based Training Data Debugging for Federated Learning |
2022 |
VLDB |
5.6156523e-05 |
| 5,930 |
ActiveClean: An Interactive Data Cleaning Framework For Modern Machine Learning |
2016 |
SIGMOD |
5.2632185e-05 |
| 5,984 |
DataPrep.EDA: Task-Centric Exploratory Data Analysis for Statistical Modeling in Python |
2021 |
SIGMOD |
5.2400405e-05 |
| 6,539 |
ConnectorX: Accelerating Data Loading From Databases to Dataframes |
2022 |
VLDB |
5.0168759e-05 |
| 6,779 |
Explaining Inference Queries with Bayesian Optimization |
2021 |
VLDB |
4.9232829e-05 |
| 6,854 |
DBease: Making Databases User-friendly and Easily Accessible |
2011 |
CIDR |
4.901546e-05 |
| 7,116 |
Crowdsourced Data Management: Overview and Challenges |
2017 |
SIGMOD |
4.8219732e-05 |
| 8,590 |
Wisteria: Nurturing Scalable Data Cleaning Infrastructure |
2015 |
VLDB |
4.4851741e-05 |
| 8,642 |
One Size Does Not Fit All: A Bandit-Based Sampler Combination Framework with Theoretical Guarantees |
2022 |
SIGMOD |
4.4734993e-05 |
| 8,674 |
Progressive Deep Web Crawling Through Keyword Queries For Data Enrichment |
2019 |
SIGMOD |
4.4659264e-05 |
| 8,703 |
Stale View Cleaning: Getting Fresh Answers from Stale Materialized Views |
2015 |
VLDB |
4.4596255e-05 |
| 8,853 |
Complaint-Driven Training Data Debugging at Interactive Speeds |
2022 |
SIGMOD |
4.4308213e-05 |
| 9,278 |
ActiveDeeper: A Model-based Active Data Enrichment System |
2020 |
VLDB |
4.3607776e-05 |
| 10,115 |
ST-Raptor: LLM-Powered Semi-Structured Table Question Answering |
2026 |
SIGMOD |
4.1905499e-05 |
| 10,591 |
A Flexible Framework for Query-oriented Interactive Community Search |
2025 |
VLDB |
4.1905499e-05 |
| 10,599 |
Accio: Bolt-on Query Federation |
2025 |
VLDB |
4.1905499e-05 |
| 10,768 |
ParSEval: Plan-aware Test Database Generation for SQL Equivalence Evaluation |
2025 |
VLDB |
4.1905499e-05 |
| 11,728 |
Deeper: A Data Enrichment System Powered by Deep Web |
2018 |
SIGMOD |
4.1905499e-05 |
| 13,208 |
Web Connector: A Unified API Wrapper to Simplify Web Data Collection |
2023 |
VLDB |
- |