| 196 |
CrowdER: Crowdsourcing Entity Resolution |
2012 |
VLDB |
0.00025780596 |
| 582 |
ActiveClean: Interactive Data Cleaning For Statistical Modeling |
2016 |
VLDB |
0.00016148948 |
| 852 |
Leveraging Transitive Relations for Crowdsourced Joins |
2013 |
SIGMOD |
0.00013604253 |
| 975 |
Can We Beat the Prefix Filtering? An Adaptive Framework for Similarity Join and Search |
2012 |
SIGMOD |
0.00012870645 |
| 1,061 |
Are We Ready For Learned Cardinality Estimation? |
2021 |
VLDB |
0.00012369764 |
| 1,248 |
Entity Matching: How Similar Is Similar |
2011 |
VLDB |
0.00011498301 |
| 1,323 |
Data Cleaning: Overview and Emerging Challenges |
2016 |
SIGMOD |
0.00011152602 |
| 1,736 |
A Sample-and-Clean Framework for Fast and Accurate Query Processing on Dirty Data |
2014 |
SIGMOD |
9.8984415e-05 |
| 1,886 |
Pass-Join: A Partition-based Method for Similarity Joins |
2012 |
VLDB |
9.5358137e-05 |
| 2,584 |
Complaint-driven Training Data Debugging for Query 2.0 |
2020 |
SIGMOD |
8.3783546e-05 |
| 2,981 |
Towards Dependable Data Repairing with Fixing Rules |
2014 |
SIGMOD |
7.8960114e-05 |
| 3,004 |
SCODED: Statistical Constraint Oriented Data Error Detection |
2020 |
SIGMOD |
7.8608629e-05 |
| 3,106 |
Skipping-oriented Partitioning for Columnar Layouts |
2017 |
VLDB |
7.7515666e-05 |
| 3,281 |
Cleaning Crowdsourced Labels Using Oracles for Statistical Classification |
2019 |
VLDB |
7.5706653e-05 |
| 3,366 |
AQP++: Connecting Approximate Query Processing With Aggregate Precomputation for Interactive Analytics |
2018 |
SIGMOD |
7.4748604e-05 |
| 3,731 |
Trie-Join: Efficient Trie-based String Similarity Joins with Edit-Distance Constraints |
2010 |
VLDB |
7.1663952e-05 |
| 3,904 |
QASCA: A Quality-Aware Task Assignment System for Crowdsourcing Applications |
2015 |
SIGMOD |
7.0304212e-05 |
| 3,971 |
CLAMShell: Speeding up Crowds for Low-latency Data Labeling |
2016 |
VLDB |
6.9835263e-05 |
| 4,837 |
PrivateClean: Data Cleaning and Differential Privacy |
2016 |
SIGMOD |
6.4845444e-05 |
| 5,131 |
Enabling SQL-based Training Data Debugging for Federated Learning |
2022 |
VLDB |
6.3537809e-05 |
| 5,620 |
DataPrep.EDA: Task-Centric Exploratory Data Analysis for Statistical Modeling in Python |
2021 |
SIGMOD |
6.1482864e-05 |
| 5,794 |
ActiveClean: An Interactive Data Cleaning Framework For Modern Machine Learning |
2016 |
SIGMOD |
6.0863045e-05 |
| 6,628 |
ConnectorX: Accelerating Data Loading From Databases to Dataframes |
2022 |
VLDB |
5.8197537e-05 |
| 6,862 |
DBease: Making Databases User-friendly and Easily Accessible |
2011 |
CIDR |
5.7514733e-05 |
| 6,884 |
Explaining Inference Queries with Bayesian Optimization |
2021 |
VLDB |
5.7465524e-05 |
| 7,357 |
Crowdsourced Data Management: Overview and Challenges |
2017 |
SIGMOD |
5.6346837e-05 |
| 8,608 |
One Size Does Not Fit All: A Bandit-Based Sampler Combination Framework with Theoretical Guarantees |
2022 |
SIGMOD |
5.4024561e-05 |
| 8,620 |
Wisteria: Nurturing Scalable Data Cleaning Infrastructure |
2015 |
VLDB |
5.3991092e-05 |
| 8,705 |
Progressive Deep Web Crawling Through Keyword Queries For Data Enrichment |
2019 |
SIGMOD |
5.3807913e-05 |
| 8,714 |
Stale View Cleaning: Getting Fresh Answers from Stale Materialized Views |
2015 |
VLDB |
5.3778009e-05 |
| 8,859 |
Complaint-Driven Training Data Debugging at Interactive Speeds |
2022 |
SIGMOD |
5.356561e-05 |
| 9,348 |
ActiveDeeper: A Model-based Active Data Enrichment System |
2020 |
VLDB |
5.2843006e-05 |
| 9,691 |
A Flexible Framework for Query-oriented Interactive Community Search |
2025 |
VLDB |
5.2351259e-05 |
| 9,945 |
ParSEval: Plan-aware Test Database Generation for SQL Equivalence Evaluation |
2025 |
VLDB |
5.1915905e-05 |
| 10,403 |
ST-Raptor: LLM-Powered Semi-Structured Table Question Answering |
2026 |
SIGMOD |
5.093636e-05 |
| 10,852 |
Accio: Bolt-on Query Federation |
2025 |
VLDB |
5.093636e-05 |
| 11,927 |
Deeper: A Data Enrichment System Powered by Deep Web |
2018 |
SIGMOD |
5.093636e-05 |
| 13,397 |
Web Connector: A Unified API Wrapper to Simplify Web Data Collection |
2023 |
VLDB |
- |