| 104 |
HoloClean: Holistic Data Repairs with Probabilistic Inference |
2017 |
VLDB |
143 |
0.00033676943 |
| 160 |
CORDS: Automatic Discovery of Correlations and Soft Functional Dependencies |
2004 |
SIGMOD |
80 |
0.00027827605 |
| 350 |
Discovering Denial Constraints |
2013 |
VLDB |
75 |
0.00020244085 |
| 514 |
Data Curation at Scale: The Data Tamer System |
2013 |
CIDR |
38 |
0.00016999652 |
| 524 |
Supporting Top-k Join Queries in Relational Databases |
2003 |
VLDB |
51 |
0.00016902116 |
| 697 |
NADEEF: A Commodity Data Cleaning System |
2013 |
SIGMOD |
58 |
0.00014687805 |
| 716 |
Guided Data Repair |
2011 |
VLDB |
48 |
0.00014546968 |
| 884 |
HoloDetect: Few-Shot Learning for Error Detection |
2019 |
SIGMOD |
45 |
0.00013263269 |
| 962 |
RankSQL: Query Algebra and Optimization for Relational Top-k Queries |
2005 |
SIGMOD |
36 |
0.00012818013 |
| 972 |
The Data Civilizer System |
2017 |
CIDR |
56 |
0.00012757732 |
| 1,044 |
Data Cleaning: Overview and Emerging Challenges |
2016 |
SIGMOD |
37 |
0.00012329478 |
| 1,098 |
KATARA: A Data Cleaning System Powered by Knowledge Bases and Crowdsourcing |
2015 |
SIGMOD |
48 |
0.00012031983 |
| 1,344 |
Detecting Data Errors: Where are we and what needs to be done? |
2016 |
VLDB |
50 |
0.00010951939 |
| 1,348 |
Sampling the Repairs of Functional Dependency Violations under Hard Constraints |
2010 |
VLDB |
24 |
0.00010942213 |
| 1,373 |
High-Throughput Vector Similarity Search in Knowledge Graphs |
2023 |
SIGMOD |
30 |
0.00010891169 |
| 1,635 |
Efficient Search for the Top-k Probable Nearest Neighbors in Uncertain Databases |
2008 |
VLDB |
17 |
0.00010012676 |
| 1,773 |
Rank-aware Query Optimization |
2004 |
SIGMOD |
21 |
9.6683429e-05 |
| 2,424 |
BigDansing: A System for Big Data Cleansing |
2015 |
SIGMOD |
34 |
8.483813e-05 |
| 2,445 |
Approximate Denial Constraints |
2020 |
VLDB |
24 |
8.4546614e-05 |
| 2,621 |
Ranking with Uncertain Scoring Functions: Semantics and Sensitivity Measures |
2011 |
SIGMOD |
16 |
8.2130891e-05 |
| 2,882 |
Expressive and Flexible Access to Web-Extracted Data: A Keyword-based Structured Query Language |
2010 |
SIGMOD |
5 |
7.9108026e-05 |
| 2,949 |
Distributed Data Deduplication |
2016 |
VLDB |
16 |
7.8193962e-05 |
| 3,015 |
Saga: A Platform for Continuous Construction and Serving of Knowledge At Scale |
2022 |
SIGMOD |
12 |
7.7504975e-05 |
| 3,051 |
Joining Ranked Inputs in Practice |
2002 |
VLDB |
10 |
7.7058408e-05 |
| 3,292 |
Kamino: Constraint-Aware Differentially Private Data Synthesis |
2021 |
VLDB |
14 |
7.445638e-05 |
| 3,426 |
Modeling and Querying Possible Repairs in Duplicate Detection |
2009 |
VLDB |
9 |
7.3053186e-05 |
| 3,521 |
Ember: No-Code Context Enrichment via Similarity-Based Keyless Joins |
2022 |
VLDB |
10 |
7.2320843e-05 |
| 3,634 |
NADEEF/ER: Generic and Interactive Entity Resolution |
2014 |
SIGMOD |
7 |
7.1454896e-05 |
| 3,784 |
Supporting Ad-hoc Ranking Aggregates |
2006 |
SIGMOD |
8 |
7.0197908e-05 |
| 4,020 |
CLAMS: Bringing Quality to Data Lakes |
2016 |
SIGMOD |
9 |
6.8481527e-05 |
| 4,242 |
Creating Competitive Products |
2009 |
VLDB |
7 |
6.705389e-05 |
| 4,417 |
Estimating Compilation Time of a Query Optimizer |
2003 |
SIGMOD |
6 |
6.6064397e-05 |
| 4,587 |
DataXFormer: An Interactive Data Transformation Tool |
2015 |
SIGMOD |
4 |
6.5148596e-05 |
| 4,904 |
Growing and Serving Large Open-domain Knowledge Graphs |
2023 |
SIGMOD |
4 |
6.3605903e-05 |
| 5,000 |
Descriptive and Prescriptive Data Cleaning |
2014 |
SIGMOD |
12 |
6.3194405e-05 |
| 5,133 |
Distributed implementations of dependency discovery algorithms |
2019 |
VLDB |
9 |
6.2573449e-05 |
| 5,185 |
APEx: Accuracy-Aware Differentially Private Data Exploration |
2019 |
SIGMOD |
15 |
6.2362923e-05 |
| 5,260 |
Top-k Nearest Neighbor Search In Uncertain Data Series |
2015 |
VLDB |
14 |
6.2041665e-05 |
| 5,515 |
KATARA: Reliable Data Cleaning with Knowledge Bases and Crowdsourcing |
2015 |
VLDB |
9 |
6.0962265e-05 |
| 5,598 |
StatAdvisor: Recommending Statistical Views |
2009 |
VLDB |
7 |
6.0690976e-05 |
| 5,703 |
A Demo of the Data Civilizer System |
2017 |
SIGMOD |
7 |
6.02892e-05 |
| 5,866 |
Properties of Inconsistency Measures for Databases |
2021 |
SIGMOD |
9 |
5.96357e-05 |
| 5,941 |
RankSQL: Supporting Ranking Queries in Relational Database Management Systems |
2005 |
VLDB |
4 |
5.937273e-05 |
| 6,139 |
DataXFormer: Leveraging the Web for Semantic Transformations |
2015 |
CIDR |
5 |
5.870906e-05 |
| 6,154 |
NADEEF: A Generalized Data Cleaning System |
2013 |
VLDB |
13 |
5.8665355e-05 |
| 6,384 |
Qualitative Data Cleaning |
2016 |
VLDB |
4 |
5.8039496e-05 |
| 8,180 |
CORDS: Automatic Generation of Correlation Statistics in DB2 |
2004 |
VLDB |
2 |
5.3814848e-05 |
| 8,520 |
URank: Formulation and Efficient Evaluation of Top-k Queries in Uncertain Databases |
2007 |
SIGMOD |
2 |
5.3227191e-05 |
| 9,686 |
FIX: Feature-based Indexing Technique for XML Documents |
2006 |
VLDB |
1 |
5.1406509e-05 |
| 10,022 |
PCOR: Private Contextual Outlier Release via Differentially Private Search |
2021 |
SIGMOD |
1 |
5.0932019e-05 |
| 12,609 |
Just-in-Time Information Extraction using Extraction Views |
2012 |
SIGMOD |
0 |
4.9769913e-05 |
| 12,748 |
QUICK: Expressive and Flexible Search over Knowledge Bases and Text Collections |
2010 |
VLDB |
0 |
4.9769913e-05 |
| 12,772 |
Building Ranked Mashups of Unstructured Sources with Uncertain Information |
2010 |
VLDB |
0 |
4.9769913e-05 |
| 13,950 |
We are Drowning in a Sea of Least Publishable Units (LPUs) |
2013 |
SIGMOD |
0 |
- |