Back to authors
Xu Chu
- Author ID
- 16471
- ORCID
-
-
- Links
-
(found by gpt-5.6-luna on jul 24 2026)
- Most Frequent Institution
- Georgia Institute of Technology
- Pagerank
- 0.23338432
- Overall Rank
- 235 | 98.89%
- Paper Count
- 27
Affiliation Timeline
Incoming Non-self Citations Over Time
Total yearly non-self incoming citations across all papers by this author.
Publications by Paper Pagerank
Showing 27 of 27 publications.
| Rank |
Title |
Year |
Venue |
Pagerank |
| 112 |
HoloClean: Holistic Data Repairs with Probabilistic Inference |
2017 |
VLDB |
0.00032801121 |
| 376 |
Discovering Denial Constraints |
2013 |
VLDB |
0.00019677674 |
| 1,101 |
KATARA: A Data Cleaning System Powered by Knowledge Bases and Crowdsourcing |
2015 |
SIGMOD |
0.00012168934 |
| 1,323 |
Data Cleaning: Overview and Emerging Challenges |
2016 |
SIGMOD |
0.00011152602 |
| 1,351 |
Detecting Data Errors: Where are we and what needs to be done? |
2016 |
VLDB |
0.00011064851 |
| 2,147 |
Nearest Neighbor Classifiers over Incomplete Information: From Certain Answers to Certain Predictions |
2021 |
VLDB |
9.0831495e-05 |
| 2,290 |
ZeroER: Entity Resolution using Zero Labeled Examples |
2020 |
SIGMOD |
8.799251e-05 |
| 2,893 |
Distributed Data Deduplication |
2016 |
VLDB |
7.983961e-05 |
| 2,978 |
Transform-Data-by-Example (TDE): An Extensible Search Engine for Data Transformations |
2018 |
VLDB |
7.9030989e-05 |
| 3,574 |
TEGRA: Table Extraction by Global Record Alignment |
2015 |
SIGMOD |
7.2958137e-05 |
| 4,201 |
CLAMS: Bringing Quality to Data Lakes |
2016 |
SIGMOD |
6.8355878e-05 |
| 4,451 |
GOGGLES: Automatic Image Labeling with Affinity Coding |
2020 |
SIGMOD |
6.6952549e-05 |
| 4,525 |
SEMA-JOIN: Joining Semantically-Related Tables Using Big Table Corpora |
2015 |
VLDB |
6.6456999e-05 |
| 4,658 |
OmniFair: A Declarative System for Model-Agnostic Group Fairness in Machine Learning |
2021 |
SIGMOD |
6.5817368e-05 |
| 5,134 |
Auto-FuzzyJoin: Auto-Program Fuzzy Similarity Joins Without Labeled Examples |
2021 |
SIGMOD |
6.3532024e-05 |
| 5,291 |
DiffPrep: Differentiable Data Preprocessing Pipeline Search for Learning over Tabular Data |
2023 |
SIGMOD |
6.2801343e-05 |
| 5,397 |
KATARA: Reliable Data Cleaning with Knowledge Bases and Crowdsourcing |
2015 |
VLDB |
6.232136e-05 |
| 6,041 |
Demonstration of Panda: A Weakly Supervised Entity Matching System |
2021 |
VLDB |
5.9976574e-05 |
| 6,392 |
Qualitative Data Cleaning |
2016 |
VLDB |
5.8883971e-05 |
| 6,854 |
iFlipper: Label Flipping for Individual Fairness |
2023 |
SIGMOD |
5.753578e-05 |
| 7,078 |
Transform-Data-by-Example (TDE): Extensible Data Transformation in Excel |
2018 |
SIGMOD |
5.7095504e-05 |
| 7,290 |
Learning to be a Statistician: Learned Estimator for Number of Distinct Values |
2022 |
VLDB |
5.6540503e-05 |
| 7,387 |
PIClean: A Probabilistic and Interactive Data Cleaning System |
2019 |
SIGMOD |
5.6270554e-05 |
| 9,408 |
SketchQL: Video Moment Querying with a Visual Query Interface |
2024 |
SIGMOD |
5.2750967e-05 |
| 9,421 |
SketchQL Demonstration: Zero-shot Video Moment Querying with Sketches |
2024 |
VLDB |
5.2725196e-05 |
| 9,560 |
Ground Truth Inference for Weakly Supervised Entity Matching |
2023 |
SIGMOD |
5.2528121e-05 |
| 13,311 |
Prompt Editor: A Taxonomy-driven System for Guided LLM Prompt Development in Enterprise Settings |
2025 |
SIGMOD |
- |
Frequent Co-authors
Co-authored at least 5 papers.