DBScholar

Back to authors

Ihab F. Ilyas

Author ID
u794
ORCID
-
Links
(found by gpt-5.6-luna on jul 24 2026)
Most Frequent Institution
University of Waterloo
Pagerank
0.4962917
Overall Rank
64 | 99.71%
Paper Count
54

Affiliation Timeline

Incoming Non-self Citations Over Time

Total yearly non-self incoming citations across all papers by this author.

Publications by Paper Pagerank

Showing all 54 publications. Total citations include self and non-self citations.

Rank Title Year Venue Total Citations Pagerank
104 HoloClean: Holistic Data Repairs with Probabilistic Inference 2017 VLDB 143 0.00033676943
160 CORDS: Automatic Discovery of Correlations and Soft Functional Dependencies 2004 SIGMOD 80 0.00027827605
350 Discovering Denial Constraints 2013 VLDB 75 0.00020244085
514 Data Curation at Scale: The Data Tamer System 2013 CIDR 38 0.00016999652
524 Supporting Top-k Join Queries in Relational Databases 2003 VLDB 51 0.00016902116
697 NADEEF: A Commodity Data Cleaning System 2013 SIGMOD 58 0.00014687805
716 Guided Data Repair 2011 VLDB 48 0.00014546968
884 HoloDetect: Few-Shot Learning for Error Detection 2019 SIGMOD 45 0.00013263269
962 RankSQL: Query Algebra and Optimization for Relational Top-k Queries 2005 SIGMOD 36 0.00012818013
972 The Data Civilizer System 2017 CIDR 56 0.00012757732
1,044 Data Cleaning: Overview and Emerging Challenges 2016 SIGMOD 37 0.00012329478
1,098 KATARA: A Data Cleaning System Powered by Knowledge Bases and Crowdsourcing 2015 SIGMOD 48 0.00012031983
1,344 Detecting Data Errors: Where are we and what needs to be done? 2016 VLDB 50 0.00010951939
1,348 Sampling the Repairs of Functional Dependency Violations under Hard Constraints 2010 VLDB 24 0.00010942213
1,373 High-Throughput Vector Similarity Search in Knowledge Graphs 2023 SIGMOD 30 0.00010891169
1,635 Efficient Search for the Top-k Probable Nearest Neighbors in Uncertain Databases 2008 VLDB 17 0.00010012676
1,773 Rank-aware Query Optimization 2004 SIGMOD 21 9.6683429e-05
2,424 BigDansing: A System for Big Data Cleansing 2015 SIGMOD 34 8.483813e-05
2,445 Approximate Denial Constraints 2020 VLDB 24 8.4546614e-05
2,621 Ranking with Uncertain Scoring Functions: Semantics and Sensitivity Measures 2011 SIGMOD 16 8.2130891e-05
2,882 Expressive and Flexible Access to Web-Extracted Data: A Keyword-based Structured Query Language 2010 SIGMOD 5 7.9108026e-05
2,949 Distributed Data Deduplication 2016 VLDB 16 7.8193962e-05
3,015 Saga: A Platform for Continuous Construction and Serving of Knowledge At Scale 2022 SIGMOD 12 7.7504975e-05
3,051 Joining Ranked Inputs in Practice 2002 VLDB 10 7.7058408e-05
3,292 Kamino: Constraint-Aware Differentially Private Data Synthesis 2021 VLDB 14 7.445638e-05
3,426 Modeling and Querying Possible Repairs in Duplicate Detection 2009 VLDB 9 7.3053186e-05
3,521 Ember: No-Code Context Enrichment via Similarity-Based Keyless Joins 2022 VLDB 10 7.2320843e-05
3,634 NADEEF/ER: Generic and Interactive Entity Resolution 2014 SIGMOD 7 7.1454896e-05
3,784 Supporting Ad-hoc Ranking Aggregates 2006 SIGMOD 8 7.0197908e-05
4,020 CLAMS: Bringing Quality to Data Lakes 2016 SIGMOD 9 6.8481527e-05
4,242 Creating Competitive Products 2009 VLDB 7 6.705389e-05
4,417 Estimating Compilation Time of a Query Optimizer 2003 SIGMOD 6 6.6064397e-05
4,587 DataXFormer: An Interactive Data Transformation Tool 2015 SIGMOD 4 6.5148596e-05
4,904 Growing and Serving Large Open-domain Knowledge Graphs 2023 SIGMOD 4 6.3605903e-05
5,000 Descriptive and Prescriptive Data Cleaning 2014 SIGMOD 12 6.3194405e-05
5,133 Distributed implementations of dependency discovery algorithms 2019 VLDB 9 6.2573449e-05
5,185 APEx: Accuracy-Aware Differentially Private Data Exploration 2019 SIGMOD 15 6.2362923e-05
5,260 Top-k Nearest Neighbor Search In Uncertain Data Series 2015 VLDB 14 6.2041665e-05
5,515 KATARA: Reliable Data Cleaning with Knowledge Bases and Crowdsourcing 2015 VLDB 9 6.0962265e-05
5,598 StatAdvisor: Recommending Statistical Views 2009 VLDB 7 6.0690976e-05
5,703 A Demo of the Data Civilizer System 2017 SIGMOD 7 6.02892e-05
5,866 Properties of Inconsistency Measures for Databases 2021 SIGMOD 9 5.96357e-05
5,941 RankSQL: Supporting Ranking Queries in Relational Database Management Systems 2005 VLDB 4 5.937273e-05
6,139 DataXFormer: Leveraging the Web for Semantic Transformations 2015 CIDR 5 5.870906e-05
6,154 NADEEF: A Generalized Data Cleaning System 2013 VLDB 13 5.8665355e-05
6,384 Qualitative Data Cleaning 2016 VLDB 4 5.8039496e-05
8,180 CORDS: Automatic Generation of Correlation Statistics in DB2 2004 VLDB 2 5.3814848e-05
8,520 URank: Formulation and Efficient Evaluation of Top-k Queries in Uncertain Databases 2007 SIGMOD 2 5.3227191e-05
9,686 FIX: Feature-based Indexing Technique for XML Documents 2006 VLDB 1 5.1406509e-05
10,022 PCOR: Private Contextual Outlier Release via Differentially Private Search 2021 SIGMOD 1 5.0932019e-05
12,609 Just-in-Time Information Extraction using Extraction Views 2012 SIGMOD 0 4.9769913e-05
12,748 QUICK: Expressive and Flexible Search over Knowledge Bases and Text Collections 2010 VLDB 0 4.9769913e-05
12,772 Building Ranked Mashups of Unstructured Sources with Uncertain Information 2010 VLDB 0 4.9769913e-05
13,950 We are Drowning in a Sea of Least Publishable Units (LPUs) 2013 SIGMOD 0 -

Frequent Co-authors

Co-authored at least 5 papers.

Co-author Shared Papers Rank Pagerank
Mourad Ouzzani 13 123 0.35665042
Nan Tang 9 41 0.5967
Ahmed K. Elmagarmid 9 219 0.24491915
Xu Chu 9 237 0.23369665
Paolo Papotti 8 144 0.32992914
Mike Stonebraker 7 7 1.0496223
Mohamed Soliman 7 570 0.11361618
Theodoros Rekatsinas 6 319 0.18964601
Ziawasch Abedjan 5 339 0.179185
Jeffrey Pound 5 1,101 0.067125792