Back to authors
Nan Tang
- Author ID
- 1257
- ORCID
-
0000-0003-2832-0295
- Links
-
(found by gpt-5.6-luna on jul 24 2026)
- Most Frequent Institution
- Qatar Computing Research Institute
- Pagerank
- 0.58891663
- Overall Rank
- 42 | 99.81%
- Paper Count
- 72
Affiliation Timeline
Incoming Non-self Citations Over Time
Total yearly non-self incoming citations across all papers by this author.
Publications by Paper Pagerank
Showing 22 of 72 publications.
| Rank |
Title |
Year |
Venue |
Pagerank |
| 7,613 |
LakeCompass: An End-to-End System for Data Maintenance, Search and Analysis in Data Lakes |
2024 |
VLDB |
5.5817761e-05 |
| 7,893 |
LakeBench: A Benchmark for Discovering Joinable and Unionable Tables in Data Lakes |
2024 |
VLDB |
5.5209866e-05 |
| 8,177 |
HAIPipe: Combining Human-generated and Machine-generated Pipelines for Data Preparation |
2023 |
SIGMOD |
5.4730821e-05 |
| 8,242 |
Data Civilizer 2.0: A Holistic Framework for Data Preparation and Analytics |
2019 |
VLDB |
5.459288e-05 |
| 8,336 |
DADER: Hands-Off Entity Resolution with Domain Adaptation |
2022 |
VLDB |
5.451516e-05 |
| 9,030 |
CerFix: A System for Cleaning Data with Certain Fixes |
2011 |
VLDB |
5.3286356e-05 |
| 9,177 |
VerifAI: Verified Generative AI |
2024 |
CIDR |
5.3078984e-05 |
| 9,298 |
VisClean: Interactive Cleaning for Progressive Visualization |
2020 |
VLDB |
5.2901384e-05 |
| 9,442 |
Interactive and Deterministic Data Cleaning: A Tossed Stone Raises a Thousand Ripples |
2016 |
SIGMOD |
5.2677992e-05 |
| 9,495 |
Debugging Large-Scale Data Science Pipelines using Dagger |
2020 |
VLDB |
5.2616335e-05 |
| 9,550 |
Data Imputation with Limited Data Redundancy Using Data Lakes |
2025 |
VLDB |
5.2528121e-05 |
| 9,653 |
CoClean: Collaborative Data Cleaning |
2020 |
SIGMOD |
5.2425585e-05 |
| 9,870 |
Natural Language to SQL: State of the Art and Open Problems |
2025 |
VLDB |
5.2043672e-05 |
| 9,983 |
Rheem: Enabling Multi-Platform Task Execution |
2016 |
SIGMOD |
5.1844786e-05 |
| 10,587 |
LEAD: Iterative Data Selection for Efficient LLM Instruction Tuning |
2026 |
VLDB |
5.093636e-05 |
| 10,705 |
Andromeda: Debugging Database Performance Issues with Retrieval-Augmented Large Language Models |
2025 |
SIGMOD |
5.093636e-05 |
| 10,867 |
Weak-to-Strong Prompts with Lightweight-to-Powerful LLMs for High-Accuracy, Low-Cost, and Explainable Data Transformation |
2025 |
VLDB |
5.093636e-05 |
| 10,931 |
AutoPrep: Natural Language Question-Aware Data Preparation with a Multi-Agent Framework |
2025 |
VLDB |
5.093636e-05 |
| 11,211 |
MisDetect: Iterative Mislabel Detection using Early Loss |
2024 |
VLDB |
5.093636e-05 |
| 11,778 |
Interactively Discovering and Ranking Desired Tuples without Writing SQL Queries |
2020 |
SIGMOD |
5.093636e-05 |
| 13,485 |
DeepTrack: Monitoring and Exploring Spatio-Temporal Data - A Case of Tracking COVID-19 - |
2020 |
VLDB |
- |
| 13,541 |
Errata for “Lightning Fast and Space Efficient Inequality Joins” (PVLDB 8(13): 2074-2085) |
2017 |
VLDB |
- |
Frequent Co-authors
Co-authored at least 5 papers.