DBScholar

Back to papers

Leveraging Aggregate Constraints For Deduplication

Summary: Leveraging aggregate constraints (vs. pairwise) to improve cross-source deduplication in data integration. Defines a restricted search space, solves optimally within it, and shows substantial accuracy gains on real data despite semantic and computational challenges. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
3933
Venue
SIGMOD
Year
2007
Pagerank
8.6126774e-05
Overall Rank
2,410 | 83.47%
DOI
10.1145/1247480.1247530

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{chaudhuri_sigmod07,
        title = {{Leveraging Aggregate Constraints For Deduplication}},
        author = {Chaudhuri, Surajit and Sarma, Anish Das and Ganti, Venkatesh and Kaushik, Raghav},
        series = {{SIGMOD} '07},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/1247480.1247530},
        url = {https://dl.acm.org/doi/10.1145/1247480.1247530},
        year = {2007}
}

Incoming Citations (Sorted by Pagerank)

Showing 8 of 8 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 6 of 6 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers