Efficient Similarity Search and Classification via Rank Aggregation
Summary: Rank-aggregation with independent voters projecting on random lines; median-rank rule yields a (1+ε)-approximate Euclidean NN. Very efficient: probes ~5% of data, no extra storage, DB-friendly fixed access order; extends to k-NN and classification. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Ronald Fagin (IBM)
- 2. Ravi Kumar (IBM)
- 3. D. Sivakumar (IBM)
BibTeX Citation
@inproceedings{fagin_sigmod03,
title = {{Efficient Similarity Search and Classification via Rank Aggregation}},
author = {Fagin, Ronald and Kumar, Ravi and Sivakumar, D.},
series = {{SIGMOD} '03},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/872757.872795},
url = {https://dl.acm.org/doi/10.1145/872757.872795},
year = {2003}
}
Incoming Citations (Sorted by Pagerank)
Showing 20 of 20 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 7 of 7 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 5 | Optimal Aggregation Algorithms for Middleware [Extended Abstract] | 2001 | PODS | 0.0010828372 |
| 21 | Similarity Search in High Dimensions via Hashing | 1999 | VLDB | 0.00056760516 |
| 85 | The X-tree: An Index Structure for High-Dimensional Data | 1996 | VLDB | 0.00035405879 |
| 466 | Automated Ranking of Database Query Results | 2003 | CIDR | 0.00018014467 |
| 3,017 | Joining Ranked Inputs in Practice | 2002 | VLDB | 7.8483041e-05 |
| 3,282 | Contrast Plots and P-Sphere Trees: Space vs. Time in Nearest Neighbor Searches | 2000 | VLDB | 7.5667715e-05 |
| 5,963 | Hierarchical Subspace Sampling: A Unified Framework for High Dimensional Data Reduction, Selectivity Estimation and Nearest Neighbor Search | 2002 | SIGMOD | 6.0267197e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 4,657 | Finding Near Neighbors Through Cluster Pruning | 2007 | PODS |
| 2 | 3,866 | Fast Parallel Similarity Search in Multimedia Databases | 1997 | SIGMOD |
| 3 | 8,655 | A Non-Linear Dimensionality-Reduction Technique for Fast Similarity Search in Large Databases | 2006 | SIGMOD |
| 4 | 1,605 | Efficient Search for the Top-k Probable Nearest Neighbors in Uncertain Databases | 2008 | VLDB |
| 5 | 12,794 | Efficiency-Quality Tradeoffs for Vector Score Aggregation | 2004 | VLDB |
| 6 | 3,363 | Efficient k-NN Search on Vertically Decomposed Data | 2002 | SIGMOD |
| 7 | 3,710 | Efficient Reverse k-Nearest Neighbor Search in Arbitrary Metric Spaces | 2006 | SIGMOD |
| 8 | 6,746 | Flexible Aggregate Similarity Search | 2011 | SIGMOD |
| 9 | 21 | Similarity Search in High Dimensions via Hashing | 1999 | VLDB |
| 10 | 2,158 | What is the nearest neighbor in high dimensional spaces? | 2000 | VLDB |