Efficient Similarity Search and Classification via Rank Aggregation
Summary: Rank-aggregation with independent voters projecting on random lines; median-rank rule yields a (1+ε)-approximate Euclidean NN. Very efficient: probes ~5% of data, no extra storage, DB-friendly fixed access order; extends to k-NN and classification. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Ronald Fagin (IBM)
- 2. Ravi Kumar (IBM)
- 3. D. Sivakumar (IBM)
BibTeX Citation
@inproceedings{fagin_sigmod03,
title = {{Efficient Similarity Search and Classification via Rank Aggregation}},
author = {Fagin, Ronald and Kumar, Ravi and Sivakumar, D.},
series = {{SIGMOD} '03},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/872757.872795},
url = {https://dl.acm.org/doi/10.1145/872757.872795},
year = {2003}
}
Incoming Citations (Sorted by Pagerank)
Showing 20 of 20 citing papers.
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 7 of 7 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 5 | Optimal Aggregation Algorithms for Middleware [Extended Abstract] | 2001 | PODS | 0.0010679641 |
| 20 | Similarity Search in High Dimensions via Hashing | 1999 | VLDB | 0.00057568153 |
| 90 | The X-tree: An Index Structure for High-Dimensional Data | 1996 | VLDB | 0.00034860244 |
| 474 | Automated Ranking of Database Query Results | 2003 | CIDR | 0.00017672562 |
| 3,049 | Joining Ranked Inputs in Practice | 2002 | VLDB | 7.7090549e-05 |
| 3,320 | Contrast Plots and P-Sphere Trees: Space vs. Time in Nearest Neighbor Searches | 2000 | VLDB | 7.4302233e-05 |
| 6,032 | Hierarchical Subspace Sampling: A Unified Framework for High Dimensional Data Reduction, Selectivity Estimation and Nearest Neighbor Search | 2002 | SIGMOD | 5.9097667e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 4,728 | Finding Near Neighbors Through Cluster Pruning | 2007 | PODS |
| 2 | 3,887 | Fast Parallel Similarity Search in Multimedia Databases | 1997 | SIGMOD |
| 3 | 1,635 | Efficient Search for the Top-k Probable Nearest Neighbors in Uncertain Databases | 2008 | VLDB |
| 4 | 8,818 | A Non-Linear Dimensionality-Reduction Technique for Fast Similarity Search in Large Databases | 2006 | SIGMOD |
| 5 | 13,084 | Efficiency-Quality Tradeoffs for Vector Score Aggregation | 2004 | VLDB |
| 6 | 3,325 | Efficient k-NN Search on Vertically Decomposed Data | 2002 | SIGMOD |
| 7 | 3,774 | Efficient Reverse k-Nearest Neighbor Search in Arbitrary Metric Spaces | 2006 | SIGMOD |
| 8 | 6,868 | Flexible Aggregate Similarity Search | 2011 | SIGMOD |
| 9 | 20 | Similarity Search in High Dimensions via Hashing | 1999 | VLDB |
| 10 | 2,113 | What is the nearest neighbor in high dimensional spaces? | 2000 | VLDB |