Efficient and Tunable Similar Set Retrieval
Summary: Formalizes similarity-based indexing for set-valued attributes, reduced to similarity-preserving binary vectors in Hamming space. Proposes two data-structure primitives and a tunable, constraint-driven index; prototype experiments on real datasets show accuracy–efficiency tradeoffs. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 12,229 | TACO: Tunable Approximate Computation of Outliers in Wireless Sensor Networks | 2010 | SIGMOD | 4.1905499e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 6 of 6 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 34 | Similarity Search in High Dimensions via Hashing | 1999 | VLDB | 0.00076824554 |
| 65 | Fast Subsequence Matching in Time-Series Databases | 1994 | SIGMOD | 0.00061977022 |
| 242 | Generalized Search Trees for Database Systems (Extended Abstract) | 1995 | VLDB | 0.00031093647 |
| 362 | Fast Similarity Search in the Presence of Noise, Scaling, and Translation in Time-Series Databases | 1995 | VLDB | 0.00025758421 |
| 2,168 | Selectivity Estimation For Boolean Queries | 2000 | PODS | 9.3877451e-05 |
| 3,017 | Evaluation of Signature Files as Set Access Facilities in OODBs | 1993 | SIGMOD | 7.7012009e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 9,327 | Indexing for Keyword Search with Structured Constraints | 2023 | PODS | 4.351469e-05 |
| 34 | Similarity Search in High Dimensions via Hashing | 1999 | VLDB | 0.00076824554 |
| 8,645 | A Non-Linear Dimensionality-Reduction Technique for Fast Similarity Search in Large Databases | 2006 | SIGMOD | 4.4725847e-05 |
| 248 | Efficient set joins on similarity predicates | 2004 | SIGMOD | 0.00030888982 |
| 3,541 | Similarity search in the blink of an eye with compressed indices | 2023 | VLDB | 6.9910982e-05 |
| 4,273 | Similarity Query Processing for High-Dimensional Data | 2020 | VLDB | 6.2932217e-05 |
| 4,598 | A General and Efficient Querying Method for Learning to Hash | 2018 | SIGMOD | 6.053944e-05 |
| 7,106 | Efficient Similarity Join and Search on Multi-Attribute Data | 2015 | SIGMOD | 4.8250163e-05 |
| 3,461 | Leveraging Set Relations in Exact Set Similarity Join | 2017 | VLDB | 7.0696567e-05 |
| 2,779 | Hashed Samples: Selectivity Estimators For Set Similarity Selection Queries | 2008 | VLDB | 8.1314377e-05 |