Streaming Similarity Self-Join
Summary: Formulates similarity self-join over unbounded streams via time-dependent similarity, enabling bounded-memory processing. Proposes MiniBatch and tightly integrated time-filtering Streaming frameworks, including the streaming-optimized L2 index; STR+L2 scales best empirically. (summarized by gpt-5.6-luna on Jul 24 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Gianmarco De Francisci Morales (Qatar Computing Research Institute)
- 2. Aristides Gionis (Aalto University)
BibTeX Citation
@article{morales_vldb16,
title = {{Streaming Similarity Self-Join}},
author = {Morales, Gianmarco De Francisci and Gionis, Aristides},
journal = {PVLDB},
series = {{VLDB} '16},
volume = {9},
number = {10},
pages = {792--803},
doi = {10.14778/2977797.2977800},
url = {https://doi.org/10.14778/2977797.2977800},
year = {2016}
}
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 11,181 | Low-Latency Adaptive Distributed Stream Join System Based on a Flexible Join Model | 2024 | SIGMOD | 5.093636e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 4 of 4 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 169 | Efficient Exact Set-Similarity Joins | 2006 | VLDB | 0.0002743469 |
| 356 | Efficient Parallel Set-Similarity Joins Using MapReduce | 2010 | SIGMOD | 0.00020303289 |
| 3,725 | Finding replicated web collections | 2000 | SIGMOD | 7.171112e-05 |
| 3,920 | Continually Evaluating Similarity-Based Pattern Queries on a Streaming Time Series | 2002 | SIGMOD | 7.0171134e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 6,014 | Scaling Similarity Joins over Tree-Structured Data | 2015 | VLDB |
| 2 | 9,471 | Efficient and Accurate SimRank-based Similarity Joins: Experiments, Analysis, and Improvement | 2024 | VLDB |
| 3 | 10,497 | Scalable Clustering Over High Dimensional Vector Streams | 2026 | SIGMOD |
| 4 | 2,390 | Streaming Similarity Search over one Billion Tweets using Parallel Locality-Sensitive Hashing | 2013 | VLDB |
| 5 | 4,857 | Similarity Join Size Estimation using Locality Sensitive Hashing | 2011 | VLDB |
| 6 | 4,061 | Efficient Top-K SimRank-based Similarity Join | 2015 | VLDB |
| 7 | 3,040 | Leveraging Set Relations in Exact Set Similarity Join | 2017 | VLDB |
| 8 | 6,012 | Parallel Index-based Stream Join on a Multicore CPU | 2020 | SIGMOD |
| 9 | 11,149 | Similarity Joins of Sparse Features | 2024 | SIGMOD |
| 10 | 9,052 | Fast Approximate Similarity Join in Vector Databases | 2025 | SIGMOD |