Executing Stream Joins on the Cell Processor
Summary: Investigates scalable windowed stream joins on the Cell processor, exploiting heterogeneous cores and SIMD for high parallelism. Key techniques: lightweight window partitioning, column-oriented windows, delay-tuned buffering, rate-aware batching; SIMD kernels boost throughput, achieving near-linear scaling (~13 GB/s) and ~8.3× faster than SSE on dual Xeon. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Bugra Gedik
- 2. Philip S. Yu
- 3. Rajesh R. Bordawekar
Incoming Citations (Sorted by Pagerank)
Showing 10 of 10 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 350 | Sort vs. Hash Revisited: Fast Join Implementation on Modern Multi-Core CPUs | 2009 | VLDB | 0.00026368305 |
| 1,121 | Improving the Performance of List Intersection | 2009 | VLDB | 0.00013838956 |
| 1,284 | Photon: Fault-tolerant and Scalable Joining of Continuous Data Streams | 2013 | SIGMOD | 0.00012820565 |
| 1,467 | SPADE: The System S Declarative Stream Processing Engine | 2008 | SIGMOD | 0.00011838188 |
| 2,279 | Streams on Wires — A Query Compiler for FPGAs | 2009 | VLDB | 9.1249844e-05 |
| 4,161 | Scalable Distributed Stream Join Processing | 2015 | SIGMOD | 6.3883784e-05 |
| 6,046 | FPGA: What's in it for a Database? | 2009 | SIGMOD | 5.2357548e-05 |
| 6,430 | Providing Streaming Joins as a Service at Facebook | 2018 | VLDB | 5.0587634e-05 |
| 7,878 | Thread Cooperation in Multicore Architectures for Frequency Counting over Multiple Data Streams | 2009 | VLDB | 4.6257581e-05 |
| 11,813 | CarStream: An Industrial System of Big Data Processing for Internet-of-Vehicles | 2017 | VLDB | 4.1905499e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 15 of 15 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 350 | Sort vs. Hash Revisited: Fast Join Implementation on Modern Multi-Core CPUs | 2009 | VLDB | 0.00026368305 |
| 1,430 | A Scalable, Predictable Join Operator for Highly Concurrent Data Warehouses | 2009 | VLDB | 0.0001202506 |
| 4,161 | Scalable Distributed Stream Join Processing | 2015 | SIGMOD | 6.3883784e-05 |
| 4,139 | Memory-Limited Execution of Windowed Stream Joins | 2004 | VLDB | 6.4126764e-05 |
| 6,322 | Revisiting Pipelined Parallelism in Multi-Join Query Processing | 2005 | VLDB | 5.1074949e-05 |
| 8,021 | Parallelizing Intra-Window Join on Multicores: An Experimental Study | 2021 | SIGMOD | 4.600223e-05 |
| 7,832 | Effective Resource Utilization for Multiprocessor Join Execution | 1989 | VLDB | 4.636825e-05 |
| 10,970 | Low-Latency Adaptive Distributed Stream Join System Based on a Flexible Join Model | 2024 | SIGMOD | 4.1905499e-05 |
| 1,693 | How Soccer Players Would do Stream Joins | 2011 | SIGMOD | 0.00010886797 |
| 6,469 | Parallel Index-based Stream Join on a Multicore CPU | 2020 | SIGMOD | 5.0448159e-05 |