Lowering the Latency of Data Processing Pipelines Through FPGA based Hardware Acceleration
Summary: FPGA-based acceleration lowers pipeline latency by reducing data movement and speeding scoring via a decision-tree ensemble. The compact FPGA engine boosts throughput, integrates with earlier stages, and delivers two orders of magnitude speedup over CPU on a real Amazon F1 baseline. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
Incoming Citations (Sorted by Pagerank)
Showing 5 of 5 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 3,328 | Pump Up the Volume: Processing Large Data on GPUs with Fast Interconnects | 2020 | SIGMOD | 7.2136181e-05 |
| 5,089 | TCUDB: Accelerating Database with Tensor Processors | 2022 | SIGMOD | 5.7017353e-05 |
| 6,280 | Cheetah: Accelerating Database Queries with Switch Pruning | 2020 | SIGMOD | 5.1239052e-05 |
| 10,516 | SwiftSpatial: Spatial Joins on Modern Hardware | 2025 | SIGMOD | 4.1905499e-05 |
| 11,593 | Making Search Engines Faster by Lowering the Cost of Querying Business Rules Through FPGAs | 2020 | SIGMOD | 4.1905499e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 10 of 10 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 1,458 | RainForest - A Framework for Fast Decision Tree Construction of Large Datasets | 1998 | VLDB | 0.00011888676 |
| 2,372 | Predictable Performance for Unpredictable Workloads | 2009 | VLDB | 8.940791e-05 |
| 2,636 | PLANET: Massively Parallel Learning of Tree Ensembles with MapReduce | 2009 | VLDB | 8.401513e-05 |
| 2,689 | BOAT—Optimistic Decision Tree Construction | 1999 | SIGMOD | 8.2965064e-05 |
| 4,040 | In-RDBMS Hardware Acceleration of Advanced Analytics | 2018 | VLDB | 6.5052227e-05 |
| 5,124 | Accelerating Generalized Linear Models with MLWeaving: A One-Size-Fits-All System for Any-Precision Learning | 2019 | VLDB | 5.6747536e-05 |
| 5,179 | FPGA-based Data Partitioning | 2017 | SIGMOD | 5.6384436e-05 |
| 6,400 | ColumnML: Column-Store Machine Learning with On-The-Fly Data Transformation | 2019 | VLDB | 5.0739311e-05 |
| 8,201 | Accelerating Pattern Matching Queries in Hybrid CPU-FPGA Architectures | 2017 | SIGMOD | 4.5555217e-05 |
| 8,417 | doppioDB: A Hardware Accelerated Database | 2017 | SIGMOD | 4.5120193e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| Overall Rank | Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,279 | Streams on Wires — A Query Compiler for FPGAs | 2009 | VLDB | 9.1249844e-05 |
| 382 | FAST: Fast Architecture Sensitive Tree Search on Modern CPUs and GPUs | 2010 | SIGMOD | 0.00024888997 |
| 9,784 | Is FPGA Useful for Hash Joins? Exploring Hash Joins on Coupled CPU-FPGA Architecture | 2020 | CIDR | 4.2806921e-05 |
| 5,732 | FPGA-based Multithreading for In-Memory Hash Joins | 2015 | CIDR | 5.3473621e-05 |
| 5,179 | FPGA-based Data Partitioning | 2017 | SIGMOD | 5.6384436e-05 |
| 4,383 | Flexible Query Processor on FPGAs | 2013 | VLDB | 6.2274297e-05 |
| 9,881 | DASH: Asynchronous Hardware Data Processing Services | 2023 | CIDR | 4.2602815e-05 |
| 948 | Data Processing on FPGAs | 2009 | VLDB | 0.0001510759 |
| 10,711 | Fast Graph Vector Search via Hardware Acceleration and Delayed-Synchronization Traversal | 2025 | VLDB | 4.1905499e-05 |
| 11,593 | Making Search Engines Faster by Lowering the Cost of Querying Business Rules Through FPGAs | 2020 | SIGMOD | 4.1905499e-05 |