| 315 |
Worst-Case Optimal Join Algorithms: Techniques, Results, and Open Problems |
2018 |
PODS |
0.00021246 |
| 521 |
Learning Linear Regression Models over Factorized Joins |
2016 |
SIGMOD |
0.00016929744 |
| 777 |
To Join or Not to Join? Thinking Twice about Joins before Feature Selection |
2016 |
SIGMOD |
0.00014054709 |
| 795 |
Random Sampling over Joins Revisited |
2018 |
SIGMOD |
0.00013938779 |
| 1,255 |
Data Management in Machine Learning: Challenges, Techniques, and Systems |
2017 |
SIGMOD |
0.00011325762 |
| 1,256 |
Towards Linear Algebra over Normalized Data |
2017 |
VLDB |
0.00011314687 |
| 1,568 |
HELIX: Holistic Optimization for Accelerating Iterative Machine Learning |
2019 |
VLDB |
0.0001021302 |
| 1,668 |
SystemDS: A Declarative Machine Learning System for the End-to-End Data Science Lifecycle |
2020 |
CIDR |
9.9371612e-05 |
| 2,160 |
DIFF: A Relational Interface for Large-Scale Data Explanation |
2019 |
VLDB |
8.9364035e-05 |
| 2,184 |
Heterogeneity-aware Distributed Parameter Servers |
2017 |
SIGMOD |
8.8958335e-05 |
| 2,199 |
Enabling and Optimizing Non-linear Feature Interactions in Factorized Linear Algebra |
2019 |
SIGMOD |
8.8750296e-05 |
| 2,264 |
An Intermediate Representation for Optimizing Machine Learning Pipelines |
2019 |
VLDB |
8.7289107e-05 |
| 2,299 |
On Functional Aggregate Queries with Additive Inequalities |
2019 |
PODS |
8.6727773e-05 |
| 2,662 |
End-to-end Optimization of Machine Learning Prediction Queries |
2022 |
SIGMOD |
8.1596229e-05 |
| 2,663 |
A Layered Aggregate Engine for Analytics Workloads |
2019 |
SIGMOD |
8.1581558e-05 |
| 2,975 |
In-Database Learning with Sparse Tensors |
2018 |
PODS |
7.7907759e-05 |
| 3,112 |
Incremental View Maintenance with Triple Lock Factorization Benefits |
2018 |
SIGMOD |
7.6357579e-05 |
| 3,330 |
SPOOF: Sum-Product Optimization and Operator Fusion for Large-Scale Machine Learning |
2017 |
CIDR |
7.4173693e-05 |
| 3,518 |
A Comparative Evaluation of Systems for Scalable Linear Algebra-based Analytics |
2018 |
VLDB |
7.2400627e-05 |
| 3,524 |
MLog: Towards Declarative In-Database Machine Learning |
2017 |
VLDB |
7.2337006e-05 |
| 3,630 |
VISTA: Optimized System for Declarative Feature Transfer from Deep CNNs at Scale |
2020 |
SIGMOD |
7.1496182e-05 |
| 3,670 |
Are Key-Foreign Key Joins Safe to Avoid when Learning High-Capacity Classifiers? |
2018 |
VLDB |
7.1108704e-05 |
| 3,737 |
In-RDBMS Hardware Acceleration of Advanced Analytics |
2018 |
VLDB |
7.0628666e-05 |
| 4,107 |
Smurf: Self-Service String Matching Using Random Forests |
2019 |
VLDB |
6.8026037e-05 |
| 4,161 |
The Relational Data Borg is Learning |
2020 |
VLDB |
6.7700593e-05 |
| 4,671 |
GeCo: Quality Counterfactual Explanations in Real Time |
2021 |
VLDB |
6.4764543e-05 |
| 4,752 |
Optimal Join Algorithms Meet Top-k |
2020 |
SIGMOD |
6.434561e-05 |
| 4,901 |
Scalable Asynchronous Gradient Descent Optimization for Out-of-Core Models |
2017 |
VLDB |
6.3647386e-05 |
| 5,439 |
Optimizing Tensor Programs on Flexible Storage |
2023 |
SIGMOD |
6.1279464e-05 |
| 5,482 |
Beyond Equi-joins: Ranking, Enumeration and Factorization |
2021 |
VLDB |
6.111411e-05 |
| 5,657 |
PGMJoins: Random Join Sampling with Graphical Models |
2021 |
SIGMOD |
6.0488437e-05 |
| 5,682 |
Demonstration of Santoku: Optimizing Machine Learning over Normalized Data |
2015 |
VLDB |
6.0397904e-05 |
| 5,877 |
BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees |
2019 |
SIGMOD |
5.9627218e-05 |
| 6,143 |
Efficient Construction of Approximate Ad-Hoc ML models Through Materialization and Reuse |
2018 |
VLDB |
5.8721471e-05 |
| 6,246 |
ColumnML: Column-Store Machine Learning with On-The-Fly Data Transformation |
2019 |
VLDB |
5.8370793e-05 |
| 6,458 |
In-Database Machine Learning with CorgiPile: Stochastic Gradient Descent without Full Data Shuffle |
2022 |
SIGMOD |
5.7773842e-05 |
| 6,592 |
Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent |
2019 |
SIGMOD |
5.7421684e-05 |
| 7,142 |
A Cost-based Optimizer for Gradient Descent Optimization |
2017 |
SIGMOD |
5.600563e-05 |
| 7,201 |
Coresets over Multiple Tables for Feature-rich and Data-efficient Machine Learning |
2023 |
VLDB |
5.5871656e-05 |
| 8,151 |
Schema Independent Relational Learning |
2017 |
SIGMOD |
5.3899758e-05 |
| 8,279 |
Subset Sampling over Joins |
2026 |
PODS |
5.3637624e-05 |
| 8,384 |
AWARE: Workload-aware, Redundancy-exploiting Linear Algebra |
2023 |
SIGMOD |
5.3421754e-05 |
| 8,682 |
Galley: Modern Query Optimization for Sparse Tensor Programs |
2025 |
SIGMOD |
5.2905577e-05 |
| 8,798 |
Towards A Polyglot Framework for Factorized ML |
2021 |
VLDB |
5.2746322e-05 |
| 9,071 |
Privacy and Accuracy-Aware AI/ML Model Deduplication |
2025 |
SIGMOD |
5.2283159e-05 |
| 9,144 |
Cerebro: A Layered Data Platform for Scalable Deep Learning |
2021 |
CIDR |
5.2201471e-05 |
| 9,556 |
Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning |
2021 |
VLDB |
5.1572248e-05 |
| 9,822 |
BlockJoin: Efficient Matrix Partitioning Through Joins |
2017 |
VLDB |
5.1254832e-05 |
| 10,181 |
In-Database Data Imputation |
2024 |
SIGMOD |
5.0653015e-05 |
| 10,370 |
Faster Relational Algorithms Using Geometric Data Structures |
2026 |
PODS |
4.9793485e-05 |