| 315 |
Worst-Case Optimal Join Algorithms: Techniques, Results, and Open Problems |
2018 |
PODS |
0.00021236408 |
| 521 |
Learning Linear Regression Models over Factorized Joins |
2016 |
SIGMOD |
0.00016923519 |
| 779 |
To Join or Not to Join? Thinking Twice about Joins before Feature Selection |
2016 |
SIGMOD |
0.00014048128 |
| 795 |
Random Sampling over Joins Revisited |
2018 |
SIGMOD |
0.00013934719 |
| 1,223 |
Data Management in Machine Learning: Challenges, Techniques, and Systems |
2017 |
SIGMOD |
0.00011468426 |
| 1,257 |
Towards Linear Algebra over Normalized Data |
2017 |
VLDB |
0.0001130959 |
| 1,568 |
HELIX: Holistic Optimization for Accelerating Iterative Machine Learning |
2019 |
VLDB |
0.00010208225 |
| 1,669 |
SystemDS: A Declarative Machine Learning System for the End-to-End Data Science Lifecycle |
2020 |
CIDR |
9.9324573e-05 |
| 2,162 |
DIFF: A Relational Interface for Large-Scale Data Explanation |
2019 |
VLDB |
8.9344773e-05 |
| 2,186 |
Heterogeneity-aware Distributed Parameter Servers |
2017 |
SIGMOD |
8.8916253e-05 |
| 2,201 |
Enabling and Optimizing Non-linear Feature Interactions in Factorized Linear Algebra |
2019 |
SIGMOD |
8.8708356e-05 |
| 2,266 |
An Intermediate Representation for Optimizing Machine Learning Pipelines |
2019 |
VLDB |
8.7248802e-05 |
| 2,302 |
On Functional Aggregate Queries with Additive Inequalities |
2019 |
PODS |
8.6687541e-05 |
| 2,661 |
End-to-end Optimization of Machine Learning Prediction Queries |
2022 |
SIGMOD |
8.1568473e-05 |
| 2,663 |
A Layered Aggregate Engine for Analytics Workloads |
2019 |
SIGMOD |
8.1542952e-05 |
| 2,978 |
In-Database Learning with Sparse Tensors |
2018 |
PODS |
7.7872011e-05 |
| 3,114 |
Incremental View Maintenance with Triple Lock Factorization Benefits |
2018 |
SIGMOD |
7.6321464e-05 |
| 3,331 |
SPOOF: Sum-Product Optimization and Operator Fusion for Large-Scale Machine Learning |
2017 |
CIDR |
7.4138851e-05 |
| 3,518 |
A Comparative Evaluation of Systems for Scalable Linear Algebra-based Analytics |
2018 |
VLDB |
7.236638e-05 |
| 3,524 |
MLog: Towards Declarative In-Database Machine Learning |
2017 |
VLDB |
7.2303094e-05 |
| 3,632 |
VISTA: Optimized System for Declarative Feature Transfer from Deep CNNs at Scale |
2020 |
SIGMOD |
7.146238e-05 |
| 3,673 |
Are Key-Foreign Key Joins Safe to Avoid when Learning High-Capacity Classifiers? |
2018 |
VLDB |
7.1075403e-05 |
| 3,739 |
In-RDBMS Hardware Acceleration of Advanced Analytics |
2018 |
VLDB |
7.0595382e-05 |
| 4,109 |
Smurf: Self-Service String Matching Using Random Forests |
2019 |
VLDB |
6.7994518e-05 |
| 4,161 |
The Relational Data Borg is Learning |
2020 |
VLDB |
6.7669004e-05 |
| 4,673 |
GeCo: Quality Counterfactual Explanations in Real Time |
2021 |
VLDB |
6.4733885e-05 |
| 4,745 |
Optimal Join Algorithms Meet Top-k |
2020 |
SIGMOD |
6.4369578e-05 |
| 4,902 |
Scalable Asynchronous Gradient Descent Optimization for Out-of-Core Models |
2017 |
VLDB |
6.3617262e-05 |
| 5,444 |
Optimizing Tensor Programs on Flexible Storage |
2023 |
SIGMOD |
6.1250455e-05 |
| 5,487 |
Beyond Equi-joins: Ranking, Enumeration and Factorization |
2021 |
VLDB |
6.1085184e-05 |
| 5,657 |
PGMJoins: Random Join Sampling with Graphical Models |
2021 |
SIGMOD |
6.0461e-05 |
| 5,683 |
Demonstration of Santoku: Optimizing Machine Learning over Normalized Data |
2015 |
VLDB |
6.0369434e-05 |
| 5,877 |
BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees |
2019 |
SIGMOD |
5.9599042e-05 |
| 6,146 |
Efficient Construction of Approximate Ad-Hoc ML models Through Materialization and Reuse |
2018 |
VLDB |
5.869373e-05 |
| 6,249 |
ColumnML: Column-Store Machine Learning with On-The-Fly Data Transformation |
2019 |
VLDB |
5.8343283e-05 |
| 6,460 |
In-Database Machine Learning with CorgiPile: Stochastic Gradient Descent without Full Data Shuffle |
2022 |
SIGMOD |
5.7746493e-05 |
| 6,594 |
Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent |
2019 |
SIGMOD |
5.7394502e-05 |
| 7,144 |
A Cost-based Optimizer for Gradient Descent Optimization |
2017 |
SIGMOD |
5.5979118e-05 |
| 7,203 |
Coresets over Multiple Tables for Feature-rich and Data-efficient Machine Learning |
2023 |
VLDB |
5.5845207e-05 |
| 8,157 |
Schema Independent Relational Learning |
2017 |
SIGMOD |
5.387426e-05 |
| 8,285 |
Subset Sampling over Joins |
2026 |
PODS |
5.3612232e-05 |
| 8,389 |
AWARE: Workload-aware, Redundancy-exploiting Linear Algebra |
2023 |
SIGMOD |
5.3396465e-05 |
| 8,690 |
Galley: Modern Query Optimization for Sparse Tensor Programs |
2025 |
SIGMOD |
5.2880532e-05 |
| 8,806 |
Towards A Polyglot Framework for Factorized ML |
2021 |
VLDB |
5.2721353e-05 |
| 9,080 |
Privacy and Accuracy-Aware AI/ML Model Deduplication |
2025 |
SIGMOD |
5.2258409e-05 |
| 9,153 |
Cerebro: A Layered Data Platform for Scalable Deep Learning |
2021 |
CIDR |
5.217676e-05 |
| 9,564 |
Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning |
2021 |
VLDB |
5.1547835e-05 |
| 9,829 |
BlockJoin: Efficient Matrix Partitioning Through Joins |
2017 |
VLDB |
5.1230568e-05 |
| 10,185 |
In-Database Data Imputation |
2024 |
SIGMOD |
5.0629036e-05 |
| 10,382 |
Faster Relational Algorithms Using Geometric Data Structures |
2026 |
PODS |
4.9769913e-05 |