| 321 |
Worst-Case Optimal Join Algorithms: Techniques, Results, and Open Problems |
2018 |
PODS |
0.00021283186 |
| 536 |
Learning Linear Regression Models over Factorized Joins |
2016 |
SIGMOD |
0.0001693369 |
| 764 |
To Join or Not to Join? Thinking Twice about Joins before Feature Selection |
2016 |
SIGMOD |
0.00014226652 |
| 802 |
Random Sampling over Joins Revisited |
2018 |
SIGMOD |
0.00013907725 |
| 1,235 |
Towards Linear Algebra over Normalized Data |
2017 |
VLDB |
0.00011548457 |
| 1,250 |
Data Management in Machine Learning: Challenges, Techniques, and Systems |
2017 |
SIGMOD |
0.00011485301 |
| 1,569 |
HELIX: Holistic Optimization for Accelerating Iterative Machine Learning |
2019 |
VLDB |
0.00010335423 |
| 1,756 |
SystemDS: A Declarative Machine Learning System for the End-to-End Data Science Lifecycle |
2020 |
CIDR |
9.8172465e-05 |
| 2,161 |
DIFF: A Relational Interface for Large-Scale Data Explanation |
2019 |
VLDB |
9.0606664e-05 |
| 2,162 |
Heterogeneity-aware Distributed Parameter Servers |
2017 |
SIGMOD |
9.0581831e-05 |
| 2,179 |
Enabling and Optimizing Non-linear Feature Interactions in Factorized Linear Algebra |
2019 |
SIGMOD |
9.0146333e-05 |
| 2,239 |
An Intermediate Representation for Optimizing Machine Learning Pipelines |
2019 |
VLDB |
8.8875753e-05 |
| 2,266 |
On Functional Aggregate Queries with Additive Inequalities |
2019 |
PODS |
8.8391372e-05 |
| 2,769 |
A Layered Aggregate Engine for Analytics Workloads |
2019 |
SIGMOD |
8.1465406e-05 |
| 2,865 |
End-to-end Optimization of Machine Learning Prediction Queries |
2022 |
SIGMOD |
8.0180243e-05 |
| 2,927 |
In-Database Learning with Sparse Tensors |
2018 |
PODS |
7.9531195e-05 |
| 3,206 |
Incremental View Maintenance with Triple Lock Factorization Benefits |
2018 |
SIGMOD |
7.6367549e-05 |
| 3,284 |
SPOOF: Sum-Product Optimization and Operator Fusion for Large-Scale Machine Learning |
2017 |
CIDR |
7.5663058e-05 |
| 3,459 |
A Comparative Evaluation of Systems for Scalable Linear Algebra-based Analytics |
2018 |
VLDB |
7.3953716e-05 |
| 3,490 |
MLog: Towards Declarative In-Database Machine Learning |
2017 |
VLDB |
7.3667971e-05 |
| 3,573 |
VISTA: Optimized System for Declarative Feature Transfer from Deep CNNs at Scale |
2020 |
SIGMOD |
7.2977194e-05 |
| 3,681 |
Are Key-Foreign Key Joins Safe to Avoid when Learning High-Capacity Classifiers? |
2018 |
VLDB |
7.2037388e-05 |
| 3,682 |
In-RDBMS Hardware Acceleration of Advanced Analytics |
2018 |
VLDB |
7.2035518e-05 |
| 4,023 |
Smurf: Self-Service String Matching Using Random Forests |
2019 |
VLDB |
6.949387e-05 |
| 4,128 |
The Relational Data Borg is Learning |
2020 |
VLDB |
6.8850804e-05 |
| 4,571 |
GeCo: Quality Counterfactual Explanations in Real Time |
2021 |
VLDB |
6.625099e-05 |
| 4,799 |
Scalable Asynchronous Gradient Descent Optimization for Out-of-Core Models |
2017 |
VLDB |
6.5024714e-05 |
| 5,414 |
Optimizing Tensor Programs on Flexible Storage |
2023 |
SIGMOD |
6.2258658e-05 |
| 5,551 |
PGMJoins: Random Join Sampling with Graphical Models |
2021 |
SIGMOD |
6.1782856e-05 |
| 5,557 |
Demonstration of Santoku: Optimizing Machine Learning over Normalized Data |
2015 |
VLDB |
6.17499e-05 |
| 5,593 |
Beyond Equi-joins: Ranking, Enumeration and Factorization |
2021 |
VLDB |
6.1552328e-05 |
| 5,601 |
Optimal Join Algorithms Meet Top-k |
2020 |
SIGMOD |
6.1540123e-05 |
| 5,785 |
BlinkML: Efficient Maximum Likelihood Estimation with Probabilistic Guarantees |
2019 |
SIGMOD |
6.0892672e-05 |
| 6,038 |
Efficient Construction of Approximate Ad-Hoc ML models Through Materialization and Reuse |
2018 |
VLDB |
5.9990929e-05 |
| 6,143 |
ColumnML: Column-Store Machine Learning with On-The-Fly Data Transformation |
2019 |
VLDB |
5.9622615e-05 |
| 6,339 |
In-Database Machine Learning with CorgiPile: Stochastic Gradient Descent without Full Data Shuffle |
2022 |
SIGMOD |
5.907165e-05 |
| 6,485 |
Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent |
2019 |
SIGMOD |
5.8657457e-05 |
| 7,000 |
A Cost-based Optimizer for Gradient Descent Optimization |
2017 |
SIGMOD |
5.7287645e-05 |
| 7,112 |
Coresets over Multiple Tables for Feature-rich and Data-efficient Machine Learning |
2023 |
VLDB |
5.6990782e-05 |
| 7,985 |
Schema Independent Relational Learning |
2017 |
SIGMOD |
5.5125969e-05 |
| 8,513 |
Galley: Modern Query Optimization for Sparse Tensor Programs |
2025 |
SIGMOD |
5.4119882e-05 |
| 8,638 |
Towards A Polyglot Framework for Factorized ML |
2021 |
VLDB |
5.395289e-05 |
| 8,794 |
AWARE: Workload-aware, Redundancy-exploiting Linear Algebra |
2023 |
SIGMOD |
5.370464e-05 |
| 8,908 |
Privacy and Accuracy-Aware AI/ML Model Deduplication |
2025 |
SIGMOD |
5.3483178e-05 |
| 8,979 |
Cerebro: A Layered Data Platform for Scalable Deep Learning |
2021 |
CIDR |
5.3399615e-05 |
| 9,371 |
Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning |
2021 |
VLDB |
5.275595e-05 |
| 9,675 |
BlockJoin: Efficient Matrix Partitioning Through Joins |
2017 |
VLDB |
5.2380072e-05 |
| 9,719 |
Subset Sampling over Joins |
2026 |
PODS |
5.2319816e-05 |
| 9,993 |
In-Database Data Imputation |
2024 |
SIGMOD |
5.1815618e-05 |
| 10,153 |
Faster Relational Algorithms Using Geometric Data Structures |
2026 |
PODS |
5.093636e-05 |