| 106 |
The MADlib Analytics Library or MAD Skills, the SQL |
2012 |
VLDB |
0.00033539462 |
| 518 |
Towards a Unified Architecture for in-RDBMS Analytics |
2012 |
SIGMOD |
0.00017167492 |
| 640 |
Materialization Optimizations for Feature Selection Workloads |
2014 |
SIGMOD |
0.00015409494 |
| 715 |
Learning Generalized Linear Models Over Normalized Data |
2015 |
SIGMOD |
0.00014655327 |
| 764 |
To Join or Not to Join? Thinking Twice about Joins before Feature Selection |
2016 |
SIGMOD |
0.00014226652 |
| 1,157 |
Cerebro: A Data System for Optimized Deep Learning Model Selection |
2020 |
VLDB |
0.00011924049 |
| 1,235 |
Towards Linear Algebra over Normalized Data |
2017 |
VLDB |
0.00011548457 |
| 1,250 |
Data Management in Machine Learning: Challenges, Techniques, and Systems |
2017 |
SIGMOD |
0.00011485301 |
| 1,488 |
Towards Model-based Pricing for Machine Learning in a Data Marketplace |
2019 |
SIGMOD |
0.00010612416 |
| 2,179 |
Enabling and Optimizing Non-linear Feature Interactions in Factorized Linear Algebra |
2019 |
SIGMOD |
9.0146333e-05 |
| 2,604 |
Brainwash: A Data System for Feature Engineering |
2013 |
CIDR |
8.3524514e-05 |
| 2,702 |
Panorama: A Data System for Unbounded Vocabulary Querying over Video |
2020 |
VLDB |
8.2342712e-05 |
| 2,987 |
Incremental and Approximate Inference for Faster Occlusion-based Deep CNN Explanations |
2019 |
SIGMOD |
7.8907997e-05 |
| 3,459 |
A Comparative Evaluation of Systems for Scalable Linear Algebra-based Analytics |
2018 |
VLDB |
7.3953716e-05 |
| 3,573 |
VISTA: Optimized System for Declarative Feature Transfer from Deep CNNs at Scale |
2020 |
SIGMOD |
7.2977194e-05 |
| 3,681 |
Are Key-Foreign Key Joins Safe to Avoid when Learning High-Capacity Classifiers? |
2018 |
VLDB |
7.2037388e-05 |
| 3,682 |
In-RDBMS Hardware Acceleration of Advanced Analytics |
2018 |
VLDB |
7.2035518e-05 |
| 3,853 |
Understanding and Benchmarking the Impact of GDPR on Database Systems |
2020 |
VLDB |
7.0733274e-05 |
| 4,067 |
Distributed Deep Learning on Data Systems: A Comparative Analysis of Approaches |
2021 |
VLDB |
6.9293511e-05 |
| 4,185 |
Bolt-on Differential Privacy for Scalable Stochastic Gradient Descent-based Analytics |
2017 |
SIGMOD |
6.8462927e-05 |
| 4,593 |
Demonstration of SpeakQL: Speech-driven Multimodal Querying of Structured Data |
2019 |
SIGMOD |
6.6138904e-05 |
| 5,051 |
Towards Benchmarking Feature Type Inference for AutoML Platforms |
2021 |
SIGMOD |
6.385354e-05 |
| 5,323 |
SNAILS: Schema Naming Assessments for Improved LLM-Based SQL Inference |
2025 |
SIGMOD |
6.2655413e-05 |
| 5,557 |
Demonstration of Santoku: Optimizing Machine Learning over Normalized Data |
2015 |
VLDB |
6.17499e-05 |
| 5,738 |
SpeakQL: Towards Speech-driven Multimodal Querying of Structured Data |
2020 |
SIGMOD |
6.1045268e-05 |
| 6,100 |
Demonstration of Nimbus: Model-based Pricing for Machine Learning in a Data Marketplace |
2019 |
SIGMOD |
5.9776732e-05 |
| 6,213 |
The future of data(base) education: Is the "cow book" dead? |
2021 |
VLDB |
5.9425753e-05 |
| 6,485 |
Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent |
2019 |
SIGMOD |
5.8657457e-05 |
| 6,588 |
Lotan: Bridging the Gap between GNNs and Scalable Graph Analytics Engines |
2023 |
VLDB |
5.833338e-05 |
| 6,928 |
How do Categorical Duplicates Affect ML? A New Benchmark and Empirical Analyses |
2024 |
VLDB |
5.7370426e-05 |
| 7,403 |
Feature Selection in Enterprise Analytics: A Demonstration using an R-based Data Analytics System |
2013 |
VLDB |
5.6249895e-05 |
| 7,656 |
Nautilus: An Optimized System for Deep Transfer Learning over Evolving Training Datasets |
2022 |
SIGMOD |
5.5740571e-05 |
| 8,373 |
Automation of Data Prep, ML, and Data Science: New Cure or Snake Oil? |
2021 |
SIGMOD |
5.4399999e-05 |
| 8,585 |
Probabilistic Management of OCR Data using an RDBMS |
2012 |
VLDB |
5.4075213e-05 |
| 8,638 |
Towards A Polyglot Framework for Factorized ML |
2021 |
VLDB |
5.395289e-05 |
| 8,979 |
Cerebro: A Layered Data Platform for Scalable Deep Learning |
2021 |
CIDR |
5.3399615e-05 |
| 9,371 |
Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning |
2021 |
VLDB |
5.275595e-05 |
| 9,372 |
Intermittent Human-in-the-Loop Model Selection using Cerebro: A Demonstration |
2021 |
VLDB |
5.275595e-05 |
| 9,737 |
Saturn: An Optimized Data System for Multi-Large-Model Deep Learning Workloads |
2024 |
VLDB |
5.227679e-05 |
| 13,375 |
Reimagining Deep Learning Systems Through the Lens of Data Systems |
2024 |
VLDB |
- |
| 13,473 |
Errata for "Cerebro: A Data System for Optimized Deep Learning Model Selection" |
2021 |
VLDB |
- |
| 13,514 |
Demonstration of Krypton: Optimized CNN Inference for Occlusion-based Deep CNN Explanations |
2019 |
VLDB |
- |