DBScholar

Back to authors

Arun Kumar

Author ID
871
ORCID
-
Links
(found by gpt-5.6-luna on jul 24 2026)
Most Frequent Institution
University of California San Diego
Pagerank
0.35039113
Overall Rank
125 | 99.41%
Paper Count
42

Affiliation Timeline

Incoming Non-self Citations Over Time

Total yearly non-self incoming citations across all papers by this author.

Publications by Paper Pagerank

Showing 42 of 42 publications.

Rank Title Year Venue Pagerank
106 The MADlib Analytics Library or MAD Skills, the SQL 2012 VLDB 0.00033539462
518 Towards a Unified Architecture for in-RDBMS Analytics 2012 SIGMOD 0.00017167492
640 Materialization Optimizations for Feature Selection Workloads 2014 SIGMOD 0.00015409494
715 Learning Generalized Linear Models Over Normalized Data 2015 SIGMOD 0.00014655327
764 To Join or Not to Join? Thinking Twice about Joins before Feature Selection 2016 SIGMOD 0.00014226652
1,157 Cerebro: A Data System for Optimized Deep Learning Model Selection 2020 VLDB 0.00011924049
1,235 Towards Linear Algebra over Normalized Data 2017 VLDB 0.00011548457
1,250 Data Management in Machine Learning: Challenges, Techniques, and Systems 2017 SIGMOD 0.00011485301
1,488 Towards Model-based Pricing for Machine Learning in a Data Marketplace 2019 SIGMOD 0.00010612416
2,179 Enabling and Optimizing Non-linear Feature Interactions in Factorized Linear Algebra 2019 SIGMOD 9.0146333e-05
2,604 Brainwash: A Data System for Feature Engineering 2013 CIDR 8.3524514e-05
2,702 Panorama: A Data System for Unbounded Vocabulary Querying over Video 2020 VLDB 8.2342712e-05
2,987 Incremental and Approximate Inference for Faster Occlusion-based Deep CNN Explanations 2019 SIGMOD 7.8907997e-05
3,459 A Comparative Evaluation of Systems for Scalable Linear Algebra-based Analytics 2018 VLDB 7.3953716e-05
3,573 VISTA: Optimized System for Declarative Feature Transfer from Deep CNNs at Scale 2020 SIGMOD 7.2977194e-05
3,681 Are Key-Foreign Key Joins Safe to Avoid when Learning High-Capacity Classifiers? 2018 VLDB 7.2037388e-05
3,682 In-RDBMS Hardware Acceleration of Advanced Analytics 2018 VLDB 7.2035518e-05
3,853 Understanding and Benchmarking the Impact of GDPR on Database Systems 2020 VLDB 7.0733274e-05
4,067 Distributed Deep Learning on Data Systems: A Comparative Analysis of Approaches 2021 VLDB 6.9293511e-05
4,185 Bolt-on Differential Privacy for Scalable Stochastic Gradient Descent-based Analytics 2017 SIGMOD 6.8462927e-05
4,593 Demonstration of SpeakQL: Speech-driven Multimodal Querying of Structured Data 2019 SIGMOD 6.6138904e-05
5,051 Towards Benchmarking Feature Type Inference for AutoML Platforms 2021 SIGMOD 6.385354e-05
5,323 SNAILS: Schema Naming Assessments for Improved LLM-Based SQL Inference 2025 SIGMOD 6.2655413e-05
5,557 Demonstration of Santoku: Optimizing Machine Learning over Normalized Data 2015 VLDB 6.17499e-05
5,738 SpeakQL: Towards Speech-driven Multimodal Querying of Structured Data 2020 SIGMOD 6.1045268e-05
6,100 Demonstration of Nimbus: Model-based Pricing for Machine Learning in a Data Marketplace 2019 SIGMOD 5.9776732e-05
6,213 The future of data(base) education: Is the "cow book" dead? 2021 VLDB 5.9425753e-05
6,485 Tuple-oriented Compression for Large-scale Mini-batch Stochastic Gradient Descent 2019 SIGMOD 5.8657457e-05
6,588 Lotan: Bridging the Gap between GNNs and Scalable Graph Analytics Engines 2023 VLDB 5.833338e-05
6,928 How do Categorical Duplicates Affect ML? A New Benchmark and Empirical Analyses 2024 VLDB 5.7370426e-05
7,403 Feature Selection in Enterprise Analytics: A Demonstration using an R-based Data Analytics System 2013 VLDB 5.6249895e-05
7,656 Nautilus: An Optimized System for Deep Transfer Learning over Evolving Training Datasets 2022 SIGMOD 5.5740571e-05
8,373 Automation of Data Prep, ML, and Data Science: New Cure or Snake Oil? 2021 SIGMOD 5.4399999e-05
8,585 Probabilistic Management of OCR Data using an RDBMS 2012 VLDB 5.4075213e-05
8,638 Towards A Polyglot Framework for Factorized ML 2021 VLDB 5.395289e-05
8,979 Cerebro: A Layered Data Platform for Scalable Deep Learning 2021 CIDR 5.3399615e-05
9,371 Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning 2021 VLDB 5.275595e-05
9,372 Intermittent Human-in-the-Loop Model Selection using Cerebro: A Demonstration 2021 VLDB 5.275595e-05
9,737 Saturn: An Optimized Data System for Multi-Large-Model Deep Learning Workloads 2024 VLDB 5.227679e-05
13,375 Reimagining Deep Learning Systems Through the Lens of Data Systems 2024 VLDB -
13,473 Errata for "Cerebro: A Data System for Optimized Deep Learning Model Selection" 2021 VLDB -
13,514 Demonstration of Krypton: Optimized CNN Inference for Occlusion-based Deep CNN Explanations 2019 VLDB -
Previous Page 1 / 1 Next

Frequent Co-authors

Co-authored at least 5 papers.

Co-author Shared Papers Rank Pagerank
Supun Nakandala 8 953 0.07506981
Jeffrey Naughton 6 8 1.0247426
Christopher RĂ© 6 63 0.50506981
Yuhao Zhang 6 1,848 0.043263423
Jignesh Patel 5 30 0.69867277
Vraj Shah 5 1,607 0.047884138
Side Li 5 1,628 0.047148077
Lingjiao Chen 5 1,682 0.045932932