DBScholar

Back to authors

Sebastian Schelter

Author ID
1585
ORCID
0000-0003-4722-5840
Links
(found by gpt-5.6-luna on jul 24 2026)
Most Frequent Institution
University of Amsterdam
Pagerank
0.18777978
Overall Rank
314 | 98.51%
Paper Count
25

Affiliation Timeline

Incoming Non-self Citations Over Time

Total yearly non-self incoming citations across all papers by this author.

Publications by Paper Pagerank

Showing 25 of 25 publications.

Rank Title Year Venue Pagerank
1,350 Automating Large-Scale Data Quality Verification 2018 VLDB 0.00011065626
1,937 Elastic Machine Learning Algorithms in Amazon SageMaker 2020 SIGMOD 9.4524758e-05
2,239 An Intermediate Representation for Optimizing Machine Learning Pipelines 2019 VLDB 8.8875753e-05
3,670 Learning to Validate the Predictions of Black Box Classifiers on Unseen Data 2020 SIGMOD 7.2118928e-05
3,960 HedgeCut: Maintaining Randomised Trees for Low-Latency Machine Unlearning 2021 SIGMOD 6.9878154e-05
4,518 MLINSPECT: A Data Distribution Debugger for Machine Learning Pipelines 2021 SIGMOD 6.6474737e-05
5,182 Probabilistic Demand Forecasting at Scale 2017 VLDB 6.3287692e-05
5,466 SchemaPile: A Large Collection of Relational Database Schemas 2024 SIGMOD 6.2075052e-05
6,188 "Amnesia" - A Selection of Machine Learning Models That Can Forget User Data Very Fast 2020 CIDR 5.9483532e-05
6,250 Lightweight Inspection of Data Preprocessing in Native Machine Learning Pipelines 2021 CIDR 5.9418015e-05
6,801 Unit Testing Data with Deequ 2019 SIGMOD 5.7690726e-05
7,395 Automating and Optimizing Data-Centric What-If Analyses on Native Machine Learning Pipelines 2023 SIGMOD 5.6257796e-05
7,490 mlwhatif: What If You Could Stop Re-Implementing Your Machine Learning Pipeline Analyses Over and Over? 2023 VLDB 5.6061425e-05
8,134 DORIAN in action: Assisted Design of Data Science Pipelines 2022 VLDB 5.4811783e-05
8,868 SemBench: A Benchmark for Semantic Query Processing Engines 2026 VLDB 5.3546848e-05
9,527 Serenade - Low-Latency Session-Based Recommendation in e-Commerce at Scale 2022 SIGMOD 5.2550158e-05
9,670 Optimistic Recovery for Iterative Dataflows in Action 2015 SIGMOD 5.2389261e-05
9,675 BlockJoin: Efficient Matrix Partitioning Through Joins 2017 VLDB 5.2380072e-05
11,043 mlidea: Interactively Improving ML Data Preparation Code via “Shadow Pipelines” 2025 VLDB 5.093636e-05
11,283 A Flexible Forecasting Stack 2024 VLDB 5.093636e-05
11,302 Snapcase – Regain Control over Your Predictions with Low-Latency Machine Unlearning 2024 VLDB 5.093636e-05
11,353 Reconstructing and Querying ML Pipeline Intermediates 2023 CIDR 5.093636e-05
11,509 Screening Native ML Pipelines with “ArgusEyes” 2022 CIDR 5.093636e-05
12,237 Iterative Parallel Data Processing with Stratosphere: An Inside Look 2013 SIGMOD 5.093636e-05
13,501 DEEM 2019: Workshop on Data Management for End-to-End Machine Learning 2019 SIGMOD -
Previous Page 1 / 1 Next

Frequent Co-authors

Co-authored at least 5 papers.

Co-author Shared Papers Rank Pagerank
Stefan Grafberger 8 1,072 0.06798417
Volker Markl 5 26 0.71769703