DBScholar

Back to authors

Sebastian Schelter

Author ID
o0000-0003-4722-5840
ORCID
0000-0003-4722-5840
Links
(found by gpt-5.6-luna on jul 24 2026)
Most Frequent Institution
University of Amsterdam
Pagerank
0.20065019
Overall Rank
288 | 98.67%
Paper Count
27

Affiliation Timeline

Incoming Non-self Citations Over Time

Total yearly non-self incoming citations across all papers by this author.

Publications by Paper Pagerank

Showing 27 of 27 publications.

Rank Title Year Venue Pagerank
1,308 Automating Large-Scale Data Quality Verification 2018 VLDB 0.0001107886
1,927 Elastic Machine Learning Algorithms in Amazon SageMaker 2020 SIGMOD 9.3607648e-05
2,264 An Intermediate Representation for Optimizing Machine Learning Pipelines 2019 VLDB 8.7289107e-05
3,735 Learning to Validate the Predictions of Black Box Classifiers on Unseen Data 2020 SIGMOD 7.0642839e-05
4,037 HedgeCut: Maintaining Randomised Trees for Low-Latency Machine Unlearning 2021 SIGMOD 6.8390354e-05
4,606 MLINSPECT: A Data Distribution Debugger for Machine Learning Pipelines 2021 SIGMOD 6.5041225e-05
5,304 Probabilistic Demand Forecasting at Scale 2017 VLDB 6.1873575e-05
5,385 SchemaPile: A Large Collection of Relational Database Schemas 2024 SIGMOD 6.1523255e-05
5,501 SemBench: A Benchmark for Semantic Query Processing Engines 2026 VLDB 6.1027188e-05
6,321 "Amnesia" - A Selection of Machine Learning Models That Can Forget User Data Very Fast 2020 CIDR 5.814903e-05
6,323 Lightweight Inspection of Data Preprocessing in Native Machine Learning Pipelines 2021 CIDR 5.8138422e-05
6,898 Unit Testing Data with Deequ 2019 SIGMOD 5.6530554e-05
7,538 Automating and Optimizing Data-Centric What-If Analyses on Native Machine Learning Pipelines 2023 SIGMOD 5.4995874e-05
7,629 mlwhatif: What If You Could Stop Re-Implementing Your Machine Learning Pipeline Analyses Over and Over? 2023 VLDB 5.4803665e-05
8,310 DORIAN in action: Assisted Design of Data Science Pipelines 2022 VLDB 5.3581954e-05
9,711 Serenade - Low-Latency Session-Based Recommendation in e-Commerce at Scale 2022 SIGMOD 5.1371426e-05
9,822 BlockJoin: Efficient Matrix Partitioning Through Joins 2017 VLDB 5.1254832e-05
9,824 Optimistic Recovery for Iterative Dataflows in Action 2015 SIGMOD 5.1254334e-05
10,898 stratum: A System Infrastructure for Massive Agent-Centric ML Workloads 2026 VLDB 4.9793485e-05
10,961 SemPiper: Interactive Code Synthesis for Semantic Operators in Machine Learning Pipelines 2026 VLDB 4.9793485e-05
11,407 mlidea: Interactively Improving ML Data Preparation Code via “Shadow Pipelines” 2025 VLDB 4.9793485e-05
11,607 A Flexible Forecasting Stack 2024 VLDB 4.9793485e-05
11,621 Snapcase – Regain Control over Your Predictions with Low-Latency Machine Unlearning 2024 VLDB 4.9793485e-05
11,671 Reconstructing and Querying ML Pipeline Intermediates 2023 CIDR 4.9793485e-05
11,818 Screening Native ML Pipelines with “ArgusEyes” 2022 CIDR 4.9793485e-05
12,528 Iterative Parallel Data Processing with Stratosphere: An Inside Look 2013 SIGMOD 4.9793485e-05
13,815 DEEM 2019: Workshop on Data Management for End-to-End Machine Learning 2019 SIGMOD -
Previous Page 1 / 1 Next

Frequent Co-authors

Co-authored at least 5 papers.

Co-author Shared Papers Rank Pagerank
Stefan Grafberger 8 1,094 0.067527232
Volker Markl 5 27 0.72645947