Back to authors
Sebastian Schelter
- Author ID
- o0000-0003-4722-5840
- ORCID
-
0000-0003-4722-5840
- Links
-
(found by gpt-5.6-luna on jul 24 2026)
- Most Frequent Institution
- University of Amsterdam
- Pagerank
- 0.20065019
- Overall Rank
- 288 | 98.67%
- Paper Count
- 27
Affiliation Timeline
Incoming Non-self Citations Over Time
Total yearly non-self incoming citations across all papers by this author.
Publications by Paper Pagerank
Showing 27 of 27 publications.
| Rank |
Title |
Year |
Venue |
Pagerank |
| 1,308 |
Automating Large-Scale Data Quality Verification |
2018 |
VLDB |
0.0001107886 |
| 1,927 |
Elastic Machine Learning Algorithms in Amazon SageMaker |
2020 |
SIGMOD |
9.3607648e-05 |
| 2,264 |
An Intermediate Representation for Optimizing Machine Learning Pipelines |
2019 |
VLDB |
8.7289107e-05 |
| 3,735 |
Learning to Validate the Predictions of Black Box Classifiers on Unseen Data |
2020 |
SIGMOD |
7.0642839e-05 |
| 4,037 |
HedgeCut: Maintaining Randomised Trees for Low-Latency Machine Unlearning |
2021 |
SIGMOD |
6.8390354e-05 |
| 4,606 |
MLINSPECT: A Data Distribution Debugger for Machine Learning Pipelines |
2021 |
SIGMOD |
6.5041225e-05 |
| 5,304 |
Probabilistic Demand Forecasting at Scale |
2017 |
VLDB |
6.1873575e-05 |
| 5,385 |
SchemaPile: A Large Collection of Relational Database Schemas |
2024 |
SIGMOD |
6.1523255e-05 |
| 5,501 |
SemBench: A Benchmark for Semantic Query Processing Engines |
2026 |
VLDB |
6.1027188e-05 |
| 6,321 |
"Amnesia" - A Selection of Machine Learning Models That Can Forget User Data Very Fast |
2020 |
CIDR |
5.814903e-05 |
| 6,323 |
Lightweight Inspection of Data Preprocessing in Native Machine Learning Pipelines |
2021 |
CIDR |
5.8138422e-05 |
| 6,898 |
Unit Testing Data with Deequ |
2019 |
SIGMOD |
5.6530554e-05 |
| 7,538 |
Automating and Optimizing Data-Centric What-If Analyses on Native Machine Learning Pipelines |
2023 |
SIGMOD |
5.4995874e-05 |
| 7,629 |
mlwhatif: What If You Could Stop Re-Implementing Your Machine Learning Pipeline Analyses Over and Over? |
2023 |
VLDB |
5.4803665e-05 |
| 8,310 |
DORIAN in action: Assisted Design of Data Science Pipelines |
2022 |
VLDB |
5.3581954e-05 |
| 9,711 |
Serenade - Low-Latency Session-Based Recommendation in e-Commerce at Scale |
2022 |
SIGMOD |
5.1371426e-05 |
| 9,822 |
BlockJoin: Efficient Matrix Partitioning Through Joins |
2017 |
VLDB |
5.1254832e-05 |
| 9,824 |
Optimistic Recovery for Iterative Dataflows in Action |
2015 |
SIGMOD |
5.1254334e-05 |
| 10,898 |
stratum: A System Infrastructure for Massive Agent-Centric ML Workloads |
2026 |
VLDB |
4.9793485e-05 |
| 10,961 |
SemPiper: Interactive Code Synthesis for Semantic Operators in Machine Learning Pipelines |
2026 |
VLDB |
4.9793485e-05 |
| 11,407 |
mlidea: Interactively Improving ML Data Preparation Code via “Shadow Pipelines” |
2025 |
VLDB |
4.9793485e-05 |
| 11,607 |
A Flexible Forecasting Stack |
2024 |
VLDB |
4.9793485e-05 |
| 11,621 |
Snapcase – Regain Control over Your Predictions with Low-Latency Machine Unlearning |
2024 |
VLDB |
4.9793485e-05 |
| 11,671 |
Reconstructing and Querying ML Pipeline Intermediates |
2023 |
CIDR |
4.9793485e-05 |
| 11,818 |
Screening Native ML Pipelines with “ArgusEyes” |
2022 |
CIDR |
4.9793485e-05 |
| 12,528 |
Iterative Parallel Data Processing with Stratosphere: An Inside Look |
2013 |
SIGMOD |
4.9793485e-05 |
| 13,815 |
DEEM 2019: Workshop on Data Management for End-to-End Machine Learning |
2019 |
SIGMOD |
- |
Frequent Co-authors
Co-authored at least 5 papers.