DBScholar

Back to papers

SHiFT: An Efficient, Flexible Search Engine for Transfer Learning

Summary: SHiFT: first downstream task-aware, flexible search engine over transfer-learning model repositories, exposing SHiFT-QL to express hybrid, task-specific model-selection criteria. Combines a cost-based optimizer and special incremental execution to evaluate queries efficiently for iterative ML workflows. (summarized by gpt-5-mini on Feb 09 2026)

Paper ID
h1b6ec6606bcfe71d
Venue
VLDB
Year
2023
Pagerank
5.336291e-05
Overall Rank
8,411 | 43.45%
DOI
10.14778/3565816.3565831

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{renggli_vldb23,
        title = {{SHiFT: An Efficient, Flexible Search Engine for Transfer Learning}},
        author = {Renggli, Cedric and Yao, Xiaozhe and Kolar, Luka and Rimanic, Luka and Klimovic, Ana and Zhang, Ce},
        journal = {PVLDB},
        series = {{VLDB} '23},
        volume = {16},
        number = {2},
        pages = {304--316},
        doi = {10.14778/3565816.3565831},
        url = {https://doi.org/10.14778/3565816.3565831},
        year = {2023}
}

Incoming Citations (Sorted by Pagerank)

Showing 2 of 2 citing papers.

Rank Citing Paper Year Venue Pagerank
11,176 Alsatian: Optimizing Model Search for Deep Transfer Learning 2025 SIGMOD 4.9793485e-05
11,303 LLMLog: Advanced Log Template Generation via LLM-driven Multi-Round Annotation 2025 VLDB 4.9793485e-05
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 16 of 16 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
104 HoloClean: Holistic Data Repairs with Probabilistic Inference 2017 VLDB 0.00033690989
205 Snorkel: Rapid Training Data Creation with Weak Supervision 2018 VLDB 0.00025181304
483 ActiveClean: Interactive Data Cleaning For Statistical Modeling 2016 VLDB 0.00017590977
1,152 Cerebro: A Data System for Optimized Deep Learning Model Selection 2020 VLDB 0.00011801961
1,255 Data Management in Machine Learning: Challenges, Techniques, and Systems 2017 SIGMOD 0.00011325762
1,409 Northstar: An Interactive Data Science System 2018 VLDB 0.00010743451
2,036 Ease.ml: Towards Multi-tenant Resource Sharing for Machine Learning Workloads 2018 VLDB 9.1520279e-05
2,323 ZeroER: Entity Resolution using Zero Labeled Examples 2020 SIGMOD 8.6348884e-05
2,598 Complaint-driven Training Data Debugging for Query 2.0 2020 SIGMOD 8.2385793e-05
3,040 Incremental and Approximate Inference for Faster Occlusion-based Deep CNN Explanations 2019 SIGMOD 7.7216143e-05
4,606 MLINSPECT: A Data Distribution Debugger for Machine Learning Pipelines 2021 SIGMOD 6.5041225e-05
5,506 Managing ML Pipelines: Feature Stores and the Coming Wave of Embedding Ecosystems 2021 VLDB 6.1016716e-05
9,438 Ease.ml in Action: Towards Multi-tenant Declarative Learning Services 2018 VLDB 5.1777033e-05
9,556 Towards an Optimized GROUP BY Abstraction for Large-Scale Machine Learning 2021 VLDB 5.1572248e-05
9,557 Intermittent Human-in-the-Loop Model Selection using Cerebro: A Demonstration 2021 VLDB 5.1572248e-05
11,818 Screening Native ML Pipelines with “ArgusEyes” 2022 CIDR 4.9793485e-05
Previous Page 1 / 1 Next

Semantically Similar Papers