DBScholar

Back to papers

ThriftLLM: On Cost-Effective Selection of Large Language Models for Classification Queries

Summary: ThriftLLM formulates budget-constrained LLM ensemble selection for classification as maximizing correctness probability—non-monotone? nondecreasing but nonsubmodular and likely NP-hard—using a submodular upper bound for high-probability guarantees. Demonstrates cost-effective gains on classification and entity matching. (summarized by gpt-5.6-luna on Jul 24 2026)

Paper ID
hc72554aee48123f2
Venue
VLDB
Year
2025
Pagerank
5.3561732e-05
Overall Rank
8,316 | 44.09%
DOI
10.14778/3749646.3749702

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{huang_vldb25,
        title = {{ThriftLLM: On Cost-Effective Selection of Large Language Models for Classification Queries}},
        author = {Huang, Keke and Shi, Yimin and Ding, Dujian and Li, Yifei and Fei, Yang and Lakshmanan, Laks and Xiao, Xiaokui},
        journal = {PVLDB},
        series = {{VLDB} '25},
        volume = {18},
        number = {11},
        pages = {4410--4423},
        doi = {10.14778/3749646.3749702},
        url = {https://doi.org/10.14778/3749646.3749702},
        year = {2025}
}

Incoming Citations (Sorted by Pagerank)

Showing 4 of 4 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 16 of 16 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
134 Deep Entity Matching with Pre-Trained Language Models 2021 VLDB 0.00030043481
158 Deep Learning for Entity Matching: A Design Space Exploration 2018 SIGMOD 0.00028046388
174 Text-to-SQL Empowered by Large Language Models: A Benchmark Evaluation 2024 VLDB 0.00026790979
199 Influence Maximization: Near-Optimal Time Complexity Meets Practical Efficiency 2014 SIGMOD 0.00025522558
457 Distributed Representations of Tuples for Entity Resolution 2018 VLDB 0.00017907103
530 Magellan: Toward Building Entity Matching Management Systems 2016 VLDB 0.00016855162
623 Entity Resolution: Theory, Practice & Open Challenges 2012 VLDB 0.00015483844
986 DBSCAN Revisited: Mis-Claim, Un-Fixability, and Approximation 2015 SIGMOD 0.00012668888
1,391 Creating Embeddings of Heterogeneous Relational Datasets for Data Integration Tasks 2020 SIGMOD 0.00010816237
3,423 AutoTQA: Towards Autonomous Tabular Question Answering through Multi-Agent Large Language Models 2024 VLDB 7.313499e-05
3,489 Combining Small Language Models and Large Language Models for Zero-Shot NL2SQL 2024 VLDB 7.2627807e-05
4,331 FinSQL: Model-Agnostic LLMs-based Text-to-SQL Framework for Financial Analysis 2024 SIGMOD 6.6590534e-05
4,676 Entity Resolution with Hierarchical Graph Attention Networks 2022 SIGMOD 6.4745561e-05
7,318 An Interactive Multi-modal Query Answering System with Retrieval-Augmented Large Language Models 2024 VLDB 5.5547304e-05
7,612 Are Large Language Models a Good Replacement of Taxonomies? 2024 VLDB 5.484001e-05
8,016 Generating Succinct Descriptions of Database Schemata for Cost-Efficient Prompting of Large Language Models 2024 VLDB 5.4065977e-05
Previous Page 1 / 1 Next

Semantically Similar Papers