Hybrid Querying Over Relational Databases and Large Language Models
Summary: Presents SWAN: the first cross-domain benchmark of 120 beyond-database questions over four real-world relational schemas for hybrid DB+LLM querying. Proposes schema-expansion and UDF-based integration, evaluates GPT‑4 Turbo (≤40% exec accuracy, 48.2% factuality) and exposes optimization needs and accuracy/factuality gaps. (summarized by gpt-5-mini on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Fuheng Zhao (University of California Santa Barbara)
- 2. Divyakant Agrawal (University of California Santa Barbara)
- 3. Amr El Abbadi (University of California Santa Barbara)
BibTeX Citation
@inproceedings{zhao_cidr25,
address = {Amsterdam, Netherlands},
series = {{CIDR} '25},
title = {{Hybrid Querying Over Relational Databases and Large Language Models}},
booktitle = {Proceedings of the {Conference} on {Innovative} {Data} {Systems} {Research}},
author = {Zhao, Fuheng and Agrawal, Divyakant and Abbadi, Amr El},
year = {2025}
}
Incoming Citations (Sorted by Pagerank)
Showing 5 of 5 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 3,717 | Logical and Physical Optimizations for SQL Query Execution over Large Language Models | 2025 | SIGMOD | 7.0712441e-05 |
| 9,313 | Sphinteract: Resolving Ambiguities in NL2SQL Through User Interaction | 2025 | VLDB | 5.1953926e-05 |
| 11,033 | Bridging LLMs and Database Systems: A Deep Dive into Enhanced Relational Operators | 2026 | VLDB | 4.9769913e-05 |
| 11,170 | ScaleLLM: A Technique for Scalable LLM-augmented Data Systems | 2025 | SIGMOD | 4.9769913e-05 |
| 11,174 | SwellDB: Dynamic Query-Driven Table Generation with Large Language Models | 2025 | SIGMOD | 4.9769913e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 11 of 11 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 71 | DuckDB: an Embeddable Analytical Database | 2019 | SIGMOD | 0.00037724477 |
| 92 | CrowdDB: Answering Queries with Crowdsourcing | 2011 | SIGMOD | 0.00034670735 |
| 174 | Text-to-SQL Empowered by Large Language Models: A Benchmark Evaluation | 2024 | VLDB | 0.00026787121 |
| 257 | Crowdsourced Databases: Query Processing with People | 2011 | CIDR | 0.00022958404 |
| 259 | Answering Queries using Humans, Algorithms and Databases | 2011 | CIDR | 0.00022916014 |
| 329 | Can Foundation Models Wrangle Your Data? | 2023 | VLDB | 0.00020867521 |
| 2,117 | NL2SQL is a solved problem... Not! | 2024 | CIDR | 9.011308e-05 |
| 2,158 | Obtaining Complete Answers from Incomplete Databases | 1996 | VLDB | 8.9413031e-05 |
| 2,430 | Deco: A System for Declarative Crowdsourcing | 2012 | VLDB | 8.4763599e-05 |
| 3,668 | Revisiting Prompt Engineering via Declarative Crowdsourcing | 2024 | CIDR | 7.1144218e-05 |
| 8,979 | What Should A Database Know? | 1988 | PODS | 5.244147e-05 |
Previous
Page 1 / 1
Next