DBScholar

Back to papers

ScienceBenchmark: A Complex Real-World Benchmark for Evaluating Natural Language to SQL Systems

Summary: ScienceBenchmark introduces the first domain-expert-validated NL-to-SQL benchmark over three complex scientific databases. It combines scarce human NL/SQL pairs with GPT-3 synthetic data, exposing severe weaknesses of Spider-trained systems under realistic schema and domain complexity. (summarized by gpt-5.6-luna on Jul 24 2026)

Paper ID
13934
Venue
VLDB
Year
2024
Pagerank
9.2721259e-05
Overall Rank
2,039 | 86.02%
DOI
10.14778/3636218.3636225

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{zhang_vldb24,
        title = {{ScienceBenchmark: A Complex Real-World Benchmark for Evaluating Natural Language to SQL Systems}},
        author = {Zhang, Yi and Deriu, Jan and Katsogiannis-Meimarakis, George and Kosten, Catherine and Koutrika, Georgia and Stockinger, Kurt},
        journal = {PVLDB},
        series = {{VLDB} '24},
        volume = {17},
        number = {4},
        pages = {685--698},
        doi = {10.14778/3636218.3636225},
        url = {https://doi.org/10.14778/3636218.3636225},
        year = {2024}
}

Incoming Citations (Sorted by Pagerank)

Showing 12 of 12 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 7 of 7 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers