ParSEval: Interactive Counterexample-driven Evaluation for Text-to-SQL
Summary: ParSEval interactively generates databases under configurable constraints to counterexample-test Text-to-SQL predictions beyond single-instance execution accuracy. It exposes witness instances and disagreement patterns across NULLs, duplicates, set/bag semantics, and key-foreign-key assumptions. (summarized by gpt-5.6-luna on Aug 28 2026)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
Authors
- 1. Chunyu Chen (Simon Fraser University)
- 2. Zhengjie Miao (Simon Fraser University)
- 3. Yong Zhang (Huawei)
- 4. Jiannan Wang (Tsinghua University)
BibTeX Citation
@article{chen_vldb26,
title = {{ParSEval: Interactive Counterexample-driven Evaluation for Text-to-SQL}},
author = {Chen, Chunyu and Miao, Zhengjie and Zhang, Yong and Wang, Jiannan},
journal = {PVLDB},
series = {{VLDB} '26},
volume = {19},
number = {12},
pages = {4846--4849},
doi = {10.14778/3827998.3828137},
url = {https://doi.org/10.14778/3827998.3828137},
year = {2026}
}
Incoming Citations (Sorted by Pagerank)
Showing 0 of 0 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 1 of 1 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 10,129 | ParSEval: Plan-aware Test Database Generation for SQL Equivalence Evaluation | 2025 | VLDB | 5.0751052e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 10,695 | NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions | 2026 | VLDB |
| 2 | 11,644 | Demonstration of the VeriEQL Equivalence Checker for Complex SQL Queries | 2024 | VLDB |
| 3 | 7,638 | Test Data Generation for Complex SQL Queries | 2026 | SIGMOD |
| 4 | 5,520 | PATSQL: Efficient Synthesis of SQL Queries from Example Tables with Quick Inference of Projected Columns | 2021 | VLDB |
| 5 | 4,847 | Explaining Wrong Queries Using Small Examples | 2019 | SIGMOD |
| 6 | 5,544 | Automated Validating and Fixing of Text-to-SQL Translation with Execution Consistency | 2025 | SIGMOD |
| 7 | 5,155 | An In-Depth Benchmarking of Text-to-SQL Systems | 2021 | SIGMOD |
| 8 | 10,893 | Developing and Benchmarking Verification Algorithms to Improve Text-to-SQL Generation | 2026 | VLDB |
| 9 | 10,969 | Text-to-SQL Evaluation Toolkit | 2026 | VLDB |
| 10 | 10,129 | ParSEval: Plan-aware Test Database Generation for SQL Equivalence Evaluation | 2025 | VLDB |