Reward-SQL: Boosting Text-to-SQL via Stepwise Execution-Aware Reasoning and Process-Supervised Rewards
Summary: Reward-SQL introduces CoCTE, progressively composing SQL with execution-validated intermediate views and structured CTEs. Its entropy-weighted, execution-aware process reward model supervises both RL training and inference, improving accuracy, interpretability, and cross-domain generalization. (summarized by gpt-5.6-luna on Jul 26 2026)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
Authors
- 1. Yuxin Zhang (Renmin University of China)
- 2. Meihao Fan (Renmin University of China)
- 3. Ju Fan (Renmin University of China)
- 4. Mingyang Yi (Renmin University of China)
- 5. Yuyu Luo (Hong Kong University of Science and Technology)
- 6. Guoliang Li (Tsinghua University)
- 7. Bin Wu (Alibaba)
- 8. Wenchao Zhou (Alibaba)
BibTeX Citation
@inproceedings{zhang_sigmod26,
title = {{Reward-SQL: Boosting Text-to-SQL via Stepwise Execution-Aware Reasoning and Process-Supervised Rewards}},
author = {Zhang, Yuxin and Fan, Meihao and Fan, Ju and Yi, Mingyang and Luo, Yuyu and Li, Guoliang and Wu, Bin and Zhou, Wenchao},
series = {{SIGMOD} '26},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/3802105},
url = {https://dl.acm.org/doi/10.1145/3802105},
year = {2026}
}
Incoming Citations (Sorted by Pagerank)
Showing 7 of 7 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 7,273 | AutoPrep: Natural Language Question-Aware Data Preparation with a Multi-Agent Framework | 2025 | VLDB | 5.5668569e-05 |
| 9,220 | Natural Language to SQL: State of the Art and Open Problems | 2025 | VLDB | 5.2036865e-05 |
| 10,445 | DeepEye-SQL: A Software-Engineering-Inspired Text-to-SQL Framework | 2026 | SIGMOD | 4.9769913e-05 |
| 10,527 | VecBench: A Controllable Benchmark for Filtered Vector Search: [Experiments & Analysis] | 2026 | SIGMOD | 4.9769913e-05 |
| 10,729 | TACO: A Benchmark for Open-Domain Text-to-SQL with Ambiguous and Cross-Database Queries | 2026 | VLDB | 4.9769913e-05 |
| 10,878 | DeepPrep: An LLM-Powered Agentic System for Autonomous Data Preparation | 2026 | VLDB | 4.9769913e-05 |
| 10,901 | Dial: A Knowledge-Grounded Dialect-Specific NL2SQL System | 2026 | VLDB | 4.9769913e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 9 of 9 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 174 | Text-to-SQL Empowered by Large Language Models: A Benchmark Evaluation | 2024 | VLDB | 0.00026787121 |
| 533 | CodeS: Towards Building Open-source Language Models for Text-to-SQL | 2024 | SIGMOD | 0.00016818605 |
| 1,777 | OmniSQL: Synthesizing High-quality Text-to-SQL Data at Scale | 2025 | VLDB | 9.6599769e-05 |
| 1,867 | ScienceBenchmark: A Complex Real-World Benchmark for Evaluating Natural Language to SQL Systems | 2024 | VLDB | 9.4771545e-05 |
| 2,467 | Few-shot Text-to-SQL Translation using Structure and Content Prompt Learning | 2023 | SIGMOD | 8.4171827e-05 |
| 3,489 | Combining Small Language Models and Large Language Models for Zero-Shot NL2SQL | 2024 | VLDB | 7.2594757e-05 |
| 7,273 | AutoPrep: Natural Language Question-Aware Data Preparation with a Multi-Agent Framework | 2025 | VLDB | 5.5668569e-05 |
| 9,220 | Natural Language to SQL: State of the Art and Open Problems | 2025 | VLDB | 5.2036865e-05 |
| 10,878 | DeepPrep: An LLM-Powered Agentic System for Autonomous Data Preparation | 2026 | VLDB | 4.9769913e-05 |
Previous
Page 1 / 1
Next