Back to papers
Automated Discovery of Test Oracles for Database Management Systems Using LLMs
Summary: Argus uses LLMs to synthesize equivalent-query skeletons, formally checked by a SQL equivalence solver, then instantiates them into reusable DBMS tests. Evaluation across five systems found 41 new bugs, including 36 logic bugs.
(summarized by gpt-5.6-luna on Jul 26 2026)
Paper ID
h7ced0dcd20d3e2c8
Venue
SIGMOD
Year
2026
Pagerank
4.9769913e-05
Overall Rank
10,427 | 29.92%
DOI
10.1145/3802017
PDF
Download
(CC BY 4.0)
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
BibTeX Citation
Copy BibTeX
@inproceedings{mang_sigmod26,
title = {{Automated Discovery of Test Oracles for Database Management Systems Using LLMs}},
author = {Mang, Qiuyang and He, Runyuan and Zhong, Suyang and Liu, Xiaoxuan and Zhang, Huanchen and Cheung, Alvin},
series = {{SIGMOD} '26},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/3802017},
url = {https://dl.acm.org/doi/10.1145/3802017},
year = {2026}
}
Incoming Citations (Sorted by Pagerank)
Showing 0 of 0 citing papers.
Rank
Citing Paper
Year
Venue
Pagerank
Outgoing Citations (Sorted by Pagerank)
Showing 29 of 29 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Rank
Cited Paper
Year
Venue
Pagerank
71
DuckDB: an Embeddable Analytical Database
2019
SIGMOD
0.00037724477
234
TiDB: A Raft-based HTAP Database
2020
VLDB
0.00023756332
379
Apache Calcite: A Foundational Framework for Optimized Query Processing Over Heterogeneous Data Sources
2018
SIGMOD
0.00019507406
383
Massive Stochastic Testing of SQL
1998
VLDB
0.00019474868
658
DocETL: Agentic Query Rewriting and Evaluation for Complex Document Processing
2025
VLDB
0.00015058738
748
QAGen: Generating Query-Aware Test Databases
2007
SIGMOD
0.00014277246
790
Cosette: An Automated Prover for SQL
2017
CIDR
0.00013971102
1,832
APOLLO: Automatic Detection and Diagnosis of Performance Regressions in Database Systems
2020
VLDB
9.5385854e-05
1,890
WeTune: Automatic Discovery and Verification of Query Rewrite Rules
2022
SIGMOD
9.4234723e-05
1,930
Detecting Logic Bugs of Join Optimizations in DBMS
2023
SIGMOD
9.3522789e-05
2,037
LLM-R^2: A Large Language Model Enhanced Rule-based Rewrite System for Boosting Query Efficiency
2025
VLDB
9.1494269e-05
2,414
Data Generation using Declarative Constraints
2011
SIGMOD
8.5054535e-05
2,963
Proving Query Equivalence Using Linear Integer Arithmetic
2023
SIGMOD
7.8035846e-05
3,652
Keep It Simple: Testing Databases via Differential Query Plans
2024
SIGMOD
7.1293371e-05
3,900
SQLStorm: Taking Database Benchmarking into the LLM Era
2025
VLDB
6.9320915e-05
3,937
Testing Graph Database Systems via Graph-Aware Metamorphic Relations
2024
VLDB
6.9130095e-05
4,104
QED: A Powerful Query Equivalence Decider for SQL
2024
VLDB
6.8038041e-05
5,058
Detecting Metadata-Related Logic Bugs in Database Systems via Raw Database Construction
2024
VLDB
6.2915045e-05
5,211
SAM: Database Generation from Query Workloads with Supervised Autoregressive Models
2022
SIGMOD
6.2232582e-05
5,546
Automated Validating and Fixing of Text-to-SQL Translation with Execution Consistency
2025
SIGMOD
6.085842e-05
5,870
Leveraging Application Data Constraints to Optimize Database-Backed Web Applications
2023
VLDB
5.9620208e-05
6,102
Constant Optimization Driven Database System Testing
2025
SIGMOD
5.8840867e-05
6,575
Can Large Language Models Be Query Optimizer for Relational Databases?
2026
SIGMOD
5.7428777e-05
8,526
Cut Costs, Not Accuracy: LLM-Powered Data Processing with Guarantees
2026
SIGMOD
5.3215522e-05
9,412
Detecting Schema-Related Logic Bugs in Relational DBMSs via Equivalent Database Construction
2025
VLDB
5.1808981e-05
10,061
Finding Logic Bugs in Graph-processing Systems via Graph-cutting
2025
SIGMOD
5.0851868e-05
10,066
Finding Logic Bugs in Spatial Database Engines via Affine Equivalent Inputs
2024
SIGMOD
5.0851868e-05
10,130
An Adaptive Benchmark for Modeling User Exploration of Large Datasets
2025
SIGMOD
5.0727027e-05
10,133
ParSEval: Plan-aware Test Database Generation for SQL Equivalence Evaluation
2025
VLDB
5.0727027e-05
Semantically Similar Papers
#
Overall Rank
Paper
Year
Venue
1
10,298
Andromeda: Debugging Database Performance Issues with Retrieval-Augmented Large Language Models
2025
SIGMOD
2
1,930
Detecting Logic Bugs of Join Optimizations in DBMS
2023
SIGMOD
3
7,773
DBG-PT: A Large Language Model Assisted Query Performance Regression Debugger
2024
VLDB
4
9,412
Detecting Schema-Related Logic Bugs in Relational DBMSs via Equivalent Database Construction
2025
VLDB
5
10,547
Testing Graph Databases with Synthesized Queries
2026
SIGMOD
6
10,902
Developing and Benchmarking Verification Algorithms to Improve Text-to-SQL Generation
2026
VLDB
7
10,371
Leveraging Query Optimizers to Verify the Soundness of LLM-based Query Rewrites for Real-World Workloads, and More!
2026
CIDR
8
3,171
Panda: Performance Debugging for Databases using LLM Agents
2024
CIDR
9
10,444
DBugScribe: Automatic Database Bug Reproduction from Community Reports
2026
SIGMOD
10
4,113
Automatic Database Configuration Debugging using Retrieval-Augmented Language Models
2025
SIGMOD