Back to papers
Chameleon: Foundation Models for Fairness-aware Multi-modal Data Augmentation to Enhance Coverage of Minorities
Summary: Chameleon leverages foundation models to generate minimal, targeted multi-modal synthetic tuples to boost coverage of under-represented groups. It couples prompt-guidance strategies with quality and outlier-detection filters to preserve semantic integrity and significantly reduce downstream unfairness.
(summarized by gpt-5-mini on Feb 09 2026)
- Paper ID
- 13558
- Venue
- VLDB
- Year
- 2024
- Pagerank
- 4.1905499e-05
- Overall Rank
- 11,071 | 23.06%
- DOI
-
10.14778/3681954.3682014
Incoming Non-self Citations Over Time
No non-self incoming citations found for this paper in this database.
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
Outgoing Citations (Sorted by Pagerank)
Showing 12 of 12 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank |
Cited Paper |
Year |
Venue |
Pagerank |
| 1,037 |
Interventional Fairness : Causal Database Repair for Algorithmic Fairness |
2019 |
SIGMOD |
0.00014514825 |
| 1,088 |
Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes |
2024 |
VLDB |
0.00014158762 |
| 4,023 |
Through the Fairness Lens: Experimental Analysis and Evaluation of Entity Matching |
2023 |
VLDB |
6.5181292e-05 |
| 4,885 |
Relational Data Synthesis using Generative Adversarial Networks: A Design Space Exploration |
2020 |
VLDB |
5.8509546e-05 |
| 5,264 |
PrivLava: Synthesizing Relational Data with Foreign Keys under Differential Privacy |
2023 |
SIGMOD |
5.5951175e-05 |
| 5,506 |
Can Large Language Models Predict Data Correlations from Column Names? |
2023 |
VLDB |
5.4711611e-05 |
| 5,785 |
ImDiffusion: Imputed Diffusion Models for Multivariate Time Series Anomaly Detection |
2024 |
VLDB |
5.3257637e-05 |
| 5,982 |
Responsible Data Integration: Next-generation Challenges |
2022 |
SIGMOD |
5.2409386e-05 |
| 6,462 |
Tailoring Data Source Distributions for Fairness-aware Data Integration |
2021 |
VLDB |
5.0479645e-05 |
| 6,712 |
Demonstrating GPT-DB: Generating Query-Specific and Customizable Code for SQL Processing with GPT-4 |
2023 |
VLDB |
4.9474017e-05 |
| 6,895 |
Identifying Insufficient Data Coverage for Ordinal Continuous-Valued Attributes |
2021 |
SIGMOD |
4.8879337e-05 |
| 7,571 |
Identifying Insufficient Data Coverage in Databases with Multiple Relations |
2020 |
VLDB |
4.7037322e-05 |
Semantically Similar Papers
| Overall Rank |
Paper |
Year |
Venue |
Pagerank |
| 6,850 |
Through the Data Management Lens: Experimental Analysis and Evaluation of Fair Classification |
2022 |
SIGMOD |
4.9036077e-05 |
| 13,115 |
Interactive Fairness Auditing: Leveraging AVOIR for Dynamic Evaluation and Mitigation |
2025 |
SIGMOD |
- |
| 8,847 |
Towards Foundation Database Models |
2025 |
CIDR |
4.4329366e-05 |
| 10,564 |
Mining the Minoria: Unknown, Under-represented, and Under-performing Minority Groups |
2025 |
VLDB |
4.1905499e-05 |
| 7,827 |
Chameleon: a Heterogeneous and Disaggregated Accelerator System for Retrieval-Augmented Language Models |
2025 |
VLDB |
4.6379617e-05 |
| 7,605 |
Causal Feature Selection for Algorithmic Fairness |
2022 |
SIGMOD |
4.6943015e-05 |
| 1,037 |
Interventional Fairness : Causal Database Repair for Algorithmic Fairness |
2019 |
SIGMOD |
0.00014514825 |
| 9,373 |
Falcon: Fair Active Learning using Multi-armed Bandits |
2024 |
VLDB |
4.3460825e-05 |
| 4,866 |
OmniFair: A Declarative System for Model-Agnostic Group Fairness in Machine Learning |
2021 |
SIGMOD |
5.8620848e-05 |
| 4,761 |
Automated Feature Engineering for Algorithmic Fairness |
2021 |
VLDB |
5.9341687e-05 |