Multi-column Substring Matching for Database Schema Translation
Summary: Unsupervised method for translating schemas from substrings in columns, without training data. Iterative algorithm deduces the substring concatenation to map between schemas; evaluated on real and synthetic data for fixed- and variable-length fields. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Robert H. Warren (University of Waterloo)
- 2. Frank Wm. Tompa (University of Waterloo)
BibTeX Citation
@article{warren_vldb06,
title = {{Multi-column Substring Matching for Database Schema Translation}},
author = {Warren, Robert H. and Tompa, Frank Wm.},
journal = {PVLDB},
series = {{VLDB} '06},
volume = {29},
number = {1},
pages = {331--342},
doi = {10.14778/1164135.1164168},
url = {https://doi.org/10.14778/1164135.1164168},
year = {2006}
}
Incoming Citations (Sorted by Pagerank)
Showing 6 of 6 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 2,980 | Auto-Join: Joining Tables by Leveraging Transformations | 2017 | VLDB | 7.8975073e-05 |
| 3,547 | Learning Semantic String Transformations from Examples | 2012 | VLDB | 7.3225973e-05 |
| 3,746 | Discovering Linkage Points over Web Data | 2013 | VLDB | 7.1561558e-05 |
| 4,525 | SEMA-JOIN: Joining Semantically-Related Tables Using Big Table Corpora | 2015 | VLDB | 6.6456999e-05 |
| 5,134 | Auto-FuzzyJoin: Auto-Program Fuzzy Similarity Joins Without Labeled Examples | 2021 | SIGMOD | 6.3532024e-05 |
| 9,629 | Auto-BI: Automatically Build BI-Models Leveraging Local Join Prediction and Global Schema Graph | 2023 | VLDB | 5.2434488e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 6 of 6 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 158 | Robust and Efficient Fuzzy Match for Online Data Cleaning | 2003 | SIGMOD | 0.00028199923 |
| 263 | Reconciling Schemas of Disparate Data Sources: A Machine-Learning Approach | 2001 | SIGMOD | 0.0002305693 |
| 297 | Generic Schema Matching with Cupid | 2001 | VLDB | 0.00022157284 |
| 1,030 | Data-Driven Understanding and Refinement of Schema Mappings | 2001 | SIGMOD | 0.00012545209 |
| 1,939 | iMAP: Discovering Complex Semantic Matches between Database Schemas | 2004 | SIGMOD | 9.4490491e-05 |
| 4,009 | Flexible String Matching Against Large Databases in Practice | 2004 | VLDB | 6.9612537e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 7,865 | Generating Succinct Descriptions of Database Schemata for Cost-Efficient Prompting of Large Language Models | 2024 | VLDB |
| 2 | 5,354 | Can Large Language Models Predict Data Correlations from Column Names? | 2023 | VLDB |
| 3 | 2,748 | Few-shot Text-to-SQL Translation using Structure and Content Prompt Learning | 2023 | SIGMOD |
| 4 | 3,456 | Automatic Discovery of Attributes in Relational Databases | 2011 | SIGMOD |
| 5 | 2,372 | Instance-based Schema Matching for Web Databases by Domain-specific Query Probing | 2004 | VLDB |
| 6 | 1,521 | On Multi-Column Foreign Key Discovery | 2010 | VLDB |
| 7 | 6,252 | Putting Context into Schema Matching | 2006 | VLDB |
| 8 | 3,942 | On the Complexity of Deriving Schema Mappings from Database Instances | 2008 | PODS |
| 9 | 808 | On Schema Matching with Opaque Column Names and Data Values | 2003 | SIGMOD |
| 10 | 454 | Using Schema Matching to Simplify Heterogeneous Data Translation | 1998 | VLDB |