Exploiting Evidence from Unstructured Data to Enhance Master Data Management
Summary: Extends master data management beyond structured sources with an architecture and IBM implementation that correlates noisy textual entity mentions from news, email, transcripts, and chats. Uses extracted evidence to improve entity resolution and relationship discovery while preserving trusted MDM data. (summarized by gpt-5.6-luna on Jul 24 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Karin Murthy (IBM)
- 2. Prasad M Deshpande (IBM)
- 3. Atreyee Dey (IBM)
- 4. Ramanujam Halasipuram (IBM)
- 5. Mukesh Mohania (IBM)
- 6. Deepak P (IBM)
- 7. Jennifer Reed (IBM)
- 8. Scott Schumacher (IBM)
BibTeX Citation
@article{murthy_vldb12,
title = {{Exploiting Evidence from Unstructured Data to Enhance Master Data Management}},
author = {Murthy, Karin and Deshpande, Prasad M and Dey, Atreyee and Halasipuram, Ramanujam and Mohania, Mukesh and P, Deepak and Reed, Jennifer and Schumacher, Scott},
journal = {PVLDB},
series = {{VLDB} '12},
volume = {5},
number = {12},
pages = {1862--1873},
doi = {10.14778/2367502.2367524},
url = {https://doi.org/10.14778/2367502.2367524},
year = {2012}
}
Incoming Citations (Sorted by Pagerank)
Showing 1 of 1 citing papers.
| Rank | Citing Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 1,947 | Data Wrangling: The Challenging Journey from the Wild to the Lake | 2015 | CIDR | 9.4326202e-05 |
Previous
Page 1 / 1
Next
Outgoing Citations (Sorted by Pagerank)
Showing 4 of 4 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 3,446 | Efficient Approximate Entity Extraction with Edit Distance Constraints | 2009 | SIGMOD | 7.4087786e-05 |
| 3,507 | Efficiently Linking Text Documents with Relevant Structured Information | 2006 | VLDB | 7.3574742e-05 |
| 5,320 | Identity Resolution – 23 Years of Practical Experience and Observations at Scale | 2006 | SIGMOD | 6.2669189e-05 |
| 5,412 | Faerie: Efficient Filtering Algorithms for Approximate Dictionary-based Entity Extraction | 2011 | SIGMOD | 6.2272563e-05 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 12,169 | Mining Latent Entity Structures from Massive Unstructured and Interconnected Data | 2014 | SIGMOD |
| 2 | 11,186 | Unstructured Data Fusion for Schema and Data Extraction | 2024 | SIGMOD |
| 3 | 12,709 | One Platform for Mining Structured and Unstructured Data: Dream or Reality? | 2006 | VLDB |
| 4 | 7,572 | Cross Modal Data Discovery over Structured and Unstructured Data Lakes | 2023 | VLDB |
| 5 | 12,481 | The Case for a Structured Approach to Managing Unstructured Data | 2009 | CIDR |
| 6 | 11,980 | Building Structured Databases of Factual Knowledge from Massive Text Corpora | 2017 | SIGMOD |
| 7 | 13,827 | Managing Information Extraction [Tutorial Outline] | 2006 | SIGMOD |
| 8 | 13,850 | Integration of Structured and Unstructured Data in IBM Content Manager | 2005 | SIGMOD |
| 9 | 8,938 | Enabling Enterprise Mashups over Unstructured Text Feeds with InfoSphere MashupHub and SystemT | 2009 | SIGMOD |
| 10 | 12,516 | SMDM: Enhancing Enterprise-Wide Master Data Management Using Semantic Web Technologies | 2009 | VLDB |