DBScholar

Back to papers

How Large Language Models Will Disrupt Data Management

Summary: LLMs provide semantic grounding of tuples, schemas, and queries, enabling automation breakthroughs in tasks that stalled (entity resolution, schema matching, data discovery, query synthesis). They also blur predictive models and IR, prompting new DB/architecture designs. (summarized by gpt-5-mini on Feb 09 2026)

Paper ID
13354
Venue
VLDB
Year
2023
Pagerank
7.3297343e-05
Overall Rank
3,536 | 75.75%
DOI
10.14778/3611479.3611527

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@article{fernandez_vldb23,
        title = {{How Large Language Models Will Disrupt Data Management}},
        author = {Fernandez, Raul Castro and Elmore, Aaron J. and Franklin, Michael J. and Krishnan, Sanjay and Tan, Chenhao},
        journal = {PVLDB},
        series = {{VLDB} '23},
        volume = {16},
        number = {11},
        pages = {3302--3309},
        doi = {10.14778/3611479.3611527},
        url = {https://doi.org/10.14778/3611479.3611527},
        year = {2023}
}

Incoming Citations (Sorted by Pagerank)

Showing 12 of 12 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 22 of 22 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
17 Provenance Semirings 2007 PODS 0.00059843817
141 Deep Entity Matching with Pre-Trained Language Models 2021 VLDB 0.0002964847
397 TURL: Table Understanding through Representation Learning 2021 VLDB 0.00019278189
420 Can Foundation Models Wrangle Your Data? 2023 VLDB 0.00018789852
493 Data Integration for the Relational Web 2009 VLDB 0.00017558709
555 NaLIR: An Interactive Natural Language Interface for Querying Relational Databases 2014 SIGMOD 0.00016568053
579 Incremental Knowledge Base Construction Using DeepDive 2015 VLDB 0.00016217563
917 Data Integration: The Teenage Years 2006 VLDB 0.00013224381
969 Web-scale Data Integration: You can only afford to Pay As You Go 2007 CIDR 0.00012892148
1,337 DB-BERT: A Database Tuning Tool that "Reads the Manual" 2022 SIGMOD 0.00011117488
1,670 MISTIQUE: A System to Store and Query Model Intermediates for Model Diagnosis 2018 SIGMOD 0.00010045615
2,223 Sato: Contextual Semantic Type Detection in Tables 2020 VLDB 8.9189986e-05
2,473 PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel 2023 VLDB 8.5326287e-05
2,521 CodexDB: Synthesizing Code for Query Processing from Natural Language Instructions using GPT-3 Codex 2022 VLDB 8.4729505e-05
3,381 Ava: From Data to Insights Through Conversation 2017 CIDR 7.4571963e-05
3,530 MiCS: Near-linear Scaling for Training Gigantic Model on Public Cloud 2023 VLDB 7.3379782e-05
3,764 FastFlow: Accelerating Deep Learning Model Training with Smart Offloading of Input Data Pipeline 2023 VLDB 7.1446931e-05
4,687 Knowledge Graphs 2021: A Data Odyssey 2021 VLDB 6.5620198e-05
4,735 Leva: Boosting Machine Learning Performance with Relational Embedding Data Augmentation 2022 SIGMOD 6.5315782e-05
5,003 Galvatron: Efficient Transformer Training over Multiple GPUs Using Automatic Parallelism 2023 VLDB 6.4065691e-05
7,220 Solo: Data Discovery Using Natural Language Questions Via A Self-Supervised Approach 2023 SIGMOD 5.6679948e-05
8,127 The Case for NLP-Enhanced Database Tuning: Towards Tuning Tools that “Read the Manual” 2021 VLDB 5.4826853e-05
Previous Page 1 / 1 Next

Semantically Similar Papers