Back to papers
JSON Tiles: Fast Analytics on Semi-Structured Data
Summary: JSON Tiles enables fast analytics on semi-structured JSON in relational DBs without fixed schemas. It auto-detects key attributes, extracts them transparently like a column store, handles heterogeneous data, and collects optimizer statistics, achieving near-columnar scan speeds.
(summarized by gpt-5-nano on Feb 09 2026)
Paper ID
6152
Venue
SIGMOD
Year
2021
Pagerank
6.9276175e-05
Overall Rank
4,069 | 72.09%
DOI
10.1145/3448016.3452809
Incoming Non-self Citations Over Time
BibTeX Citation
Copy BibTeX
@inproceedings{durner_sigmod21,
title = {{JSON Tiles: Fast Analytics on Semi-Structured Data}},
author = {Durner, Dominik and Leis, Viktor and Neumann, Thomas},
series = {{SIGMOD} '21},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/3448016.3452809},
url = {https://dl.acm.org/doi/10.1145/3448016.3452809},
year = {2021}
}
Incoming Citations (Sorted by Pagerank)
Showing 12 of 12 citing papers.
Rank
Citing Paper
Year
Venue
Pagerank
7,058
BOSS - An Architecture for Database Kernel Composition
2024
VLDB
5.7154673e-05
8,508
JEDI: These aren't the JSON documents you're looking for...
2022
SIGMOD
5.4123872e-05
8,682
Columnar Formats for Schemaless LSM-based Document Stores
2022
VLDB
5.3857524e-05
9,012
Towards Theory for Real-World Data
2022
PODS
5.3323444e-05
9,436
GIO: Generating Efficient Matrix and Frame Readers for Custom Data Formats by Example
2023
SIGMOD
5.2692207e-05
9,673
High-Ratio Compression for Machine-Generated Data
2023
SIGMOD
5.2383683e-05
9,893
ReCG: Bottom-Up JSON Schema Discovery Using a Repetitive Cluster-and-Generalize Framework
2024
VLDB
5.1997534e-05
9,989
GpJSON: High-performance JSON Data Processing on GPUs
2025
VLDB
5.1826377e-05
10,526
TurboLynx: Schemaless Graph Engine Strikes Back for General-Purpose Analytics
2026
VLDB
5.093636e-05
10,771
Nested Parquet Is Flat, Why Not Use It? How To Scan Nested Data With On-the-Fly Key Generation and Joins
2025
SIGMOD
5.093636e-05
11,274
Partition, Don’t Sort! Compression Boosters for Cloud Data Ingestion Pipelines
2024
VLDB
5.093636e-05
11,393
dsJSON: A Distributed SQL JSON Processor
2023
SIGMOD
5.093636e-05
Outgoing Citations (Sorted by Pagerank)
Showing 30 of 30 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
Rank
Cited Paper
Year
Venue
Pagerank
18
How Good Are Query Optimizers, Really?
2016
VLDB
0.00059284255
24
Spark SQL: Relational Data Processing in Spark
2015
SIGMOD
0.00054865648
32
Hive - A Warehousing Solution Over a Map-Reduce Framework
2009
VLDB
0.00050111008
51
Dremel: Interactive Analysis of Web-Scale Datasets
2010
VLDB
0.0004291425
66
The Snowflake Elastic Data Warehouse
2016
SIGMOD
0.00038561587
137
Relational Databases for Querying XML Documents: Limitations and Opportunities
1999
VLDB
0.00029877792
161
Mining Frequent Patterns without Candidate Generation
2000
SIGMOD
0.00027981772
165
DB2 with BLU Acceleration: So Much More than Just a Column Store
2013
VLDB
0.00027693424
318
Storing Semistructured Data with STORED
1999
SIGMOD
0.00021380325
422
Umbra: A Disk-Based System with In-Memory Performance
2020
CIDR
0.00018732744
941
Data Blocks: Hybrid OLTP and OLAP on Compressed Storage using both Vectorization and Compilation
2016
SIGMOD
0.00013078348
1,070
NoDB: Efficient Query Execution on Raw Data Files
2012
SIGMOD
0.0001232307
1,738
Sinew: A SQL System for Multi-Structured Data
2014
SIGMOD
9.8823494e-05
1,903
Instant Loading for Main Memory Databases
2013
VLDB
9.5049156e-05
1,927
Here are my Data Files. Here are my Queries. Where are my Results?
2011
CIDR
9.4703074e-05
2,202
Quantifying TPC-H Choke Points and Their Optimizations
2020
VLDB
8.9639459e-05
2,389
Parallel Data Analysis Directly on Scientific File Formats
2014
SIGMOD
8.6439053e-05
2,391
Mison: A Fast JSON Parser for Data Analytics
2017
VLDB
8.6413407e-05
2,414
Filter Before You Parse: Faster Analytics on Raw Data with Sparser
2018
VLDB
8.6078841e-05
2,435
Parallel In-Situ Data Processing with Speculative Loading
2014
SIGMOD
8.5811022e-05
2,800
JSON Data Management – Supporting Schema-less Development in RDBMS
2014
SIGMOD
8.1124117e-05
3,074
Speculative Distributed CSV Data Parsing for Big Data Analytics
2019
SIGMOD
7.7844208e-05
3,638
Fast Queries Over Heterogeneous Data Through Engine Customization
2016
VLDB
7.2338361e-05
4,014
Automatic Generation of Normalized Relational Schemas from Nested Key-Value Data
2016
SIGMOD
6.956076e-05
4,764
ReCache: Reactive Caching for Fast Analytics over Heterogeneous Data
2018
VLDB
6.51896e-05
5,048
FAD.js: Fast JSON Data Access Using JIT-based Speculative Optimizations
2017
VLDB
6.3870952e-05
5,803
An LSM-based Tuple Compaction Framework for Apache AsterixDB
2020
VLDB
6.0841986e-05
7,557
STEED: An Analytical Database System for TrEE-structured Data
2017
VLDB
5.599684e-05
8,799
FishStore: Faster Ingestion with Subset Hashing
2019
SIGMOD
5.369237e-05
9,216
Management of Flexible Schema Data in RDBMSs - Opportunities and Limitations for NoSQL
2015
CIDR
5.3051647e-05
Semantically Similar Papers
#
Overall Rank
Paper
Year
Venue
1
1,738
Sinew: A SQL System for Multi-Structured Data
2014
SIGMOD
2
7,766
Scalable Structural Index Construction for JSON Analytics
2021
VLDB
3
1,591
SQLGraph: An Efficient Relational-Based Property Graph Store
2015
SIGMOD
4
11,393
dsJSON: A Distributed SQL JSON Processor
2023
SIGMOD
5
6,557
Semistructured Models, Queries and Algebras in the Big Data Era
2016
SIGMOD
6
7,312
Closing the functional and Performance Gap between SQL and NoSQL
2016
SIGMOD
7
3,287
JSON: Data model, Query languages and Schema specification
2017
PODS
8
4,014
Automatic Generation of Normalized Relational Schemas from Nested Key-Value Data
2016
SIGMOD
9
11,771
JSON Schema Matching: Empirical Observations
2020
SIGMOD
10
2,800
JSON Data Management – Supporting Schema-less Development in RDBMS
2014
SIGMOD