DBScholar

Back to papers

Improving Time Series Data Compression in Apache IoTDB

Summary: Introduces homomorphic compression theory for time-series queries and CompressIoTDB, an Apache IoTDB integration using CompColumn to support filtering, aggregation, and windows without decompression. Late decompression and dynamic auxiliary management yield 53.4% higher throughput and 20% lower memory. (summarized by gpt-5.6-luna on Jul 24 2026)

Paper ID
h53c2fb4102591fb4
Venue
VLDB
Year
2025
Pagerank
4.9769913e-05
Overall Rank
11,323 | 23.90%
DOI
10.14778/3748191.3748204
PDF
Download (CC BY-NC-ND 4.0)

Incoming Non-self Citations Over Time

No non-self incoming citations found for this paper in this database.

Authors

BibTeX Citation

@article{tang_vldb25,
        title = {{Improving Time Series Data Compression in Apache IoTDB}},
        author = {Tang, Yuxin and Zhang, Feng and Guan, Jiawei and Tian, Yuan and Huang, Xiangdong and Wang, Chen and Wang, Jianmin and Du, Xiaoyong},
        journal = {PVLDB},
        series = {{VLDB} '25},
        volume = {18},
        number = {10},
        pages = {3406--3420},
        doi = {10.14778/3748191.3748204},
        url = {https://doi.org/10.14778/3748191.3748204},
        year = {2025}
}

Incoming Citations (Sorted by Pagerank)

Showing 0 of 0 citing papers.

Rank Citing Paper Year Venue Pagerank
Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 43 of 43 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Rank Cited Paper Year Venue Pagerank
61 Integrating Compression and Execution in Column-Oriented Database Systems 2006 SIGMOD 0.00039236924
148 Gorilla: A Fast, Scalable, In-Memory Time Series Database 2015 VLDB 0.00029001141
163 DB2 with BLU Acceleration: So Much More than Just a Column Store 2013 VLDB 0.00027480091
594 Linear Road: A Stream Data Management Benchmark 2004 VLDB 0.00015816384
1,347 Chimp: Efficient Lossless Floating Point Compression for Time Series Databases 2022 VLDB 0.00010945231
1,375 Query Preserving Graph Compression 2012 SIGMOD 0.00010872356
1,770 How to Wring a Table Dry: Entropy Compression of Relations and Querying of Compressed Relations 2006 VLDB 9.6822825e-05
1,794 Fast Approximate Correlation for Massive Time-series Data 2010 SIGMOD 9.6198563e-05
1,912 ModelarDB: Modular Model-Based Time Series Management with Spark and Cassandra 2018 VLDB 9.3869908e-05
1,947 Aurora: A Data Stream Management System 2003 SIGMOD 9.320509e-05
1,974 Decomposed Bounded Floats for Fast Compression and Queries 2021 VLDB 9.2806652e-05
2,920 Monarch: Google’s Planet-Scale In-Memory Time Series Database 2020 VLDB 7.8519381e-05
2,929 ALP: Adaptive Lossless floating-Point Compression 2023 SIGMOD 7.838721e-05
2,972 Time Series Data Cleaning: From Anomaly Detection to Anomaly Repairing 2017 VLDB 7.793491e-05
3,108 SABER: Window-Based Hybrid Stream Processing for Heterogeneous Architectures 2016 SIGMOD 7.6392673e-05
3,182 The FastLanes Compression Layout: Decoding >100 Billion Integers per Second with Scalar Code 2023 VLDB 7.5589507e-05
3,344 LittleTable: A Time-Series Database and Its Uses 2017 SIGMOD 7.3968028e-05
3,434 Finding Semantics in Time Series 2011 SIGMOD 7.2993169e-05
3,582 Elf: Erasing-based Lossless Floating-Point Compression 2023 VLDB 7.1889857e-05
3,607 Plato: Approximate Analytics over Compressed Time Series with Tight Deterministic Error Guarantees 2020 VLDB 7.1645803e-05
3,623 Apache IoTDB: A Time Series Database for IoT Applications 2023 SIGMOD 7.150278e-05
4,184 YADING: Fast Clustering of Large-Scale Time Series Data 2015 VLDB 6.7508911e-05
4,595 Sim-Piece: Highly Accurate Piecewise Linear Approximation through Similar Segment Merging 2023 VLDB 6.5097157e-05
4,858 Time Series Data Encoding for Efficient Storage: A Comparative Analysis in Apache IoTDB 2022 VLDB 6.3770911e-05
4,996 Managing Massive Time Series Streams with Multi-Scale Compressed Trickles 2009 VLDB 6.3207572e-05
5,477 Good to the Last Bit: Data-Driven Encoding with CodecDB 2021 SIGMOD 6.1152706e-05
5,812 Column Stores For Wide and Sparse Data 2007 CIDR 5.9844567e-05
6,452 Heracles: An Efficient Storage Model and Data Flushing for Performance Monitoring Timeseries 2021 VLDB 5.7792352e-05
6,529 CompressGraph: Efficient Parallel Graph Analytics with Rule-Based Compression 2023 SIGMOD 5.7532491e-05
6,686 Hierarchical Residual Encoding for Multiresolution Time Series Compression 2023 SIGMOD 5.7066636e-05
6,698 MOST: Model-Based Compression with Outlier Storage for Time Series Data 2023 SIGMOD 5.7040999e-05
6,926 Industrial-Strength OLTP Using Main Memory and Many Cores 2020 VLDB 5.6405735e-05
7,363 PairwiseHist: Fast, Accurate and Space-Efficient Approximate Query Processing with Data Compression 2024 VLDB 5.5404386e-05
8,031 TSM-Bench: Benchmarking Time Series Database Systems for Monitoring Applications 2023 VLDB 5.4019638e-05
8,166 PIDS: Attribute Decomposition for Improved Compression and Query Performance in Columnar Storage 2020 VLDB 5.3846089e-05
8,521 TSGBench: Time Series Generation Benchmark 2024 VLDB 5.3226157e-05
8,805 Camel: Efficient Compression of Floating-Point Time Series 2024 SIGMOD 5.272343e-05
9,323 Apache TsFile: An IoT-native Time Series File Format 2024 VLDB 5.1934797e-05
9,879 Clean4TSDB: A Data Cleaning Tool for Time Series Databases 2024 VLDB 5.115241e-05
9,881 MTSClean: Efficient Constraint-based Cleaning for Multi-Dimensional Time Series Data 2024 VLDB 5.115241e-05
10,283 Mining and Forecasting of Big Time-series Data 2015 SIGMOD 5.0461162e-05
11,702 Grouping Time Series for Efficient Columnar Storage 2023 SIGMOD 4.9769913e-05
11,745 Homomorphic Compression: Making Text Processing on Compression Unlimited 2023 SIGMOD 4.9769913e-05
Previous Page 1 / 1 Next

Semantically Similar Papers