BIRCH: An Efficient Data Clustering Method for Very Large Databases
Summary: BIRCH offers incremental, memory-efficient clustering for very large databases; first DB clustering method to effectively handle noise. Shows strong time/space efficiency and single-scan quality; outperforms CLARANS on large datasets. (summarized by gpt-5-nano on Feb 09 2026)
Incoming Non-self Citations Over Time
Authors
- 1. Tian Zhang (University of Wisconsin)
- 2. Raghu Ramakrishnan (University of Wisconsin)
- 3. Miron Livny (University of Wisconsin)
BibTeX Citation
@inproceedings{zhang_sigmod96,
title = {{BIRCH: An Efficient Data Clustering Method for Very Large Databases}},
author = {Zhang, Tian and Ramakrishnan, Raghu and Livny, Miron},
series = {{SIGMOD} '96},
booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
publisher = {Association for Computing Machinery},
doi = {10.1145/233269.233324},
url = {https://dl.acm.org/doi/10.1145/233269.233324},
year = {1996}
}
Incoming Citations (Sorted by Pagerank)
Showing 34 of 84 citing papers.
Outgoing Citations (Sorted by Pagerank)
Showing 1 of 1 cited papers.
Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.
| Rank | Cited Paper | Year | Venue | Pagerank |
|---|---|---|---|---|
| 88 | Efficient and Effective Clustering Methods for Spatial Data Mining | 1994 | VLDB | 0.00035240327 |
Previous
Page 1 / 1
Next
Semantically Similar Papers
| # | Overall Rank | Paper | Year | Venue |
|---|---|---|---|---|
| 1 | 304 | Automatic Subspace Clustering of High Dimensional Data for Data Mining Applications | 1998 | SIGMOD |
| 2 | 11,581 | ParChain: A Framework for Parallel Hierarchical Agglomerative Clustering using Nearest-Neighbor Chain | 2022 | VLDB |
| 3 | 3,511 | Scalable Discovery of Best Clusters on Large Graphs | 2010 | VLDB |
| 4 | 88 | Efficient and Effective Clustering Methods for Spatial Data Mining | 1994 | VLDB |
| 5 | 7,291 | Incremental and Effective Data Summarization for Dynamic Hierarchical Clustering | 2004 | SIGMOD |
| 6 | 6,452 | Efficient Implementation of Large-Scale Multi-Structural Databases | 2005 | VLDB |
| 7 | 7,342 | Efficient Search in Very Large Databases | 1988 | VLDB |
| 8 | 3,663 | Optimal Grid-Clustering: Towards Breaking the Curse of Dimensionality in High-Dimensional Clustering | 1999 | VLDB |
| 9 | 9,013 | Data Bubbles: Quality Preserving Performance Boosting for Hierarchical Clustering | 2001 | SIGMOD |
| 10 | 14,127 | Clustering Methods for Large Databases: From the Past to the Future | 1999 | SIGMOD |