DBScholar

Back to papers

Mining Frequent Patterns without Candidate Generation

Summary: Proposes FP-tree, a compact prefix-tree that compresses frequent-pattern data and eliminates candidate generation. FP-growth mines all patterns via pattern fragment growth and divide-and-conquer on conditional databases, cutting scans and outperforming Apriori by roughly tenfold. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
ha3035396a0d819f4
Venue
SIGMOD
Year
2000
Pagerank
0.000273994
Overall Rank
164 | 98.90%
DOI
10.1145/342009.335372

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{han_sigmod00,
        title = {{Mining Frequent Patterns without Candidate Generation}},
        author = {Han, Jiawei and Pei, Jian and Yin, Yiwen},
        series = {{SIGMOD} '00},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/342009.335372},
        url = {https://dl.acm.org/doi/10.1145/342009.335372},
        year = {2000}
}

Incoming Citations (Sorted by Pagerank)

Showing 50 of 64 citing papers.

Rank Citing Paper Year Venue Pagerank
124 Approximate Frequency Counts over Data Streams 2002 VLDB 0.00030586757
714 Effective Community Search for Large Attributed Graphs 2016 VLDB 0.00014562182
988 Dense Subgraph Maintenance under Streaming Edge Weight Updates for Real-time Story Identification 2012 VLDB 0.00012654453
1,247 SAPPER: Subgraph Indexing and Approximate Matching in Large Graphs 2010 VLDB 0.00011348258
1,488 Efficient Discovery of Approximate Dependencies 2018 VLDB 0.00010513265
1,833 MacroBase: Prioritizing Attention in Fast Data 2017 SIGMOD 9.5376219e-05
1,943 Interpretable Data-Based Explanations for Fairness Debugging 2022 SIGMOD 9.3254131e-05
2,329 SliceLine: Fast, Linear-Algebra-based Slice Finding for ML Model Debugging 2021 SIGMOD 8.6268411e-05
2,417 On Differentially Private Frequent Itemset Mining 2013 VLDB 8.4922039e-05
2,544 Star-Cubing: Computing Iceberg Cubes by Top-Down and Bottom-Up Integration 2003 VLDB 8.3140577e-05
3,150 Explanation-Based Auditing 2012 VLDB 7.5885752e-05
3,263 Looking for Trouble: Analyzing Classifier Behavior via Pattern Divergence 2021 SIGMOD 7.4779877e-05
3,294 Privacy-preserving Anonymization of Set-valued Data 2008 VLDB 7.4454742e-05
3,384 Mining Compressed Frequent-Pattern Sets 2005 VLDB 7.3517702e-05
3,917 JSON Tiles: Fast Analytics on Semi-Structured Data 2021 SIGMOD 6.925328e-05
4,026 Cache-conscious Frequent Pattern Mining on a Modern Processor 2005 VLDB 6.8430008e-05
4,650 Mining Graph Patterns Efficiently via Randomized Summaries 2009 VLDB 6.4837167e-05
4,786 Distributed Processing of k Shortest Path Queries over Dynamic Road Networks 2020 SIGMOD 6.4129621e-05
5,179 Explainable AI: Foundations, Applications, Opportunities for Data Management Research 2022 SIGMOD 6.2382214e-05
5,408 Mining Document Collections to Facilitate Accurate Approximate Entity Matching 2009 VLDB 6.1429991e-05
5,483 Towards Proximity Pattern Mining in Large Graphs 2010 SIGMOD 6.1113443e-05
5,621 rho-uncertainty: Inference-Proof Transaction Anonymization 2010 VLDB 6.0588309e-05
6,145 Data Mining with the SAP NetWeaver BI Accelerator 2006 VLDB 5.8694977e-05
6,573 An Optimal Algorithm for l1-Heavy Hitters in Insertion Streams and Related Problems 2016 PODS 5.7431378e-05
6,682 Approximate Algorithms for k-Anonymity 2007 SIGMOD 5.7079897e-05
6,901 REDS: Rule Extraction for Discovering Scenarios 2021 SIGMOD 5.6503148e-05
6,914 EAGr: Supporting Continuous Ego-centric Aggregate Queries over Large Dynamic Graphs 2014 SIGMOD 5.6458225e-05
7,480 Interesting-Phrase Mining for Ad-Hoc Text Analytics 2010 VLDB 5.513824e-05
7,817 Mining Frequent Itemsets over Uncertain Databases 2012 VLDB 5.4462922e-05
7,821 Mining Tree-Structured Data on Multicore Systems 2009 VLDB 5.4456089e-05
8,204 WISK: A Workload-aware Learned Index for Spatial Keyword Queries 2023 SIGMOD 5.3769515e-05
8,732 A Condensed Representation to Find Frequent Patterns 2001 PODS 5.2880532e-05
8,744 Progressive Deep Web Crawling Through Keyword Queries For Data Enrichment 2019 SIGMOD 5.287931e-05
8,826 PARAS: A Parameter Space Framework for Online Association Mining 2013 VLDB 5.2679852e-05
8,949 Evaluating Clustering in Subspace Projections of High Dimensional Data 2009 VLDB 5.2522445e-05
9,161 Relative Risk and Odds Ratio: A Data Mining Perspective 2005 PODS 5.2154318e-05
9,197 SourceSight: Enabling Effective Source Selection 2016 SIGMOD 5.208891e-05
9,399 Feasible Itemset Distributions in Data Mining: Theory and Application 2003 PODS 5.1843659e-05
9,454 MAIDS: Mining Alarming Incidents from Data Streams 2004 SIGMOD 5.1733617e-05
9,509 Scalable Topical Phrase Mining from Text Corpora 2015 VLDB 5.168414e-05
9,616 Discovering Top-k Rules using Subjective and Objective Criteria 2023 SIGMOD 5.1503279e-05
10,442 CoShap: A Scalable Coalition Growth Approach to Shapley Value Approximation 2026 SIGMOD 4.9769913e-05
10,541 Outliers: The Good, the Bad and the Ugly 2026 SIGMOD 4.9769913e-05
11,057 Elastic Index Selection for Label-Hybrid AKNN Search 2026 VLDB 4.9769913e-05
11,128 SHARQ: Explainability Framework for Association Rules on Relational Data 2025 SIGMOD 4.9769913e-05
11,197 Incremental Rule Discovery in Response to Parameter Updates 2025 SIGMOD 4.9769913e-05
11,205 Subgroup Discovery with Small and Alternative Feature Sets 2025 SIGMOD 4.9769913e-05
11,241 Explaining Black-Box Clustering Pipelines With Cluster-Explorer 2025 VLDB 4.9769913e-05
12,156 Top-k Queries over Digital Traces 2019 SIGMOD 4.9769913e-05
12,172 Finding Theme Communities from Database Networks 2019 VLDB 4.9769913e-05
Previous Page 1 / 2 Next

Outgoing Citations (Sorted by Pagerank)

Showing 8 of 8 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers