DBScholar

Back to papers

Mining Frequent Patterns without Candidate Generation

Summary: Proposes FP-tree, a compact prefix-tree that compresses frequent-pattern data and eliminates candidate generation. FP-growth mines all patterns via pattern fragment growth and divide-and-conquer on conditional databases, cutting scans and outperforming Apriori by roughly tenfold. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
3230
Venue
SIGMOD
Year
2000
Pagerank
0.00027981772
Overall Rank
161 | 98.90%
DOI
10.1145/342009.335372

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{han_sigmod00,
        title = {{Mining Frequent Patterns without Candidate Generation}},
        author = {Han, Jiawei and Pei, Jian and Yin, Yiwen},
        series = {{SIGMOD} '00},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/342009.335372},
        url = {https://dl.acm.org/doi/10.1145/342009.335372},
        year = {2000}
}

Incoming Citations (Sorted by Pagerank)

Showing 50 of 64 citing papers.

Rank Citing Paper Year Venue Pagerank
122 Approximate Frequency Counts over Data Streams 2002 VLDB 0.00031260115
706 Effective Community Search for Large Attributed Graphs 2016 VLDB 0.00014789612
1,018 Dense Subgraph Maintenance under Streaming Edge Weight Updates for Real-time Story Identification 2012 VLDB 0.0001263063
1,231 SAPPER: Subgraph Indexing and Approximate Matching in Large Graphs 2010 VLDB 0.00011571594
1,500 Efficient Discovery of Approximate Dependencies 2018 VLDB 0.00010561098
1,792 MacroBase: Prioritizing Attention in Fast Data 2017 SIGMOD 9.7436856e-05
1,951 Interpretable Data-Based Explanations for Fairness Debugging 2022 SIGMOD 9.4252389e-05
2,273 SliceLine: Fast, Linear-Algebra-based Slice Finding for ML Model Debugging 2021 SIGMOD 8.8230899e-05
2,366 On Differentially Private Frequent Itemset Mining 2013 VLDB 8.6866148e-05
2,496 Star-Cubing: Computing Iceberg Cubes by Top-Down and Bottom-Up Integration 2003 VLDB 8.5025699e-05
3,105 Explanation-Based Auditing 2012 VLDB 7.7528586e-05
3,190 Looking for Trouble: Analyzing Classifier Behavior via Pattern Divergence 2021 SIGMOD 7.6532441e-05
3,232 Privacy-preserving Anonymization of Set-valued Data 2008 VLDB 7.6179491e-05
3,324 Mining Compressed Frequent-Pattern Sets 2005 VLDB 7.5198821e-05
3,969 Cache-conscious Frequent Pattern Mining on a Modern Processor 2005 VLDB 6.9837297e-05
4,069 JSON Tiles: Fast Analytics on Semi-Structured Data 2021 SIGMOD 6.9276175e-05
4,581 Mining Graph Patterns Efficiently via Randomized Summaries 2009 VLDB 6.6198548e-05
4,685 Distributed Processing of k Shortest Path Queries over Dynamic Road Networks 2020 SIGMOD 6.5628805e-05
5,054 Explainable AI: Foundations, Applications, Opportunities for Data Management Research 2022 SIGMOD 6.3843089e-05
5,279 Mining Document Collections to Facilitate Accurate Approximate Entity Matching 2009 VLDB 6.2849137e-05
5,401 Towards Proximity Pattern Mining in Large Graphs 2010 SIGMOD 6.2318138e-05
5,494 rho-uncertainty: Inference-Proof Transaction Anonymization 2010 VLDB 6.1997164e-05
6,017 Data Mining with the SAP NetWeaver BI Accelerator 2006 VLDB 6.0061791e-05
6,454 An Optimal Algorithm for l1-Heavy Hitters in Insertion Streams and Related Problems 2016 PODS 5.8746921e-05
6,554 Approximate Algorithms for k-Anonymity 2007 SIGMOD 5.8416386e-05
6,759 REDS: Rule Extraction for Discovering Scenarios 2021 SIGMOD 5.7827401e-05
6,775 EAGr: Supporting Continuous Ego-centric Aggregate Queries over Large Dynamic Graphs 2014 SIGMOD 5.7776218e-05
7,333 Interesting-Phrase Mining for Ad-Hoc Text Analytics 2010 VLDB 5.6430503e-05
7,657 Mining Frequent Itemsets over Uncertain Databases 2012 VLDB 5.5739357e-05
7,710 Mining Tree-Structured Data on Multicore Systems 2009 VLDB 5.5618769e-05
8,048 WISK: A Workload-aware Learned Index for Spatial Keyword Queries 2023 SIGMOD 5.5003171e-05
8,557 A Condensed Representation to Find Frequent Patterns 2001 PODS 5.4119882e-05
8,656 PARAS: A Parameter Space Framework for Online Association Mining 2013 VLDB 5.3909781e-05
8,705 Progressive Deep Web Crawling Through Keyword Queries For Data Enrichment 2019 SIGMOD 5.3807913e-05
8,778 Evaluating Clustering in Subspace Projections of High Dimensional Data 2009 VLDB 5.3753402e-05
8,993 Relative Risk and Odds Ratio: A Data Mining Perspective 2005 PODS 5.3372281e-05
9,019 SourceSight: Enabling Effective Source Selection 2016 SIGMOD 5.3309706e-05
9,212 Feasible Itemset Distributions in Data Mining: Theory and Application 2003 PODS 5.3058708e-05
9,271 MAIDS: Mining Alarming Incidents from Data Streams 2004 SIGMOD 5.2945481e-05
9,427 Discovering Top-k Rules using Subjective and Objective Criteria 2023 SIGMOD 5.271035e-05
10,214 CoShap: A Scalable Coalition Growth Approach to Shapley Value Approximation 2026 SIGMOD 5.093636e-05
10,324 Outliers: The Good, the Bad and the Ugly 2026 SIGMOD 5.093636e-05
10,601 Elastic Index Selection for Label-Hybrid AKNN Search 2026 VLDB 5.093636e-05
10,679 SHARQ: Explainability Framework for Association Rules on Relational Data 2025 SIGMOD 5.093636e-05
10,766 Incremental Rule Discovery in Response to Parameter Updates 2025 SIGMOD 5.093636e-05
10,778 Subgroup Discovery with Small and Alternative Feature Sets 2025 SIGMOD 5.093636e-05
10,825 Explaining Black-Box Clustering Pipelines With Cluster-Explorer 2025 VLDB 5.093636e-05
11,850 Top-k Queries over Digital Traces 2019 SIGMOD 5.093636e-05
11,866 Finding Theme Communities from Database Networks 2019 VLDB 5.093636e-05
11,927 Deeper: A Data Enrichment System Powered by Deep Web 2018 SIGMOD 5.093636e-05
Previous Page 1 / 2 Next

Outgoing Citations (Sorted by Pagerank)

Showing 8 of 8 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers