DBScholar

Back to papers

Clustering by Pattern Similarity in Large Data Sets

Summary: pCluster: pattern-based similarity; clustering by coherent patterns on a subset of dimensions, not by value proximity. Applies to gene expression and collaborative filtering; shows an efficient algorithm with real and synthetic data demonstrations. (summarized by gpt-5-nano on Feb 09 2026)

Paper ID
3428
Venue
SIGMOD
Year
2002
Pagerank
6.4980201e-05
Overall Rank
4,812 | 66.99%
DOI
10.1145/564691.564737

Incoming Non-self Citations Over Time

Authors

BibTeX Citation

@inproceedings{wang_sigmod02,
        title = {{Clustering by Pattern Similarity in Large Data Sets}},
        author = {Wang, Haixun and Wang, Wei and Yang, Jiong and Yu, Philip S.},
        series = {{SIGMOD} '02},
        booktitle = {Proceedings of the {ACM} {SIGMOD} International Conference on Management of Data},
        publisher = {Association for Computing Machinery},
        doi = {10.1145/564691.564737},
        url = {https://dl.acm.org/doi/10.1145/564691.564737},
        year = {2002}
}

Incoming Citations (Sorted by Pagerank)

Showing 4 of 4 citing papers.

Previous Page 1 / 1 Next

Outgoing Citations (Sorted by Pagerank)

Showing 6 of 6 cited papers.

Citations counted here include only citations to other VLDB/SIGMOD/CIDR/PODS papers in this database.

Previous Page 1 / 1 Next

Semantically Similar Papers